Text / byteshape
Qwen3.8 27B GGUF
Qwen3.8 27B GGUF from byteshape. Source-based hardware guidance from its published configuration.
EstimatedRepository opened Aug 18, 2026Source checked 9/24/2026Version: 3fdfbd9b
LOCALRENTED GPUOPEN WEIGHTS
At a glance
- Architecture
- unknown
- License
- apache-2.0
- Software
- llama.cpp
Will it run on your machine?
Save your machine to see a personalized rating and its reasoning.
Add my hardwareWays to run it
llama.cpp · Windows, macOS, Linux
Repository-specific command in publisher documentation; confirm dependencies and hardware in the source.
llama-server -hf byteshape/Qwen3.8-27B-GGUF:Qwen3.8-27B-IQ4_XS-3.84bpw -hfd incoai/Qwen3.8-27B-DFlash2-GGUF:Q4_K_M --spec-type draft-dflash --spec-draft-n-max 7 --no-mmprojExplore its uses
GGUF
Keep exploring
Text
Prism MLBonsai 2 · 27B
A compressed 27B-class reasoning model with publisher-provided GGUF packs and a dedicated llama.cpp fork for CUDA, Metal, and CPU.
MODEL SIZE27.36B
Text
DeepSeekDeepSeek V4.1 Flash
A multimodal reasoning model with a compressed key-value cache, published as open weights.
Text
QwenQwen3.8-Flash-Next
An experimental open-weight multimodal model with sparse attention and 262K native context.
MODEL SIZE125B language · 6B active