z-lab/Qwen3.8-27B-DFlash2
1mês
99
1K
- Publicado
- Coletado
Quanto custa rodar isto
1.9
| Precisão | Precisa de | Mais barato que cabe |
|---|---|---|
| 16 bits | 4.8 | GTX 1660 S |
| 8 bits | 2.4 | GTX 1660 S |
| 4 bits | 1.2 | GTX 1660 S |
Pesos mais cache de atenção, com 15% de folga para ativações e o alocador. Tudo aqui é aritmética — não medimos velocidade, então não classificamos chip por ela. Ver todos os preços de GPU
New text-generation model. Tags: qwen3, dflash2, speculative-decoding, block-diffusion, draft-model, sglang, vllm, text-generation
O overfeed.news indexa e aponta. Publicamos um trecho curto — o artigo completo fica em Hugging Face New Models.
Mais de Hugging Face New Models
Entre para seguir esta fontealesha-pro/Qwen3.8-Flash-Next-abliterated-GSQ-RCO-Strata-GGUF
New image-text-to-text model. Tags: gguf, gsq, rco, strata, abliterated, control-vector, image-text-to-text, arxiv:2604.18556
JetBrains/Mellum2.1-12B-A2.5B-Thinking
New text-generation model. Tags: mellum, text-generation, conversational, en, arxiv:2605.31268, base_model:JetBrains/Mellum2-12B-A2.5B-Base, base_model:finetune:JetBrains/Mellum2-12B-A2.5B-Base, license:apache-2.0
AtomicChat/Qwen-Image-2.1-Turbo-Uncensored-GGUF
New text-to-image model. Tags: gguf, atomic-chat, qwen, qwen-image, qwen3-vl, text-encoder, stable-diffusion.cpp, abliteration
speridlabs/iris-3b
New text-to-image model. Tags: pytorch, text-to-image, diffusion, pixel-space, flow-matching, general-vision-learner, vision-foundation-model, depth-estimation