inclusionAI/Ling-3.0-tiny
2mo
330
12K
- Published
- Collected
What it costs to run this
7.9
| Precision | Needs | Cheapest that fits |
|---|---|---|
| 16-bit | 24 | RTX 3090 |
| 8-bit | 12 | Tesla P100 |
| 4-bit | 6.0 | Tesla P100 |
Weights plus attention cache, with 15% added for activations and allocator overhead. Everything here is arithmetic — we do not measure speed, so we do not rank chips by it. See all GPU prices
New text-generation model. Tags: bailing_hybrid, text-generation, conversational, custom_code, license:mit
overfeed.news indexes and links. We publish a short excerpt — the full article stays at Hugging Face New Models.
More from Hugging Face New Models
Log in to follow this sourcealesha-pro/Qwen3.8-Flash-Next-abliterated-GSQ-RCO-Strata-GGUF
New image-text-to-text model. Tags: gguf, gsq, rco, strata, abliterated, control-vector, image-text-to-text, arxiv:2604.18556
JetBrains/Mellum2.1-12B-A2.5B-Thinking
New text-generation model. Tags: mellum, text-generation, conversational, en, arxiv:2605.31268, base_model:JetBrains/Mellum2-12B-A2.5B-Base, base_model:finetune:JetBrains/Mellum2-12B-A2.5B-Base, license:apache-2.0
AtomicChat/Qwen-Image-2.1-Turbo-Uncensored-GGUF
New text-to-image model. Tags: gguf, atomic-chat, qwen, qwen-image, qwen3-vl, text-encoder, stable-diffusion.cpp, abliteration
speridlabs/iris-3b
New text-to-image model. Tags: pytorch, text-to-image, diffusion, pixel-space, flow-matching, general-vision-learner, vision-foundation-model, depth-estimation