yandex/AliceAI-Foundation-80B-A3B-Base
18d
116
676
- Coletado
Quanto custa rodar isto
81.3
| Precisão | Precisa de | Mais barato que cabe |
|---|---|---|
| 16 bits | 178 | MI300X |
| 8 bits | 89 | RTX PRO 6000 Max-Q |
| 4 bits | 44 | Q RTX 8000 |
Pesos mais cache de atenção, com 15% de folga para ativações e o alocador. Tudo aqui é aritmética — não medimos velocidade, então não classificamos chip por ela. Ver todos os preços de GPU
New text-generation model. Tags: alice_ai, text-generation, custom_code, mixture-of-experts, vllm, ru, en, license:apache-2.0
O overfeed.news indexa e aponta. Publicamos um trecho curto — o artigo completo fica em Hugging Face New Models.
Mais de Hugging Face New Models
Entre para seguir esta fonteperplexity-ai/pplx-decider-v1.1-27b
New text-classification model. Tags: pytorch, qwen3_5, classification, multimodal, custom-code, decider, decision-model, text-classification
LiquidAI/d1-omni-600M
New image-text-to-text model. Tags: d1_omni, feature-extraction, liquid, lfm2.5, edge, decision, classification, calibration
nerkyor/Qwen3.8-27B-Coder390-EfficientThink-Opus5.5-GPT6Astra-Grok4.7-DSV4Pro-K3-SFT-RLOO-MTP-DFlash2
New image-text-to-text model. Tags: gguf, qwen3.8, efficient-thinking, reasoning, coding, uncensored, sft, simpo
Qwen/Qwen-Image-2.1-Turbo
New text-to-image model. Tags: diffusers, qwen, image-generation, image-editing, text-to-image, base_model:Qwen/Qwen-Image-2.1, base_model:finetune:Qwen/Qwen-Image-2.1, license:other