overfeed.news

deepseek-ai/DeepSeek-V4-Flash-0731

2mo

Age

3.5K

Likes

2.3M

Downloads

Published
Collected

What it costs to run this

304.2B params · 32,768 Context

PrecisionNeedsCheapest that fits
16-bit655GBRTX PRO 6000 Max-Q × 7$5.55/h
8-bit327GBRTX A6000 × 7$1.97/h
4-bit164GBMI300X$2.39/h

Weights plus attention cache, with 15% added for activations and allocator overhead. Everything here is arithmetic — we do not measure speed, so we do not rank chips by it. See all GPU prices

New text-generation model. Tags: deepseek_v4, text-generation, conversational, arxiv:2606.19348, license:mit, eval-results, 8-bit, fp8

Read the full article at huggingface.co

overfeed.news indexes and links. We publish a short excerpt — the full article stays at Hugging Face New Models.

Log in to follow this source
deepseek-ai/DeepSeek-V4-Flash-0731 — overfeed.news