overfeed.news

deepseek-ai/DeepSeek-V4-Flash-Vision-Exp

1mo

Age

84

Likes

Collected

What it costs to run this

304.6B params · 32,768 Context

PrecisionNeedsCheapest that fits
16-bit656GBRTX PRO 6000 Max-Q × 7$5.55/h
8-bit328GBRTX A6000 × 7$1.97/h
4-bit164GBMI300X$2.39/h

Weights plus attention cache, with 15% added for activations and allocator overhead. Everything here is arithmetic — we do not measure speed, so we do not rank chips by it. See all GPU prices

New text-generation model. Tags: deepseek_v4, text-generation, license:mit, 8-bit, fp8

Read the full article at huggingface.co

overfeed.news indexes and links. We publish a short excerpt — the full article stays at Hugging Face New Models.

Log in to follow this source
deepseek-ai/DeepSeek-V4-Flash-Vision-Exp — overfeed.news