overfeed.news

openbmb/MiniCPM5-2B

1mo

Age

81

Likes

13

Downloads

Collected

What it costs to run this

2.5B params · 32,768 Context

PrecisionNeedsCheapest that fits
16-bit6.9GBTesla P100$0.03/h
8-bit3.5GBGTX 1660 S$0.02/h
4-bit1.7GBGTX 1660 S$0.02/h

Weights plus attention cache, with 15% added for activations and allocator overhead. Everything here is arithmetic — we do not measure speed, so we do not rank chips by it. See all GPU prices

New text-generation model. Tags: llama, text-generation, minicpm, minicpm5, long-context, tool-calling, on-device, edge-ai

Read the full article at huggingface.co

overfeed.news indexes and links. We publish a short excerpt — the full article stays at Hugging Face New Models.

Log in to follow this source
openbmb/MiniCPM5-2B — overfeed.news