overfeed.news

inclusionAI/Ling-3.0-tiny

2mo

Age

330

Likes

12K

Downloads

Published
Collected

What it costs to run this

7.9B params · 32,768 Context

PrecisionNeedsCheapest that fits
16-bit24GBRTX 3090$0.11/h
8-bit12GBTesla P100$0.03/h
4-bit6.0GBTesla P100$0.03/h

Weights plus attention cache, with 15% added for activations and allocator overhead. Everything here is arithmetic — we do not measure speed, so we do not rank chips by it. See all GPU prices

New text-generation model. Tags: bailing_hybrid, text-generation, conversational, custom_code, license:mit

Read the full article at huggingface.co

overfeed.news indexes and links. We publish a short excerpt — the full article stays at Hugging Face New Models.

Log in to follow this source
inclusionAI/Ling-3.0-tiny — overfeed.news