r/LocalLLaMA · · 1 min read

Ling 3.0 Tiny is the strongest, fastest and greatest model on my low end PC!

Mirrored from r/LocalLLaMA for archival readability. Support the source by reading on the original site.

Ling 3.0 Tiny is the strongest, fastest and greatest model on my low end PC!

This Ling 3.0 Tiny 8b param with 1.3b active is the fastest, smartest model I can run on my poor old pc, with 4gb vram. It actually runs lightning fast, like 36 token / sec, as smart as Qwen 3.5 9b / Gemma 12, (Very close), and even faster because of 1.3b active parameters.

The Qwen 3.5 9b is running with like 5 token / sec, but this with 36 is finally the speed that i want. Very good open source model, I hope we'll get more of this tiny and really fast models, thank you! :)

submitted by /u/cosmos_hu
[link] [comments]

Discussion (0)

Sign in to join the discussion. Free account, 30 seconds — email code or GitHub.

Sign in →

No comments yet. Sign in and be the first to say something.

More from r/LocalLLaMA