Ling 3.0 Tiny 8b runs fastest on low-end PC with 4GB VRAM, 36 tokens/sec; claims strong performance vs Qwen 3.5 9b and Gemma 12
Read the original at old.reddit.com→This Ling 3.0 Tiny 8b param with 1.3b active is the fastest, smartest model I can run on my poor old pc, with 4gb vram. It actually runs lightning fast, like 36 token / sec, as smart as Qwen 3.5 9b / Gemma 12, (Very...
Original headline: "Ling 3.0 Tiny is the strongest, fastest and greatest model on my low end PC!"
Coverage timeline
- Aug 17, 16:34 UTC r/LocalLLaMA lead source Ling 3.0 Tiny is the strongest, fastest and greatest model on my low end PC!
- Aug 18, 14:03 UTC r/LocalLLaMA Is Ling 3 tiny underrated for its size?
- Aug 19, 02:41 UTC r/LocalLLaMA Ling-3.0-tiny is a very interesting model. Run on NVIDIA Orin Nano Super 8GB at 128K context with IQ4_NL quant.