c-runtime

Tag

Cards List
#c-runtime

I ran Qwen3.5-0.8B on a sub-$20 CPU chip in under 512MiB of memory

Reddit r/LocalLLaMA ↗ · 2026-08-12

The author ran Qwen3.5-0.8B on a $10-20 Amlogic A113X CPU chip with a custom C runtime, achieving 1.82 tok/s decode and under 490 MiB peak RSS, demonstrating that small LLM inference can run on deployed edge hardware without a GPU.

0 favorites 0 likes
← Back to home

Submit Feedback