Show HN: PicoLM v1.0-rc1
By gabucino · 2026-09-03 · 1 points · 0 comments
https://github.com/whoreson/picolm/
PicoLM is an LLM inference engine written in C99. It currently supports llama-2, GPT-2, Qwen 3.6/3.8(+MoE) and Gemma-3n models. Significant amount of work went into CPU SIMD acceleration/testing/correctness, and wide cross-platform availability with constant testi…
Open the full discussion on BetterNews