-
Colibri: 소비자 머신의 25GB RAM에서 GLM-5.2 실행 (순수 C, 의존성 없음)
GitHub - JustVugg/colibri: Run GLM-5.2 (744B MoE) on a 25GB-RAM consumer machine — pure C, zero deps, experts streamed from disk. Tiny engine, immense model. 🐦
<p>A few days ago I found myself trying out GLM 5.2 and was really positively impressed. The capabilities and security I was getting from this LLM are similar to those I've gotten …
-
Ternlight - 브라우저에서 실행되는 7MB 임베딩 모델 (WASM)
Ternlight – 7 MB embedding model that runs in browser (WASM)
<p>Article URL: <a href="https://ternlight-demo.vercel.app/">https://ternlight-demo.vercel.app/</a></p> <p>Comments URL: <a href="https://news.ycombinator.com/item?id=48811644">htt…