Curated by
M
More in Local AI & GPUs
See all 36 →More from Mike Boscia
See all stacks →Run GLM-5.2 on a Consumer Machine with Colibri
Colibri allows you to run the GLM-5.2 model (744B MoE) on a consumer machine with 25GB of RAM using pure C and zero dependencies. This tiny engine streams experts from disk, making it efficient and powerful for various applications.
Built for AI agentsNo ACO on this card