Getting GLM 5.2 running on my slow computer
Streams 744B MoE experts from disk to run on 25GB RAM—no GPU, pure C.
Run GLM-4.5-Air (110B) on a 16GB-RAM consumer machine - identify the best memory allocation to overcome standard hardware limitations in Local LLM applications. Placement beats budget. Falsification-Tested laws, probes and recipes for LLMs on commodity hardware
Runs 110B models on 16GB RAM by proving placement beats budget with measured laws.
Hobbyists running large language models on limited consumer GPUs
llama.cpp · Ollama · vLLM
Streams 744B MoE experts from disk to run on 25GB RAM—no GPU, pure C.
Just a model update announcement and giveaway, no new tool or code.
Another context management pitch when Cursor and Continue already solve this.
Convenient billing wrapper, but you can still bring your own API keys.
GLM-5.2 deployment recipe for 4x DGX Spark with 655K context support.
Yet another AQI wrapper, but the plain-English email briefs are genuinely useful.