Todays Local #LLM / #AI Inference experiment; Running a 753 Billion parameter model hooked up to a local coding agent.No...

Todays Local #LLM / #AI Inference experiment; Running a 753 Billion parameter model hooked up to a local coding agent.No, I've not come into a couple of money trucks full of cash, I'm using an old Dell T5810 with 256GB of RAM (bought well before the memory price spike), a 512GB NVMe SSD, a quantised version of GLM 5.2, and Colibri (https://github.com/JustVugg/colibri)Yes, it's slow, but I'm well past the belief that all agentic coding needs to be interactive, and I'm not hooked on dopamine hit cycles.

Read Original

Related