Welcome to the grand finale! In Part 1, we learned how to walk (GEMM). In Part 2, we learned how to...
Advanced GPU Optimization: How to tech an LLM with CUDA and ROCm? - Part 4
Welcome to the grand finale! In Part 1, we learned how to walk (GEMM). In Part 2, we learned how to...
DSPy's pitch is that you stop hand-writing prompts and let an optimizer compile them for you. You...
In the last post, on 2026-08-18, I published what 14 MCP servers cost a context window before an...
Welcome to the grand finale! In Part 1, we learned how to walk (GEMM). In Part 2, we learned how to...