CPU inference for Kimi K3, a 2.78T-parameter MoE LLM, in pure Rust. No GPU, no BLAS, no PyTorch. Streams the checkpoint from disk. Byte-identical port of kimi-k3-in-c.
Related
alphaparkinc/genpark-multimodal-generative-art-prompt-synthesis-engine-skill: Multimodal generative art prompt synthesis & diffusion renderer (SeaArt style)
Multimodal generative art prompt synthesis & diffusion renderer (SeaArt style)
Greninja9257/LabLLM: A native macOS lab for teaching tiny language models to think — build the architecture, train the weights, and watch a small LLM emerge from scratch, locally on Apple Silicon with custom data, tokenizers, checkpoints, and MLX acceleration.
A native macOS lab for teaching tiny language models to think — build the architecture, train the weights, and watch a small LLM emerge from scratch, locally on Apple Silicon with ...
cneuralnetwork/oplogs: Local-first, open-source experiment tracking for machine learning and agents
Local-first, open-source experiment tracking for machine learning and agents