Tools & releases
Magnitude Ships On-Device Inference Engine for AI Coding Agents
Magnitude released an open-source inference engine that compiles and tunes its kernels on the user's own hardware. It achieves up to 92% faster decode on Apple Silicon than llama.cpp and connects directly to Claude Code and Codex.
October 1, 2026 2 min read
Curated by Oleksandr Kuzmenko, AI Product EngineerUpdated October 1, 2026Sources cited on every story
AI-assisted · editor-reviewedHow we use AI

Why it matters
Magnitude released an open-source inference engine that compiles and tunes its kernels on the user's own hardware. It achieves up to 92% faster decode on Apple Silicon than llama.cpp and connects directly to Claude Code and Codex.