Magnitude: An open source inference engine that tunes itself to your hardware
Magnitude compiles and tunes kernels on-device to run open models up to 2x faster than llama.cpp across Apple Silicon, NVIDIA, AMD, and CPU setups.
Daily coverage of AI, developer tools and infrastructure. Each story explains what happened and why it matters, with a link to the original source.
Magnitude compiles and tunes kernels on-device to run open models up to 2x faster than llama.cpp across Apple Silicon, NVIDIA, AMD, and CPU setups.