Run the best local models for your machine
Magnitude is an open source inference engine optimized for consumer hardware. Runs on Apple Silicon, NVIDIA, AMD, or nothing but a CPU.
Why Magnitude?
- Knows your machine: profiles your hardware and estimates tok/s before download
- Recommends the best models: ranked by speed, accuracy, intelligence, and memory
- Tuned end to end: speculative decoding and more, all set for your hardware
- Works with your agent: one click to connect Pi, OpenCode, Hermes, and more
- Free to run: no token costs, API keys, or rate limits
- Fully private and offline: models, prompts, and files stay on your machine
- Models on demand: loaded on request, unloaded when idle or memory fills
- Open source: Apache 2.0, yours to modify
FAQ
What is Magnitude?
An open source inference engine optimized for consumer hardware. The desktop app profiles your machine, recommends the best models for it, then downloads, tunes, and runs them. One click connects the agent you already use.
How does it know what my machine can run?
Magnitude profiles your hardware and estimates tok/s for every model in the catalog before you download anything. It ranks them by speed, accuracy, intelligence, and memory so you can pick.
How is this different from Ollama or LM Studio?
They run whatever model you pick. Magnitude helps you pick. It estimates how every model and quant will perform on your machine before you download, then tunes the one you choose for your exact hardware, from context size to speculative decoding.
What hardware do I need?
There's no fixed minimum. Magnitude profiles your machine and recommends what runs well on it. More memory lets you run larger models.
What systems does Magnitude support?
The desktop app is native on macOS, Linux, and Windows. It runs on Apple Silicon, NVIDIA and AMD GPUs, and CPU-only machines, including unified-memory boxes like DGX Spark and Strix Halo.
Which harnesses work with it?
Pi, OpenCode, Hermes, OpenClaw, Codex, Claude Code, Oh My Pi, and Cline. Pick a model and connect your harness in one click.
Do I need to manage it after setup?
No. It runs in the background, loads models when your agent needs them, and unloads them when idle or memory gets tight.
Is it private?
Yes. Prompts, files, and models stay on your machine. Once a model is downloaded, no internet connection is needed.