StackMap
Subscribe
Explore / magnitude
magnitudedev

magnitude

Local inference desktop app + CLI: profiles your hardware, estimates tok/s per model before download, tunes the one you pick, and connects Pi, OpenCode, Hermes, Codex or Claude Code in a click.

5,387 383 Rust Apache-2.0updated today
View on GitHubDispute this mapping →
Curator's take

The on-ramp for people who do not know which model their machine can actually run: Magnitude profiles your hardware, ranks catalog models and quants by speed, accuracy and memory before you download, then tunes context size and speculative decoding for the one you pick and wires it into your coding agent. Models load on demand and unload when idle. That is the real difference from Ollama or LM Studio, which run whatever you choose. If you already know your model and want a scriptable server with a huge ecosystem, Ollama is the safer default; for multi-GPU serving, use vLLM. Young (launched June 2026) and moving fast.

Mapped by ShipWithAI editors · links verified

Continue your stack

What teams reach for next — and why each earns a place beside magnitude. Ranked by curator confidence.

alternativealternativeOllamaODSmagnitude
pairs wellalternativebuilt withpick a node for the why · open it from the panel
Weekly digest
README.md2 min read

Magnitude icon

Magnitude

Run the best open models for your machine

Download Magnitude Documentation Discord Follow Magnitude on Twitter GitHub Repo stars

Magnitude is the open source inference engine for the hardware you already own. It profiles your machine, recommends the best open models for it, and tunes them for your exact hardware. One click connects your favorite agent (Pi, OpenCode, Hermes, etc.). Works on Apple Silicon, NVIDIA, AMD, or nothing but a CPU.

Download Magnitude for macOS, Windows, or Linux

⭐ Help us reach more developers and grow the Magnitude community. Star this repo!

https://github.com/user-attachments/assets/8317d05b-8a6e-40e0-b45d-81011ecbc329

Get started

  1. Download Magnitude, install it, and open the app.
  2. Choose a recommended model in Discover and download it.
  3. Connect your agent in Connections and start using it.

The desktop app includes the magnitude CLI. No separate installation is needed.

Why Magnitude?

  • Knows your machine: profiles your hardware and estimates tok/s before you download
  • Recommends the best models: ranked by speed, accuracy, intelligence, and memory
  • Tuned end to end: speculative decoding and more, all set for your hardware
  • Works with your agent: one click to connect Pi, OpenCode, Hermes, and more
  • Free to run: no token costs, API keys, or rate limits
  • Fully private and offline: models, prompts, and files stay on your machine
  • Models on demand: loaded on request, unloaded when idle or memory fills
  • Open source: Apache 2.0, yours to modify

FAQ

What is Magnitude?

An open source inference engine optimized for consumer hardware. The desktop app profiles your machine, recommends the best models for it, then downloads, tunes, and runs them. One click connects the agent you already use.

How does it know what my machine can run?

Magnitude profiles your hardware and estimates tok/s for every model in the catalog before you download anything. It ranks them by speed, accuracy, intelligence, and memory so you can pick.

How is this different from Ollama or LM Studio?

They run whatever model you pick. Magnitude helps you pick. It estimates how every model and quant will perform on your machine before you download, then tunes the one you choose for your exact hardware, from context size to speculative de