Which models can your machine actually run?
Magnitude now answers that. The new model catalog:
- Profiles your hardware automatically
- Estimates tok/s for every model before you download
- Recommends the best models for your machine
Pick one and Magnitude handles the rest:
- Downloads the model and quant from Hugging Face
- Loads it into the built-in inference engine
- Configures speculative decoding (MTP, DFlash, etc.)
- Sets concurrency based on your memory
Try it on your hardware:
npm i -g @magnitudedev/cli