Your machine.
One clear shortlist.
Prism finds the local models your PC can actually run,
with the right quantization, memory cost, and honest limits.

Analyze your machine.
Enter your GPU, VRAM, CPU, and RAM. Prism returns a ranked shortlist of models that actually fit — with quantization, memory cost, and honest limits.
Compare what actually runs.
RTX 4070 · 12 GB · 32 GB
Speed and memory figures are planning estimates, not measured benchmarks.
Questions, answered.
Not yet. Throughput values are planning estimates derived from model size, quantization, VRAM class, and CPU offload. They are labeled everywhere they appear.
No. No account required. You enter your hardware manually — nothing is installed, uploaded, or granted access to your machine.
The catalog syncs from the public Ollama library (daily when CI is enabled). VRAM and speed figures stay planning estimates — formula-based with published Ollama file sizes when available.
For sizing, yes — enter the unified memory total as both VRAM and RAM. Chip-specific throughput calibration is still in development.
A ranked shortlist of local models with the right quantization, memory cost, runtime, context size, estimated speed, and the main limitation to expect.