even the tiniest model here wants ~2 GB. double-check your numbers above?
What's my local AI?
Check what's the best model you can run locally!
Your machine
We auto-detect what we can from your browser; everything else is a best guess. Tweak any value below to see what else could run.
Hardware
Memory
Platform
Your local model
Other options
nothing smaller on the list — your pick above is the sweet spot.
nothing matches — try clearing the search or filters.
What can you do with it?
Local models are private by design — everything stays on your machine. A few things people actually do:
Private drafting
Brainstorm, rewrite, or translate text with no data ever leaving your laptop. Perfect for work that isn't yours to share.
Offline code help
Autocomplete, explain a gnarly snippet, or write a quick script — even on a plane or in a dead zone.
Read your own files
Summarize that 50-page report, pull the action items from a meeting transcript, or chat with your notes — without uploading them anywhere.
Zero-cost experiments
Try open-weight models (reasoning, vision, tool calls) on your own hardware before paying for an API — the model list on this page is your playground.
No throttle, no quota
Run as much as you want. No rate limits, no token metering, no monthly bill — just your electricity.
Learning & tinkering
The fastest way to build a mental model of what different model sizes are good at — swap models and feel the difference.
How detection works
- GPU
- WebGL renderer string
- VRAM
- GPU database lookup, unified memory on Apple, or estimate from RAM
- RAM
navigator.deviceMemory, cross-checked against the detected chip on Apple Silicon- CPU
navigator.hardwareConcurrency- OS
- user agent parsing
- WebGPU
navigator.gpupresence