// model_catalog
Every model is quantized for Apple Silicon, tested by us, and runs fully on-device. Grouped by where it runs, then by family. Memory figures are on-device peaks — weights are download size estimates.
Models that fit within iOS's ~5 GB per-app memory budget. “Tight” models still run, but sit close to the limit on some devices.
Ultra-light chat and quick PDF Q&A; long answers cut off
Fast reasoning, PDF Q&A, and doc lookup at a tiny size
Solid everyday chat and PDF work; doc lookup can miss specifics
Reasoning plus accurate doc lookup in a light model
Current-gen Qwen chat with reasoning; forgot cross-chat facts in tests
Top all-rounder — reasoning, web search, PDF tools; passed every test
Fast, dependable chat and PDF/doc Q&A (Liquid AI)
Light general chat and PDF summaries; weak doc lookup in tests
Reliable general chat and doc Q&A; PDF summaries occasionally thin
Quality chat at a tiny footprint (QAT); mixed doc lookup in tests
Efficient chat with strong doc lookup; long answers sometimes cut off
Higher-quality Gemma chat and PDF summaries (QAT-tuned)
Reasoning, PDFs, and doc lookup — passed every on-device test
Default model — balanced chat and PDF Q&A; long answers cut off in tests
Short chat and simple doc questions; unstable on iOS and long PDF tasks
Reasoning, web search, and doc Q&A — near-perfect in tests
Reasoning and math; skipped web search and failed PDF summaries in tests
Compact instruction-following chat and doc Q&A
Every catalog model runs on a Mac with enough memory — the card shows the minimum RAM each one needs.
Ultra-light chat and quick PDF Q&A; long answers cut off
Fast reasoning, PDF Q&A, and doc lookup at a tiny size
Solid everyday chat and PDF work; doc lookup can miss specifics
Reasoning plus accurate doc lookup in a light model
Current-gen Qwen chat with reasoning; forgot cross-chat facts in tests
Top all-rounder — reasoning, web search, PDF tools; passed every test
Strong reasoning, web search, and PDF tools — Mac only (OOM on iOS)
High-accuracy PDF tools, web search, and doc lookup (needs 16 GB Mac)
Strongest reasoning in class — passed every on-device test
Fast, dependable chat and PDF/doc Q&A (Liquid AI)
Fast mixture-of-experts chat and PDFs on Mac; doc lookup missed specifics
Light general chat and PDF summaries; weak doc lookup in tests
Reliable general chat and doc Q&A; PDF summaries occasionally thin
Dependable general chat and doc Q&A on Mac
Quality chat at a tiny footprint (QAT); mixed doc lookup in tests
Efficient chat with strong doc lookup; long answers sometimes cut off
Higher-quality Gemma chat and PDF summaries (QAT-tuned)
Reasoning, PDFs, and doc lookup — passed every on-device test
Default model — balanced chat and PDF Q&A; long answers cut off in tests
Short chat and simple doc questions; unstable on iOS and long PDF tasks
High-accuracy reasoning and math on Mac
Reasoning, web search, and doc Q&A — near-perfect in tests
Reasoning and math; skipped web search and failed PDF summaries in tests
Long-context reasoning, web search, and doc Q&A on Mac
Reasoning, web search, and docs — passed every on-device test
MoE flagship — ~3B active/token, strong reasoning on Mac
Top Llama Nemotron instruct quality on high-RAM Macs
Largest Nemotron 3 — ~12B active/token for M3 Ultra class Macs
Compact instruction-following chat and doc Q&A
// sovereign_vault_protocol
Little Brother is coming soon. No account required. Nothing leaves your device.