| Llama 3.2 1B llama3.2:1b | reasoning | Phone | Suitable when a mobile runtime is present | ~1.3 GB | ~2 GB + OS | Smallest Llama option for short questions on phones and low-RAM PCs. Strengths: Low memory · Fast on CPU Limits: Weak on long contracts |
| Llama 3.2 3B llama3.2:3b | reasoning | Tablet | Suitable when a mobile runtime is present | ~2 GB | ~4 GB + OS | Common first local model for cited Q&A on tablets and modest PCs. Strengths: Fits many 8 GB machines · Clear short answers Limits: Limited nuance |
| Gemma 2 2B gemma2:2b | reasoning | Phone | Suitable when a mobile runtime is present | ~1.6 GB | ~3 GB + OS | Very small instruction model for phones and tight RAM. Strengths: Small download Limits: Shallow multi-document reasoning |
| Phi-3 Mini phi3:mini | reasoning | Phone | Suitable when a mobile runtime is present | ~2.3 GB | ~4 GB + OS | Compact reasoning for phones and tablets; needs chunking on long files. Strengths: Better logic than 1–2B class Limits: Short context |
| Qwen2.5 3B qwen2.5:3b | reasoning | Tablet | Suitable when a mobile runtime is present | ~2 GB | ~4 GB + OS | Small Qwen with a longer window than most tiny models. Suitable for tablets when RAM is reported. Strengths: Longer context than Phi-3 Mini Limits: Still a small model |
| Gemma 2 9B gemma2:9b | reasoning | Capable PC | Desktop / workstation only | ~5.4 GB | ~10 GB + OS | Mid-size general model for clearer multi-clause answers. Strengths: Better instruction quality than 2–3B Limits: Needs a 16 GB-class PC |
| Qwen2.5 7B qwen2.5:7b | reasoning | Capable PC | Desktop / workstation only | ~4.7 GB | ~8 GB + OS | Balanced local reasoning model for document excerpts. Strengths: Long context · Structured answers Limits: Slower on CPU-only |
| Mistral 7B mistral:7b | reasoning | Capable PC | Desktop / workstation only | ~4.1 GB | ~8 GB + OS | General 7B instruction model. Strengths: Clear short prose Limits: Shorter context than Qwen2.5 7B |
| Llama 3.1 8B llama3.1:8b | reasoning | Capable PC | Desktop / workstation only | ~4.9 GB | ~10 GB + OS | Strong general model. Do not load a whole 200-page PDF at once. Strengths: Broad instruction quality Limits: Long context is not free RAM |
| DeepSeek R1 8B deepseek-r1:8b | reasoning | Capable PC | Desktop / workstation only | ~4.9 GB | ~10 GB + OS | More deliberate reasoning on supplied excerpts. Strengths: Useful for “what differs” questions Limits: Can be slower |
| Mistral Nemo mistral-nemo | reasoning | High-performance | Desktop / workstation only | ~7.1 GB | ~12 GB + OS | Larger Mistral model when RAM allows. Strengths: Richer long-form answers Limits: Heavy for 16 GB machines |
| Qwen2.5 14B qwen2.5:14b | reasoning | High-performance | Desktop / workstation only | ~9 GB | ~16 GB + OS | Stronger reasoning for large excerpt sets. Strengths: Better clause discrimination Limits: Poor experience under ~32 GB |
| Phi-4 phi4 | reasoning | High-performance | Desktop / workstation only | ~9.1 GB | ~16 GB + OS | Higher-quality reasoning model. Still needs grounded excerpts. Strengths: Structured answers Limits: Heavy for laptops |
| DeepSeek R1 14B deepseek-r1:14b | reasoning | High-performance | Desktop / workstation only | ~9 GB | ~16 GB + OS | Heavier reasoning for comparison of cited passages. Strengths: Useful for contradictions Limits: Slow without GPU |
| Qwen2.5 32B qwen2.5:32b | reasoning | Workstation | Desktop / workstation only | ~20 GB | ~28 GB + OS | Large local model. Only when the machine can hold it. Strengths: Highest quality in this catalog when it fits Limits: Very large download |
| Mixtral 8x7B mixtral | reasoning | Workstation | Desktop / workstation only | ~26 GB | ~32 GB + OS | Mixture-of-experts. Often incompatible on ordinary PCs. Strengths: Strong general quality when it runs Limits: Huge RAM and download |
| Moondream moondream | vision | Tablet | Suitable when a mobile runtime is present | ~1.7 GB | ~4 GB + OS | Small vision model for stamps or diagrams. Not OCR and not a contract reasoner. Strengths: Tiny vision option Limits: Not a replacement for Tesseract |
| LLaVA 7B llava:7b | vision | Capable PC | Desktop / workstation only | ~4.7 GB | ~10 GB + OS | Vision-language model for page images. Strengths: Can describe visual page content Limits: Weak as a text-only contract model |
| Llama 3.2 Vision llama3.2-vision | vision | High-performance | Desktop / workstation only | ~7.9 GB | ~14 GB + OS | Vision-capable Llama for mixed scans. Strengths: Vision plus some reasoning Limits: Prefer selectable text when it exists |
| nomic-embed-text nomic-embed-text | embedding | Phone | Suitable when a mobile runtime is present | ~0.3 GB | ~1 GB | Local embeddings for retrieving passages. Does not answer questions. Strengths: Tiny · Good retrieval helper Limits: Not a reasoning model |
| mxbai-embed-large mxbai-embed-large | embedding | Capable PC | Desktop / workstation only | ~0.7 GB | ~1 GB | Larger embedding model for retrieval. Strengths: Stronger retrieval than MiniLM-class Limits: Cannot analyse contracts alone |
| BGE-M3 bge-m3 | embedding | Capable PC | Desktop / workstation only | ~1.2 GB | ~2 GB | Multilingual embedding model. Strengths: Multilingual retrieval Limits: Not a Q&A model |
| all-minilm all-minilm | embedding | Phone | Suitable when a mobile runtime is present | ~0.05 GB | ~1 GB | Smallest embedding option for phones and low-RAM devices. Strengths: Tiny download Limits: Weaker retrieval |