Gemma 4 e4b
GoogleQ4_K_MCurrentollama pull gemma4:e4b
License:Gemma Terms of Use
Google's 4B multimodal model with native tool-calling, vision inputs, and configurable reasoning tokens.
Source Manifest SizeOllama listed: 6.6GB
Context Token Window128k tokens(131,072 tokens)
Memory Requirement8 GB System RAMOfficial Vendor Guidance
Verified Capabilities:✓ Tool-calling✓ Multimodal VisionReasoning tokens
Gemma 4 12B
GoogleQ4_K_MCurrentollama pull gemma4:12b
License:Gemma Terms of Use
Flagship local model from Google offering high reasoning, multimodal perception, and 256k context.
Source Manifest SizeOllama listed: 7.7GB
Context Token Window256k tokens(262,144 tokens)
Memory Requirement16 GB System RAMOfficial Vendor Guidance
Verified Capabilities:✓ Tool-calling✓ Multimodal VisionReasoning tokens
Qwen 3.5 4B
Alibaba QwenQ4_K_MCurrentollama pull qwen3.5:4b
License:Apache-2.0
Compact Qwen 3.5 release featuring hybrid thinking tokens, strong code logic, vision input, and 256k context.
Source Manifest SizeOllama listed: 3.4GB
Context Token Window256k tokens(262,144 tokens)
Memory Requirement~6 GB ComfortStarter Empirical Estimate
Verified Capabilities:✓ Tool-calling✓ Multimodal VisionReasoning tokens
Qwen 3.5 9B
Alibaba QwenQ4_K_MCurrentollama pull qwen3.5:9b
License:Apache-2.0
Balanced 9B model excelling at complex multi-step reasoning, coding tasks, and 256k context window.
Source Manifest SizeOllama listed: 6.6GB
Context Token Window256k tokens(262,144 tokens)
Memory Requirement~10 GB ComfortStarter Empirical Estimate
Verified Capabilities:✓ Tool-calling✓ Multimodal VisionReasoning tokens
Qwen 3.6 27B
Alibaba QwenQ4_K_MCurrentollama pull qwen3.6:27b
License:Apache-2.0
Dense 27B model from Qwen 3.6 series with extended 256k context window and multimodal vision.
Source Manifest SizeOllama listed: 18GB
Context Token Window256k tokens(262,144 tokens)
Memory Requirement~24 GB ComfortStarter Empirical Estimate
Verified Capabilities:✓ Tool-calling✓ Multimodal VisionReasoning tokens
Qwen 3.6 35B
Alibaba QwenQ4_K_MCurrentollama pull qwen3.6:35b
License:Apache-2.0
35B parameter model with 23 GB download footprint and full 256k context for high-memory setups.
Source Manifest SizeOllama listed: 23GB
Context Token Window256k tokens(262,144 tokens)
Memory Requirement~30 GB ComfortStarter Empirical Estimate
Verified Capabilities:✓ Tool-calling✓ Multimodal VisionReasoning tokens
Qwen 3.8 27B
Alibaba QwenQ4_K_MCurrentollama pull qwen3.8:27b
License:Apache-2.0
Qwen 3.8 release (August 2026) with refined reasoning architecture, 256k context, and multimodal vision.
Source Manifest SizeOllama listed: 18GB
Context Token Window256k tokens(262,144 tokens)
Memory Requirement~24 GB ComfortStarter Empirical Estimate
Verified Capabilities:✓ Tool-calling✓ Multimodal VisionReasoning tokens
Qwen 3 8B
Alibaba QwenQ4_K_MCurrentollama pull qwen3:8b
License:Apache-2.0
Standard dense 8B model with fast latency and proven tool calling for intermediate laptops.
Source Manifest SizeOllama listed: 5.2GB
Context Token Window40k tokens(40,960 tokens)
Memory Requirement~8 GB ComfortStarter Empirical Estimate
Verified Capabilities:✓ Tool-calling✕ Text only
GPT-OSS 20B
Open-Weight CommunityQ4_K_MCurrentollama pull gpt-oss:20b
License:Apache-2.0
Open-weight architecture with 21B total parameters (3.6B active) and 14 GB Ollama download.
Source Manifest SizeOllama listed: 14GB
Context Token Window128k tokens(131,072 tokens)
Memory Requirement~18 GB ComfortStarter Empirical Estimate
Verified Capabilities:✓ Tool-calling✕ Text onlyReasoning tokens
Phi-4 Mini
MicrosoftQ4_K_MCurrentollama pull phi4-mini
License:MIT
Microsoft's 3.8B model with strong synthetic math, reasoning, and tool support for lightweight machines.
Source Manifest SizeOllama listed: 2.5GB
Context Token Window128k tokens(131,072 tokens)
Memory Requirement~4 GB ComfortStarter Empirical Estimate
Verified Capabilities:✓ Tool-calling✕ Text only
Nemotron 3 Nano 4B
NVIDIAQ4_K_MCurrentollama pull nemotron-3-nano:4b
License:NVIDIA Open Model License
NVIDIA's 4B edge model with 2.8 GB footprint and 256k context engineered for low-latency tool calling.
Source Manifest SizeOllama listed: 2.8GB
Context Token Window256k tokens(262,144 tokens)
Memory Requirement~5 GB ComfortStarter Empirical Estimate
Verified Capabilities:✓ Tool-calling✕ Text only
Qwen 3 Coder Next
Alibaba QwenQ4_K_MCurrentollama pull qwen3-coder-next:latest
License:Apache-2.0
80B sparse MoE coding model with 3B active parameters, 52 GB footprint, and 256k context for high-end workstations.
Source Manifest SizeOllama listed: 52GB
Context Token Window256k tokens(262,144 tokens)
Memory Requirement~60 GB ComfortStarter Empirical Estimate
Verified Capabilities:✓ Tool-calling✕ Text onlyReasoning tokens
DeepSeek R1 1.5B
DeepSeekQ4_K_MCurrentollama pull deepseek-r1:1.5b
License:MIT
Distilled 1.5B reasoning model with chain-of-thought outputs. Runs fast on CPU; does not support native function calling.
Source Manifest SizeOllama listed: 1.1GB
Context Token Window128k tokens(131,072 tokens)
Memory Requirement~3 GB ComfortStarter Empirical Estimate
Verified Capabilities:✕ No tools✕ Text onlyReasoning tokens
GLM-4 9B (Unverified Candidate)
Zhipu AIQ4_0Currentollama pull glm4:9b
License:GLM-4 License
Candidate model entry pending full verification audit.
Source Manifest SizeOllama listed: 5.5GB
Context Token Window128k tokens(131,072 tokens)
Memory Requirement~9 GB ComfortStarter Empirical Estimate
Verified Capabilities:✓ Tool-calling✕ Text only
Llama 3.2 3B
MetaQ4_K_MCurrent✓ Runtime Verifiedollama pull llama3.2:3b
License:Llama 3.2 Community License
Meta's 3B lightweight model. Runtime verified live through Hack Day Starter Chat + Tool-calling Agent.
Source Manifest SizeOllama listed: 2.0GB
Context Token Window128k tokens(131,072 tokens)
Memory Requirement~4 GB ComfortStarter Empirical Estimate
Verified Capabilities:✓ Tool-calling✕ Text only
Llama 3.1 8B
MetaQ4_K_MCurrent✓ Runtime Verifiedollama pull llama3.1:8b
License:Llama 3.1 Community License
Meta's foundational 8B model with 128k context. Runtime verified live through Hack Day Starter Chat + Tool-calling Agent.
Source Manifest SizeOllama listed: 4.9GB
Context Token Window128k tokens(131,072 tokens)
Memory Requirement~8 GB ComfortStarter Empirical Estimate
Verified Capabilities:✓ Tool-calling✕ Text only
Qwen 2.5 Coder 7B
Alibaba QwenQ4_K_MCurrent⏳ Runtime Pendingollama pull qwen2.5-coder:7b
License:Apache-2.0
Alibaba's dedicated 7B code generation model with 32k context. Source/capability verified; runtime verification pending.
Source Manifest SizeOllama listed: 4.7GB
Context Token Window32k tokens(32,768 tokens)
Memory Requirement~8 GB ComfortStarter Empirical Estimate
Verified Capabilities:✓ Tool-calling✕ Text only
Mistral NeMo 12B
Mistral AI / NVIDIAQ4_K_MCurrent⏳ Runtime Pendingollama pull mistral-nemo:12b
License:Apache-2.0
12B multilingual model with Tekken tokenizer and 1000k context. Source/capability verified; runtime verification pending.
Source Manifest SizeOllama listed: 7.1GB
Context Token Window1000k tokens(1,024,000 tokens)
Memory Requirement~11 GB ComfortStarter Empirical Estimate
Verified Capabilities:✓ Tool-calling✕ Text only
LFM 2.5 8B
Liquid AIQ4_K_MCurrent⏳ Runtime Pendingollama pull lfm2.5:8b
License:Liquid AI Model License
Liquid Foundation Model with reasoning tokens and 125k context. Source/capability verified; runtime verification pending.
Source Manifest SizeOllama listed: 5.2GB
Context Token Window125k tokens(128,000 tokens)
Memory Requirement~9 GB ComfortStarter Empirical Estimate
Verified Capabilities:✓ Tool-calling✕ Text onlyReasoning tokens
Devstral Small 2
Mistral AI CommunityQ4_K_MCurrent⏳ Runtime Pendingollama pull devstral-small-2:latest
License:Apache-2.0
24B multimodal code agent model with 384k context. Source/capability verified; runtime verification pending.
Source Manifest SizeOllama listed: 15GB
Context Token Window384k tokens(393,216 tokens)
Memory Requirement~19 GB ComfortStarter Empirical Estimate
Verified Capabilities:✓ Tool-calling✓ Multimodal Vision