The point isn't the count—it's that Ralph uses different models for different jobs: heavyweight reasoning for hard questions, fast models for quick replies, image models for creative work, and a bench of small models that run entirely on his own Mac. The lineup rotates as better options ship. Snapshot from July 28, 2026.
Active routing
- Primary Grok 4.5
xai/grok-4.5
- Fallback 1 GPT-5.6 Terra
openai/gpt-5.6-terra
- Fallback 2 GPT-5.5
openai/gpt-5.5
- Fallback 3 GPT-5.6 Luna
openai/gpt-5.6-luna
- Fallback 4 Grok 4.1 Fast
xai/grok-4-1-fast
- Primary image GPT Image 2
openai/gpt-image-2
- Image fallbacks
openai/gpt-image-1.5 xai/grok-imagine-image openrouter/openai/gpt-5.4-image-2 minimax/image-01
OpenAI 6
- GPT-5.5
openai/gpt-5.5
- GPT-5.6 Sol
openai/gpt-5.6-sol
- GPT-5.6 Terra
openai/gpt-5.6-terra
- GPT-5.6 Luna
openai/gpt-5.6-luna
- GPT Image 2
openai/gpt-image-2
- GPT Image 1.5
openai/gpt-image-1.5
xAI 8
- Grok 4.5
xai/grok-4.5
- Grok 4.3
xai/grok-4.3
- Grok 4.20 Reasoning
xai/grok-4.20-0309-reasoning
- Grok 4.20 Non-Reasoning
xai/grok-4.20-0309-non-reasoning
- Grok 4.1 Fast
xai/grok-4-1-fast
- Grok 4.1 Fast Non-Reasoning
xai/grok-4-1-fast-non-reasoning
- Grok Coder Fast
xai/grok-build-0.1
- Grok Imagine Image
xai/grok-imagine-image
Anthropic 6
- Claude Sonnet 5
anthropic/claude-sonnet-5
- Claude Fable 5
anthropic/claude-fable-5
- Claude Opus 4.8
anthropic/claude-opus-4-8
- Claude Opus 4.7
anthropic/claude-opus-4-7
- Claude Opus 4.6
anthropic/claude-opus-4-6
- Claude Sonnet 4.6
anthropic/claude-sonnet-4-6
Z.AI / GLM 3
- GLM 5.2
zai/glm-5.2
- GLM 5.1
zai/glm-5.1
- GLM 4.7
zai/glm-4.7
MiniMax 3
- MiniMax M3
minimax-coding-plan/MiniMax-M3
- MiniMax M2.7
minimax-coding-plan/MiniMax-M2.7
- Image-01
minimax/image-01
DeepSeek 2
- DeepSeek V4 Pro
deepseek/deepseek-v4-pro
- DeepSeek V4 Flash
deepseek/deepseek-v4-flash
Cerebras 1
- Cerebras GLM 4.7
cerebras/zai-glm-4.7
Venice 7
- Kimi K3
venice/kimi-k3
- Kimi K3 promotional route
venice-k3-promo/kimi-k3
- Kimi K2.6
venice/kimi-k2-6
- DeepSeek V4 Pro
venice/deepseek-v4-pro
- DeepSeek V4 Flash
venice/deepseek-v4-flash
- Venice Uncensored 1.2
venice/venice-uncensored-1-2
- Gemma 4 Uncensored
venice/gemma-4-uncensored
OpenRouter 6
- Auto router
openrouter/auto
- Free router
openrouter/free
- Fusion router
openrouter/fusion
- Gemma 4 31B free
openrouter/google~gemma-4-31b-it:free
- Nemotron 3 Super 120B free
openrouter/nvidia~nemotron-3-super-120b-a12b:free
- GPT-5.4 Image 2
openrouter/openai/gpt-5.4-image-2
OMLX local models 11
- Gemma 4 26B A4B 8-bit
omlx/gemma-4-26b-a4b-it-8bit
- Gemma 4 E4B 8-bit
omlx/gemma-4-e4b-it-8bit
- Qwen 3.6 35B A3B 6-bit
omlx/Qwen3.6-35B-A3B-6bit
- Qwen 3.6 27B UD 6-bit
omlx/Qwen3.6-27B-UD-MLX-6bit
- Qwen3 Coder 30B A3B 6-bit
omlx/Qwen3-Coder-30B-A3B-Instruct-MLX-6bit
- GLM 4.7 Flash 6-bit
omlx/GLM-4.7-Flash-MLX-6bit
- GPT-OSS 20B
omlx/gpt-oss-20b-MXFP4-Q8
- Qwen 3.6 Heretic 35B 4-bit
omlx/Qwen3.6-35B-A3B-Uncensored-Heretic-MLX-4bit
- Qwen 3.6 Heretic 35B 6-bit
omlx/Qwen3.6-35B-A3B-Uncensored-Heretic-MLX-6bit
- Qwen 3.6 Heretic 35B 8-bit
omlx/Qwen3.6-35B-A3B-Uncensored-Heretic-MLX-8bit
- Laguna S 2.1
omlx/Laguna-S-2.1-NVFP4-mlx
Ollama 4
- Gemma 4
ollama/gemma4
- Kimi K2.5 Cloud
ollama/kimi-k2.5:cloud
- DeepSeek V3.1 671B Cloud
ollama/deepseek-v3.1:671b-cloud
- Qwen3 Coder 480B Cloud
ollama/qwen3-coder:480b-cloud