Models
Models are the AI Alfred can use. Your model choice affects answer quality, speed, usage cost, supported inputs, and how much context Alfred can consider.
You set your preferred model from the Settings Panel on the right of the Dashboard.
How to choose
Section titled “How to choose”Use a stronger model for hard reasoning, coding, detailed writing, complex file analysis, and high-stakes planning.
Use a faster or lower-cost model for quick questions, drafts, simple summaries, and high-volume work.
Most people are best served by picking one capable model for daily use and switching to a cheaper one when they know they’re going to have a long or repetitive conversation.
Availability
Section titled “Availability”Model availability depends on your plan. Plans are covered in Patreon.
| Model | Best for | Inputs | Availability |
|---|---|---|---|
| GPT-5.6-Terra | Advanced reasoning, professional writing, complex analysis | Text, images | Elite and above |
| GPT-5.6-Luna | Lighter, faster work that still needs solid reasoning | Text, images | Power and above |
| Claude-Sonnet | Coding, careful reasoning, long structured writing | Text, images | Elite and above |
| Claude-Haiku | Fast everyday work with good reasoning | Text, images | Power and above |
| Gemini-Pro | Large context, complex analysis, multimodal work | Text, images, audio, video | Elite and above |
| Gemini-Flash | Fast, cost-effective general work | Text, images, audio, video | All plans |
| Gemini-Flash-Lite | Lowest-cost conservation mode | Text, images, audio, video | All plans |
| Grok-4.6 | General work with a more direct answering style | Text, images | Power and above |
| Grok-4.3 | The previous Grok generation, still available | Text, images | Power and above |
| DeepSeek-Pro | Higher-quality text-only work | Text | All plans |
| DeepSeek-Flash | Fast text-only work | Text | All plans |
| Minimax-M3 | General work with image and video input | Text, images, video | Power and above |
| MiMo-V2.5 | Cheap multimodal work | Text, images, audio, video | All plans |
| MiMo-V2.5-Pro | Cheap text-only work with a bit more depth | Text | All plans |
| GLM-5.2 | Text-only work, strong at structured output | Text | Power and above |
Model names, availability and the lineup itself change over time as new models release and old ones get sunset. The Dashboard model picker is the source of truth for your account because it only shows what your current plan can use. If a model is locked, the picker shows the plan you’d need for it.
Alfred also uses a few models internally that you don’t select, for example image generation models and the models his agents run on. Those don’t appear in the picker.
Context
Section titled “Context”Every model has a context limit, which is how much of a conversation it can consider at once. Most models Alfred offers carry a large context, but a longer conversation still costs more per message because the whole thing gets sent along each time.
If a conversation gets very long, start a new one. It’s cheaper, and answers are usually better because Alfred isn’t carrying a pile of unrelated history. You can also turn on Memory Compression in the Settings Panel to keep long conversations lighter.
Automatic model switching
Section titled “Automatic model switching”Your selected model is a preference, not a guarantee. When your usage gets high, Alfred may temporarily switch to a more cost-effective model to keep your daily allowance available for longer.
You may see Alfred mention that he switched models. This usually means you’re approaching a usage threshold. It does not mean your subscription changed, and it does not remove your preferred model setting.
Input support
Section titled “Input support”Some models can work with images, audio, video, or large context better than others. Models that can’t process images are given a Vision Agent that mediates vision tasks on their behalf, so uploading an image to a text-only model still works, it just goes through an extra step.
Text-only models tend to be the cheapest, so if you never send images they’re often the better pick.
Cost and speed
Section titled “Cost and speed”More capable models use more allowance. Larger prompts, long files, web research, generated outputs, and tool use all add to it too.
For simple work a cheaper model is often the better experience, because it’s faster and lets you do more before reaching your limit.
See Usage for how allowance works.