Model availability, names, and capabilities can change over time. Use the model picker as the source of truth for what is available in your account.
Choose by task
Recently added models
Use the model picker as the source of truth for your account, but these recent additions can help when you are choosing quickly: Claude Fable 5.1, GLM 5.3 Flash, and Gemini 3.7 Flash joined Babbily on September 1, 2026. Each has a 1-million-token context window and supports image, document, and audio attachments in Babbily. Available tools and file limits still depend on your model, plan, and account state.Capability signals
The model picker and composer can show what a model supports. Common capability differences include:- File types and file size behavior.
- Context window size.
- Tool availability.
- Image generation or editing.
- Video generation or editing.
- Memory personalization availability.
- Speed and reasoning depth.
- Multi-step or agentic task behavior.
- Codebase-scale context handling.
- Fast mode availability.
- Sunsetting status for older model options.
- API usage budget impact.
Sunsetting and removed models
A model that still appears in the picker with a Sunsetting label remains selectable during its transition. Choose a newer option for new work. Once an older model is removed from the picker, you can no longer select it there, even if its name still appears in saved chat history. As of September 1, 2026, these four models remain selectable as Sunsetting:
GLM 5.2 retains Fast mode during this transition. Gemini 3.6 Flash and Gemini 3.5 Flash Lite remain in the chat picker but are no longer options in Translate.
The following older options were removed from the picker on September 1:
These are alternatives for new work, not promises of identical output or automatic replacements in every chat. Recheck file support, tools, and API usage budget when switching models.
Use Fast mode
Fast mode is available only when your selected model supports it. A lightning icon next to a model name means Fast mode is available for that model. Claude Opus 5 and Veo 3.1 support Fast mode, as does the sunsetting GLM 5.2. Claude Fable 5.1, GLM 5.3 Flash, and Gemini 3.7 Flash do not currently have a separate Fast mode control. “Flash” is part of a model’s name, not an indication that it has this toggle.1
Select a supported model
Open the model picker and choose a model with the Fast mode lightning icon, such as Claude Opus 5.
2
Turn on Fast mode
Choose Fast mode using the lightning control near the composer. Babbily shows a usage notice the first time you turn it on.
3
Send your prompt
Keep working with the selected model using its faster response path.
Choose reasoning effort
Open the composer menu, choose Manual, and turn on Reasoning. If the selected model supports effort levels, a Reasoning effort selector appears beside the microphone.
In Auto, Babbily chooses reasoning effort for supported work. In Off, Babbily does not request optional reasoning. Higher effort can take longer and use more API usage budget. The available levels can change when you switch models.
Advanced reasoning and coding work
Some models are built for specialized advanced work, such as comparing multiple approaches, reviewing code, working across large files, or planning multi-step engineering tasks. Choose these models when the task benefits from depth, synthesis, or a larger working context. For better results:- State the goal, constraints, and expected output.
- Include only the files or excerpts the model needs.
- Ask for a plan before large edits or refactors.
- Ask the model to call out uncertainty, risks, and tradeoffs.
Switch models during work
You can switch models when the task changes. Good switching patterns:- Start with a faster model to brainstorm.
- Switch to a stronger reasoning model for decisions or plans.
- Switch to a file-capable model before attaching documents.
- Switch to a media-capable model for image or video work.
- Use Fast mode when it appears and you want quicker supported replies, then check your API usage budget for heavier work.
- Start a new thread if older context is steering the answer in the wrong direction.
Usage considerations
More advanced model work can use more of your API usage budget, especially when it includes tools, long file context, image generation, video generation, or deep research. For budget-sensitive work:- Use shorter prompts when the task is simple.
- Avoid unnecessary file attachments.
- Turn tools Off for model-only drafting.
- Use search or deep research only when sources matter.
- Check your usage card before running image or video generation.
Troubleshooting
If a capability is missing:- Switch to a model that supports it.
- Change tool mode from Off to Auto or Manual.
- Check whether your plan supports the feature.
- Check your remaining API usage budget.
- Try a new thread if the current one has a long or confusing context history.
Chat and models
Learn the basics of choosing models and continuing threads.
Known limits
Understand why capabilities can vary by model, plan, file, tool, and plugin.
