Start with auto
auto chooses from frontier open models for each new user turn. It is the default on a standard Fireworks key.
Use auto-instant when you want a latency-first open-model mix. If fast models are disabled by model governance, use auto.
/model, and select Auto. No --model flag is required.
Choose a family, tier, or version
Use--model only when you want a specific family, tier, or version:
Prefer
*-latest unless you need a pinned version.
Current aliases
These are all serverless aliases as of September 27, 2026. Aliases can move to newer models or be retired, so treat this list as a snapshot. The serverless models endpoint always shows the current set, andfireconnect model list shows the current coding aliases.
qwen-max-latest is not in the coding catalog, so fireconnect model list does not show it.
Aliases may change as Fireworks adds and retires models. For example,
deepseek-pro-latest is deprecated. Check the current list before you
hard-code an alias, as shown in Find a model ID.Find a model ID
List current models and aliases with FireConnect:--refresh fetches the latest copy.
Without FireConnect, call the serverless models endpoint:
aliases field lists the aliases that point to that model, such as accounts/fireworks/routers/glm-latest for glm-5p3. Use the last part, glm-latest, as the model ID. Remove use_cases=coding to list every serverless model.
This command lists models for FireConnect. To construct a custom
firerouter/... ID without FireConnect, use the
FireRouter supported models.
Check image support
glm-latest and glm-fast-latest are text-only. Image support for other aliases depends on the model currently behind each alias. See the Input column in Current aliases, and check fireconnect model list before sending images.
If an image reaches a text-only model in Claude Code, use /rewind or switch to a vision-capable model such as Kimi or glm-5p3-flash.