> ## Documentation Index
> Fetch the complete documentation index at: https://docs.fireworks.ai/llms.txt
> Use this file to discover all available pages before exploring further.

> ## Agent Instructions
> For Fireworks Nexus, start at https://docs.fireworks.ai/nexus.
> Use https://docs.fireworks.ai/nexus/quickstart for coding harnesses, custom agents, APIs, SDKs, and LLM gateways.
> Use https://docs.fireworks.ai/nexus/firerouter for how model routers work, the supported model list, composition, closed-model credentials, and pricing.
> Prefer canonical short model IDs such as firerouter/opus. In LiteLLM litellm_params.model, use the full path fireworks_ai/accounts/fireworks/routers/firerouter/opus.
> Family names such as opus track the latest evaluated family version; do not describe them as fixed model versions.

# Open Models

> Use auto, version-tracking aliases, fast tiers, or pinned Fireworks open models in coding workflows.

Use a Fireworks open model directly when you want Fireworks to serve every turn. FireConnect automatically registers the coding-ready catalog when you connect a harness.

## Start with `auto`

`auto` chooses from frontier open models for each new user turn. It is the default on a standard Fireworks key.

Use `auto-instant` when you want a latency-first open-model mix. If fast models are disabled by model governance, use `auto`.

```bash wrap theme={null}
fireconnect claude
```

Restart Claude Code, open `/model`, and select **Auto**. No `--model` flag is required.

## Choose a family, tier, or version

Use `--model` only when you want a specific family, tier, or version:

```bash wrap theme={null}
fireconnect claude --model glm-latest
```

| Kind | Example IDs | Behavior |
| - | - | - |
| **Latest** | `kimi-latest`, `glm-latest`, `deepseek-flash-latest` | Tracks the current version in the model family |
| **Fast** | `kimi-fast-latest`, `glm-fast-latest` | Favors lower latency at a higher token price |
| **Pinned** | `kimi-k3`, `glm-5p3`, `glm-5p3-flash` | Stays on the named version |

Prefer `*-latest` unless you need a pinned version.

### Current aliases

These are all serverless aliases as of September 27, 2026. Aliases can move to newer models or be retired, so treat this list as a snapshot. The [serverless models endpoint](#find-a-model-id) always shows the current set, and `fireconnect model list` shows the current coding aliases.

| Alias | Current model | Tier | Input |
| - | - | - | - |
| `kimi-latest` | `kimi-k3` | Standard | Text and images |
| `kimi-fast-latest` | `kimi-k3` | Fast | Text and images |
| `glm-latest` | `glm-5p3` | Standard | Text |
| `glm-fast-latest` | `glm-5p3` | Fast | Text |
| `glm-flash-latest` | `glm-5p3-flash` | Standard | Text and images |
| `deepseek-flash-latest` | `deepseek-v4p1-flash` | Standard | Text and images |
| `minimax-latest` | `minimax-m3` | Standard | Text |
| `qwen-max-latest` | `qwen3p8-max` | Standard | Text and images |

`qwen-max-latest` is not in the coding catalog, so `fireconnect model list` does not show it.

<Note>
  Aliases may change as Fireworks adds and retires models. For example,
  `deepseek-pro-latest` is deprecated. Check the current list before you
  hard-code an alias, as shown in [Find a model ID](#find-a-model-id).
</Note>

## Find a model ID

List current models and aliases with FireConnect:

```bash wrap theme={null}
fireconnect model list
fireconnect model list --search glm
fireconnect model list --refresh
```

The catalog is cached for one hour. `--refresh` fetches the latest copy.

Without FireConnect, call the serverless models endpoint:

```bash wrap theme={null}
curl "https://api.fireworks.ai/v1/serverless/models?use_cases=coding" \
  -H "Authorization: Bearer $FIREWORKS_API_KEY"
```

Each entry's `aliases` field lists the aliases that point to that model, such as `accounts/fireworks/routers/glm-latest` for `glm-5p3`. Use the last part, `glm-latest`, as the model ID. Remove `use_cases=coding` to list every serverless model.

This command lists models for FireConnect. To construct a custom
`firerouter/...` ID without FireConnect, use the
[FireRouter supported models](/nexus/firerouter#supported-models).

## Check image support

`glm-latest` and `glm-fast-latest` are text-only. Image support for other aliases depends on the model currently behind each alias. See the Input column in [Current aliases](#current-aliases), and check `fireconnect model list` before sending images.

If an image reaches a text-only model in Claude Code, use `/rewind` or switch to a vision-capable model such as Kimi or `glm-5p3-flash`.

## Use US-only models

Use the current model IDs, endpoint, and pricing guidance from [US-only Serverless](/serverless/us-only-serverless#available-models).

## Pricing

Open models bill to your Fireworks account at [Serverless Pricing](/serverless/pricing). For per-session estimates and account usage, see [Usage and Cost](/nexus/metrics).
