Skip to main content

ScaleMax models

Human-readable model names and copyable API IDs are not the same thing.

Use the display name to compare models, then copy the exact API ID from the live authenticated catalog. Availability follows the key’s scope and active production routes.

Catalog rules

Three checks before selecting a model.

01

Check key scope

Claude-only, OpenAI-only, combined, Gemini, Grok, Qwen, Local Models, API-credit, and exact-model keys can return different catalogs.

02

Copy the API ID

Client configuration uses the exact machine-readable ID, including punctuation and version suffixes.

03

Check capabilities

Tools, images, reasoning controls, context limits, and output limits are model- and route-specific.

Reference names

Models represented in the shared ScaleMax catalog.

This reference explains names and IDs. Local Models lists every id in the live family roster. It is not an availability guarantee; the live catalog filters models to currently active routes for the requesting key.

Anthropic

7 reference IDs

Claude Haiku 4.5

claude-haiku-4-5-20251001

Fast Anthropic model for low-latency chat and everyday work

Claude Opus 4.6

claude-opus-4-6

High-capability Anthropic reasoning model for complex work

Claude Opus 4.7

claude-opus-4-7

High-capability Anthropic model for advanced coding and reasoning

Claude Opus 4.8

claude-opus-4-8

Premium Anthropic model for deep analysis and demanding tasks

Claude Opus 5

claude-opus-5

Latest premium Anthropic model for advanced coding, agents, and deep reasoning

Claude Sonnet 5

claude-sonnet-5

Balanced Anthropic model for strong reasoning and production workflows

Fable 5

fable-5

Anthropic-family Fable model through ScaleMax

ScaleMax

9 reference IDs

Codex Auto Review

codex-auto-review

ScaleMax-optimized automated code review and PR analysis

GPT-5.3 Codex

gpt-5.3-codex

Budget-friendly coding assistant

GPT-5.4

gpt-5.4

Advanced long-context coding and analysis

GPT-5.4 mini

gpt-5.4-mini

Fast low-cost everyday development

GPT-5.5

gpt-5.5

Premium reasoning and agentic coding

GPT-5.6 Sol

gpt-5.6-sol

Available through ScaleMax

GPT-5.6 Terra

gpt-5.6-terra

Available through ScaleMax

GPT-6 astra

gpt-6-astra

Available through ScaleMax

GPT-5.6 Luna

gpt-5.6-luna

Available through ScaleMax

Gemini

5 reference IDs

Gemini 3 Flash Preview

gemini-3-flash-preview

Preview Google Gemini 3 model tuned for fast responses

Gemini 3 Pro Preview

gemini-3-pro-preview

Preview Google Gemini 3 model for advanced reasoning and long context

Gemini 3.1 Flash Lite Preview

gemini-3.1-flash-lite-preview

Preview lightweight Google Gemini model for high-volume work

Gemini 3.1 Pro Preview

gemini-3.1-pro-preview

Preview Google Gemini model for demanding reasoning and agentic work

Gemini 3.6 Flash

gemini-3.6-flash

Latest fast Google Gemini model for everyday coding and chat

Grok

2 reference IDs

Grok 4.5

grok-4.5

xAI Grok model for reasoning, coding, and general assistance

Grok Build

grok-build

xAI Grok model for coding and agentic work

Qwen

2 reference IDs

Qwen 3.6 Plus

qwen3.6-plus

Alibaba Qwen model for general reasoning and coding

Qwen 3.7 Plus

qwen3.7-plus

Latest Alibaba Qwen model for general reasoning and coding

Local Models

51 reference IDs

DeepSeek V4 Flash

deepseek-v4-flash

Fast general-purpose model for everyday coding and chat

DeepSeek V4 Flash New

deepseek-v4-flash-new

Newer DeepSeek V4 Flash generation with a 1M-token context window

DeepSeek V4 Flash 0731 Unlimited

dsv4-flash-0731-unlimited

DeepSeek V4 Flash 0731 build with tool calling, sold in the Local Models pack

Sonnet 4.6

claude-sonnet-4-6

Fast general-purpose model with tool calling, sold in the Local Models pack

DeepSeek V4 Flash 1M

deepseek-v4-flash-1m

Long-context general-purpose model for large codebases and documents

DeepSeek V4 Pro

deepseek-v4-pro

High-capability general-purpose model for demanding reasoning and coding

deepseek-v4-pro-0813

deepseek-v4-pro-0813

DeepSeek · Local Models pack

DeepSeek V3.2

deepseek-v3.2

General-purpose model for reasoning and coding

deepseek-v4-pro-v3

deepseek-v4-pro-v3

DeepSeek · Local Models pack

deepseek-v4-flash-v3

deepseek-v4-flash-v3

DeepSeek · Local Models pack

Gemini 3 Flash Preview

gemini-3-flash-preview

Preview general-purpose model tuned for fast responses

Gemini 3 Pro Preview

gemini-3-pro-preview

Preview general-purpose model for advanced reasoning and long context

Gemini 3.1 Flash Lite Preview

gemini-3.1-flash-lite-preview

Preview lightweight general-purpose model for high-volume work

Gemini 3.1 Pro Preview

gemini-3.1-pro-preview

Preview general-purpose model for demanding reasoning and agentic work

Gemini 3.6 Flash

gemini-3.6-flash

Latest fast general-purpose model for everyday coding and chat

gemini-3.7-flash

gemini-3.7-flash

Gemini · Local Models pack

Gemini 3.8 Flash

gemini-3.8-flash

Fast general-purpose model with tool calling, sold in the Local Models pack

Grok 4.5

grok-4.5

General-purpose model for reasoning, coding, and everyday assistance

Grok 4.6

grok-4.6

Latest general-purpose model for reasoning, coding, and everyday assistance

Grok Build

grok-build

Grok model for coding and agentic work

glm-5

glm-5

GLM · Local Models pack

glm-5.1

glm-5.1

GLM · Local Models pack

GLM 5.2

glm-5.2

Balanced general-purpose model for production workflows

glm-5.3

glm-5.3

GLM · coming soon (priced, not yet routed)

glm-5.3-flash

glm-5.3-flash

GLM · Local Models pack

glm-5.2-v2

glm-5.2-v2

GLM · Local Models pack

glm-5.3-v2

glm-5.3-v2

GLM · Local Models pack

glm-5.3-flash-v2

glm-5.3-flash-v2

GLM · Local Models pack

glm-5.3-flash-v3

glm-5.3-flash-v3

GLM · Local Models pack

MiniMax M2.5

minimax-m2.5

General-purpose model for reasoning and coding

MiniMax M2.7

minimax-m2.7

Latest general-purpose model for reasoning and coding

minimax-m3

minimax-m3

MiniMax · Local Models pack

Kimi K2.6

kimi-k2.6

Long-context general-purpose model for extended conversations and large codebases

kimi-k2.7-code

kimi-k2.7-code

Kimi · Local Models pack

kimi-k3

kimi-k3

Kimi · Local Models pack

Hunyuan 3

hunyuan-3

General-purpose model for reasoning and everyday coding

Hunyuan 4

hy4

Reasoning model with tool calling, sold in the Local Models pack

hy4-preview

hy4-preview

Hunyuan · Local Models pack

MiMo V2.5

mimo-v2.5

Lightweight general-purpose model for fast, low-cost work

mimo-v2.5-pro

mimo-v2.5-pro

MiMo · Local Models pack

mistral-large-3

mistral-large-3

Mistral · Local Models pack

nemotron-lightning-3.5

nemotron-lightning-3.5

Nemotron · Local Models pack

qwen-3.7-flash

qwen-3.7-flash

Qwen · Local Models pack

qwen-3.7-max

qwen-3.7-max

Qwen · Local Models pack

qwen-3.8-27b

qwen-3.8-27b

Qwen · Local Models pack

qwen3.8

qwen3.8

Qwen · Local Models pack

qwen3.8-max-v2

qwen3.8-max-v2

Qwen · Local Models pack

muse-spark-1.2

muse-spark-1.2

Muse · Local Models pack

step-3.5-flash

step-3.5-flash

StepFun · Local Models pack

LongCat 2.0

longcat-2.0

Free reasoning model with tool calling, sold in the Local Models pack

gpt-oss-120b

gpt-oss-120b

OpenAI · Local Models pack

The authenticated response wins.

Before a production request, fetch the catalog with the same key and base URL the client will use.

Check current models