Quick StartModel List

Model Overview

OriginRouter aggregates dozens of models from OpenAI, Anthropic, Google, Zhipu, DeepSeek, MiniMax, Kimi, and more — covering conversation, reasoning, coding, embeddings, and image generation. This page gives you a quick overview of the model categories and how to choose between them.


Model Categories

International flagship chat models

Top-tier general conversation and reasoning, ideal for complex tasks, deep analysis, and high-quality content generation.

  • GPT-5.5 / GPT-5.4 / GPT-5.2 — OpenAI's latest generation, spanning flagship to efficient performance tiers
  • Claude Opus 4.6 / Sonnet 4.6 / Haiku 4.5 — Anthropic's Claude series; Opus for the most complex reasoning, Sonnet excels at coding, Haiku is fast and lightweight
  • Gemini 3.1 Pro / Flash — Google multimodal models with long context windows (up to 1M tokens)

Coding and reasoning models

Deeply optimized for code generation, math reasoning, and logical analysis.

  • GPT-5.4 / GPT-5.3 Codex — OpenAI's code-optimized series, great for understanding and refactoring large codebases
  • Claude Sonnet 4.6 — Top performer on coding benchmarks; the primary recommended model for the coding plan
  • DeepSeek V3.2 — Open-source model with strong competitiveness in coding and math reasoning

Domestic models

Unique advantages in Chinese comprehension, local knowledge, and cost-effectiveness.

  • GLM-5.1 / GLM-4.7 / GLM-4.5-Air — Zhipu's GLM series, strong Chinese understanding; Air is lightweight and efficient
  • Kimi K2.6 / K2.5 / K2 Thinking — Moonshot series; Thinking variant excels at deep reasoning
  • MiniMax M2.7 / M2.5 / M2 — MiniMax series with well-rounded general capabilities

Specialized models

  • Embedding models: text-embedding-3-large / text-embedding-3-small — convert text to vectors for semantic search and clustering
  • Image models: gpt-image-1 / gemini-2.5-flash-image-preview — generate images from text descriptions

How to Choose a Model

By use case

Use caseRecommended modelsNotes
Everyday codingclaude-sonnet-4-6, gpt-5.4Best coding experience, balanced performance and speed
Lightweight coding / completionclaude-haiku-4-5, gpt-5.4-miniFast responses, quota-efficient
Architecture design / complex reasoningclaude-opus-4-6, gpt-5.5Maximum reasoning power
Chinese content creationGLM-5.1, kimi-k2.6Natural, fluent Chinese output
Long document processinggemini-3.1-pro-preview1M token context window
Image generationgpt-image-1, gemini-2.5-flash-image-previewMultimodal capability

Cost vs. performance trade-offs

  • Performance first: choose models on direct or hybrid routing, such as claude-opus-4-6 or gpt-5.5
  • Cost-efficiency first: use third-party channels or domestic models — costs can drop by about 85-95%
  • Quota-conscious: use claude-haiku-4-5 or GLM-4.5-Air for simple tasks, then switch to flagship models when needed

Discover Available Models

Query via API

Call the GET /beta/v1/models endpoint to get the full list of currently available models, including model ID, owner, and creation time.

curl https://api.easytransnote.com/beta/v1/models \
  -H "Authorization: Bearer {YOUR_API_KEY}"

Browse in the console

Log in to the developer console to visually browse all models along with their routing status and pricing.

Check the coding plan model table

Coding plan users can visit Model Routing Details to see the routing strategy for each model at different subscription tiers.


Model Routing

A single model on OriginRouter is usually connected to multiple upstream providers. You can use the system-recommended default route, or pin a specific provider by appending a route suffix. If the request does not specify a route suffix, the system automatically uses the model's default route to handle the call.

Default vs. Explicit

Default — when model is <model_id>, OriginRouter calls the model on its default provider.

{ "model": "claude-opus-4-6" }

Explicit — when model is <model_id>:<router>, the system uses that provider's instance of the model.

{ "model": "claude-opus-4-6:anthropic" }

Supported Endpoints

Model routing is currently available on the following inference endpoints:

  • /beta/v1/chat/completions
  • /beta/v1/messages
  • /beta/v1/responses

If you call an endpoint that does not support routing, the system returns 404 invalid_model. See Errors for details.

Model Fallback

Beta endpoints enable experimental features by default, such as model fallback. Even with a valid <model_id>:<router>, your preferred provider may still fail (e.g. on internal thinking signatures or tool-call records) and the system will fall through to the next available provider.

You can force-disable fallback by setting fallback: "disabled" in the request body, or read more about model fallback.

To see which provider actually handled a request, check the response headers:

  • X-Originrouter-Model-Id — the model ID actually used
  • X-Originrouter-Actual-Route — the model provider actually used

Pricing

Different providers may price the same model differently. The Console → Model List page lists, for each model, the model ID, the default provider, all supported providers, and the price on each provider.


Next Steps

Now that you know what models are available, see how to access them at the best price through the coding plan.

Was this page helpful?