MULTI-MODEL API

Leading AI model families behind one API account

Keep your application architecture stable while comparing model quality, latency and cost. API Models supports familiar OpenAI, Anthropic and Gemini request formats.

One Base URLMultiple providersLive model availability
Live catalog

GPT

OpenAI model families for general language, reasoning, coding, multimodal and tool-driven workloads.

OpenAI
Live catalog

Claude

Anthropic model families for long-form reasoning, coding, document analysis and agent workflows.

Anthropic
Live catalog

Gemini

Google model families for multimodal input, long context, reasoning and content generation.

Google
Live catalog

DeepSeek

DeepSeek model families for reasoning, coding, agents and cost-sensitive production workloads.

DeepSeek
Check availability

Kimi

Moonshot AI model families for reasoning, long-context work, coding and tool-enabled applications.

Moonshot AI
Live catalog

Qwen

Alibaba model families for multilingual, coding, multimodal and general-purpose AI workloads.

Alibaba
Live catalog

GLM

Zhipu AI model families for multilingual reasoning, coding and tool-based applications.

Zhipu AI
Live catalog

MiniMax

MiniMax model families for text, long-context and multimodal generation workloads.

MiniMax
Live catalog

Grok

xAI model families available through the exact IDs and endpoints shown in the live catalog.

xAI

Availability belongs to the exact model ID

A provider family on this page describes the catalog scope; it does not guarantee that every upstream model or version is available. Open Models and Pricing to verify the exact ID, supported endpoints, current availability and billing.

Copy, do not guess

Model IDs can include dates, capability labels and provider-specific suffixes. Copy the full value from the console before sending a request.

Choose the request format separately

FormatPrimary routeGuide
OpenAI-compatible/v1/responses or /v1/chat/completionsResponses and Chat Completions
Anthropic Messages/v1/messagesMessages API guide
Gemini generateContent/v1beta/models/{model}:generateContentgenerateContent guide

Evaluate models by completed task

  1. Define a redacted evaluation set that matches real usage.
  2. Compare quality and failure modes, not only benchmark summaries.
  3. Record time to first token and total latency at realistic concurrency.
  4. Measure total cost per successful task, including retries and long outputs.
  5. Keep a fallback model and test it with the same request format.

Find the exact model you can call

The live catalog contains current IDs, prices and endpoint support.

Open live model catalog