Available models

Explore each model's capabilities, pricing, model ID, and supported protocols. Click a protocol badge in a model card to copy its corresponding base URL.

37 / 37

Text Generation

36

Qwen3.7 Flash

Alibaba Cloud

Text Generation

A fast multimodal Qwen3.7 model with stronger vision, agent execution, and coding throughput than Qwen3.6 Flash. Built for low-latency multimodal understanding and agent workflows.

Context

1M

Max output

128K

Price / MTok

In$0.04Out$0.16

Model ID

qwen3.7-flash

Protocols

Qwen3.7 Max

Alibaba Cloud

Text Generation

The Qwen3.7 flagship, optimized for coding, productivity, and agent-centric workloads. Its agent-oriented design emphasizes coding, office productivity, and complex multi-step work.

Context

1M

Max output

128K

Price / MTok

In$2.5Out$7.5

Model ID

qwen3.7-max

Protocols

Qwen3.7 Plus

Alibaba Cloud

Text Generation

A cost-efficient multimodal Qwen model for general reasoning, vision, and production applications. It combines text and image understanding with a cost-conscious design for production-scale multimodal reasoning.

Context

1M

Max output

128K

Price / MTok

In$0.4Out$1.6

Model ID

qwen3.7-plus

Protocols

Qwen3.8 Flash

Alibaba Cloud

Text Generation

Context

1M

Max output

128K

Price / MTok

In$0.16Out$0.47

Model ID

qwen3.8-flash

Protocols

Qwen3.8 Max

Alibaba Cloud

Text Generation

A 2.4T-parameter MoE flagship for coding, office productivity, vision, and long-horizon agents. Native multimodal understanding supports long documents and video, with strong end-to-end delivery on professional tasks.

Context

1M

Max output

128K

Price / MTok

In$2.5Out$7.5

Model ID

qwen3.8-max

Protocols

Claude Fable 5

Anthropic

Text Generation

An Anthropic model built for autonomous knowledge work, coding, and multimodal agents. It is positioned for autonomous knowledge work and coding with multimodal input, reasoning, and tool-oriented workflows.

Context

1M

Max output

128K

Price / MTok

In$10Out$50

Model ID

claude-fable-5

Protocols

Claude Haiku 4.5

Anthropic

Text Generation

A fast, efficient Claude model for latency-sensitive chat, coding, and agent tasks. It targets near-frontier intelligence at substantially lower latency and cost for high-volume interactive workloads.

Context

200K

Max output

64K

Price / MTok

In$1Out$5

Model ID

claude-haiku-4.5

Protocols

Claude Opus 4.6

Anthropic

Text Generation

A powerful Claude model for end-to-end coding and long-running professional workflows. It is optimized for agents that operate across complete software and professional workflows instead of isolated prompts.

Context

1M

Max output

128K

Price / MTok

In$5Out$25

Model ID

claude-opus-4.6

Protocols

Claude Opus 4.7

Anthropic

Text Generation

An Opus model designed for long-running asynchronous agents and complex software workflows. The model is designed for asynchronous agents that sustain coding and professional work over extended runs.

Context

1M

Max output

128K

Price / MTok

In$5Out$25

Model ID

claude-opus-4.7

Protocols

Claude Opus 4.8

Anthropic

Text Generation

A high-capability Opus model for long-context reasoning, coding, and multimodal analysis. It combines multimodal input, reasoning support, and a long context window for demanding general-purpose workflows.

Context

1M

Max output

128K

Price / MTok

In$5Out$25

Model ID

claude-opus-4.8

Protocols

Claude Opus 5

Anthropic

Text Generation

Anthropic's flagship for demanding reasoning, coding, visual analysis, and long-running agents. It is especially capable at end-to-end software tasks, code review and bug finding, document and chart analysis, complex office deliverables, tool use, and coordinating parallel subagents.

Context

1M

Max output

128K

Price / MTok

In$5Out$25

Model ID

claude-opus-5

Protocols

Claude Sonnet 4.6

Anthropic

Text Generation

A balanced frontier model for iterative development, agents, and professional knowledge work. Strengths include iterative development, complex codebase navigation, and end-to-end project execution with tools.

Context

1M

Max output

128K

Price / MTok

In$3Out$15

Model ID

claude-sonnet-4.6

Protocols

Claude Sonnet 5

Anthropic

Text Generation

A frontier Sonnet model for coding, professional work, and adaptive agentic reasoning. Adaptive reasoning levels let applications balance frontier coding and agent performance against latency and token use.

Context

1M

Max output

128K

Price / MTok

In$3Out$15

Model ID

claude-sonnet-5

Protocols

DeepSeek V4 Flash

DeepSeek

Text Generation

The official DeepSeek V4 Flash model for fast long-context reasoning, coding, and agents. The 0731 post-training refresh keeps the mixture-of-experts architecture while strengthening agentic coding and tool use.

Context

1M

Max output

384K

Price / MTok

In$0.22Out$0.66

Model ID

deepseek-v4-flash

Protocols

DeepSeek V4 Pro

DeepSeek

Text Generation

The official DeepSeek V4 Pro flagship for advanced reasoning, coding, and sustained agent workloads. The 0813 production checkpoint keeps 1M context and dual thinking modes under the same API name.

Context

1M

Max output

384K

Price / MTok

In$0.66Out$1.98

Model ID

deepseek-v4-pro

Protocols

Gemini 3 Flash Preview

Google

Text Generation

A fast Gemini thinking model for multi-turn chat, coding, and agent workflows. It offers near-Pro reasoning and tool use at Flash-class speed for multi-turn chat, coding, and agents.

Context

1M

Max output

64K

Price / MTok

In$0.5Out$3

Model ID

gemini-3-flash-preview

Protocols

Gemini 3.1 Flash Lite

Google

Text Generation

A high-efficiency Gemini preview optimized for high-volume, latency-sensitive applications. It is tuned for high-volume workloads while approaching the quality of larger Flash models.

Context

1M

Max output

64K

Price / MTok

In$0.25Out$1.5

Model ID

gemini-3.1-flash-lite

Protocols

Gemini 3.1 Pro Preview

Google

Text Generation

A frontier Gemini reasoning preview for software engineering and reliable agentic workflows. The preview improves software-engineering performance, agent reliability, and token efficiency across complex workflows.

Context

1.1M

Max output

65.5K

Price / MTok

In$2Out$12

Model ID

gemini-3.1-pro-preview

Protocols

Gemini 3.5 Flash

Google

Text Generation

An efficient multimodal Gemini model with strong coding, reasoning, and parallel-agent performance. It brings near-Pro coding and reasoning to a faster tier and is tuned for parallel agent execution.

Context

1.1M

Max output

65.5K

Price / MTok

In$1.5Out$9

Model ID

gemini-3.5-flash

Protocols

Gemini 3.5 Flash Lite

Google

Text Generation

A lightweight Gemini model for focused subagents and high-volume production tasks. The lightweight design is especially suited to focused subagents inside larger multi-agent systems.

Context

1M

Max output

64K

Price / MTok

In$0.3Out$2.5

Model ID

gemini-3.5-flash-lite

Protocols

Gemini 3.6 Flash

Google

Text Generation

A high-efficiency Gemini model for coding, agents, and web or application development. It focuses on polished coding and application outputs with fewer unnecessary edits across agentic workflows.

Context

1M

Max output

64K

Price / MTok

In$1.5Out$7.5

Model ID

gemini-3.6-flash

Protocols

Gemini 3.7 Flash

Google

Text Generation

A high-efficiency Gemini workhorse for coding, agents, and knowledge work. It improves first-pass code accuracy, multi-step tool use, and production-ready application output over Gemini 3.6 Flash.

Context

1M

Max output

64K

Price / MTok

In$1.5Out$7.5

Model ID

gemini-3.7-flash

Protocols

Kimi K2.5

Moonshot AI

Text Generation

A native multimodal Kimi model for visual coding and coordinated agent workflows. It pairs native multimodal understanding with visual coding and a self-directed agent-swarm approach.

Context

256K

Max output

256K

Price / MTok

In$0.6Out$3

Model ID

kimi-k2.5

Protocols

Kimi K2.6

Moonshot AI

Text Generation

A multimodal Kimi model for long-horizon coding, UI generation, and agent orchestration. It targets long-horizon coding, code-driven UI generation, and coordinated multi-agent execution across complex projects.

Context

256K

Max output

256K

Price / MTok

In$0.95Out$4

Model ID

kimi-k2.6

Protocols

Kimi K2.7 Code

Moonshot AI

Text Generation

A coding-focused Kimi model for reliable end-to-end programming across long contexts. The coding-focused design aims to complete end-to-end programming tasks reliably across long contexts.

Context

256K

Max output

256K

Price / MTok

In$0.95Out$4

Model ID

kimi-k2.7-code

Protocols

Kimi K3

Moonshot AI

Text Generation

An open-weight multimodal Kimi model for complex coding, reasoning, and agentic work. The open-weight multimodal reasoning model is positioned for complex coding, knowledge work, and long-running agents.

Context

1M

Max output

1M

Price / MTok

In$3Out$15

Model ID

kimi-k3

Protocols

GPT-5.4

OpenAI

Text Generation

A flagship GPT model for advanced reasoning, coding, multimodal input, and professional workflows. It unifies general GPT and coding capabilities for large-context professional work and tool-driven execution.

Context

1.1M

Max output

128K

Price / MTok

In$2.5Out$15

Model ID

gpt-5.4

Protocols

GPT-5.4 mini

OpenAI

Text Generation

A faster GPT-5.4 variant for high-throughput reasoning, coding, and multimodal applications. The smaller variant keeps strong reasoning and coding capabilities while improving speed and throughput.

Context

400K

Max output

128K

Price / MTok

In$0.75Out$4.5

Model ID

gpt-5.4-mini

Protocols

GPT-5.4 nano

OpenAI

Text Generation

A compact GPT model optimized for low-latency, high-volume, and cost-sensitive tasks. It is tuned for speed-critical classification, extraction, routing, and other high-volume lightweight tasks.

Context

400K

Max output

128K

Price / MTok

In$0.2Out$1.25

Model ID

gpt-5.4-nano

Protocols

GPT-5.5

OpenAI

Text Generation

A frontier GPT model for complex professional work with stronger reasoning and reliability. It strengthens reliability and token efficiency on difficult professional tasks while retaining long-context support.

Context

1.1M

Max output

128K

Price / MTok

In$5Out$30

Model ID

gpt-5.5

Protocols

GPT-5.6 Luna

OpenAI

Text Generation

A fast, efficient GPT-5.6 model for chat, classification, and lightweight agents. The efficient tier targets high-volume chat, classification, and lightweight agent workflows with low latency.

Context

1.1M

Max output

128K

Price / MTok

In$0.22Out$1.2

Model ID

gpt-5.6-luna

Protocols

GPT-5.6 Sol

OpenAI

Text Generation

The GPT-5.6 flagship for demanding reasoning, coding, and multi-step agent tasks. The flagship tier is particularly strong at command-line work and complex multi-step coding workflows.

Context

1.1M

Max output

128K

Price / MTok

In$5Out$30

Model ID

gpt-5.6-sol

Protocols

GPT-5.6 Terra

OpenAI

Text Generation

A balanced GPT-5.6 model for everyday coding, reasoning, and agentic applications. The balanced tier sits between flagship quality and cost efficiency for everyday coding, reasoning, and agents.

Context

1.1M

Max output

128K

Price / MTok

In$2.2Out$13.2

Model ID

gpt-5.6-terra

Protocols

Grok 4.6

xAI (Grok)

Text Generation

A frontier Grok model for long-running agents, coding, and knowledge work. It focuses on multi-step agents, ambitious interactive and visual work, and sustained coding trajectories.

Context

500K

Max output

128K

Price / MTok

In$2Out$6

Model ID

grok-4.6

Protocols

GLM-5.2

Zhipu AI

Text Generation

A long-context GLM reasoning model for project-level software engineering and agents. The large-scale reasoning model is intended for project-level software engineering and long-horizon agent workflows.

Context

1M

Max output

128K

Price / MTok

In$1.4Out$4.4

Model ID

glm-5.2

Protocols

GLM-5.3

Zhipu AI

Text Generation

A flagship GLM model for long-horizon coding, agents, and security review. It keeps the GLM-5.2 base and raises coding and agent performance through post-training.

Context

1M

Max output

128K

Price / MTok

In$1.4Out$4.4

Model ID

glm-5.3

Protocols

Image Generation

1

GPT Image 2

OpenAI

Image Generation

OpenAI's high-fidelity image generation and editing model for the dedicated Images API.

Price / MTok

In$8Out$30

Model ID

gpt-image-2

Protocols