Amux

Model comparison

Compare models side by side on price, context, performance, endpoints and features. Each column is one model on one provider — put the same model twice to compare providers.

Chat

Overview

DeveloperOpenAI
DeveloperOpenAI
Context1.05M token
Context1.05M token
Max output128K token
Max output128K token
ReasoningSupported
ReasoningSupported
Input
Input
Output
Output

Pricing

Input≤272K$5$4/1M tokens>272K$10$8/1M tokens
Input≤272K$5$0.25/1M tokens>272K$10$0.5/1M tokens
Output≤272K$30$24/1M tokens>272K$45$36/1M tokens
Output≤272K$30$1.5/1M tokens>272K$45$2.25/1M tokens
Read≤272K$0.5$0.4/1M tokens>272K$1$0.8/1M tokens
Read≤272K$0.5$0.025/1M tokens>272K$1$0.05/1M tokens
Write≤272K$6.25$5/1M tokens>272K$12.5$10/1M tokens
Write≤272K$6.25$0.3125/1M tokens>272K$12.5$0.625/1M tokens
Web search$0.01$0.008/request
Web search$0.01$0.0005/request

Performance

Avg. time to first token2.96 s
Avg. time to first token5.80 s
Avg. throughput57.4 tok/s
Avg. throughput45.2 tok/s
Availability99.8%
Availability99.9%

Averages over the recent window, weighted by request volume — not percentiles. Hours with no successful samples are left out rather than counted as zero.

Endpoints

OpenAI Chat Completions

/v1/chat/completions
/v1/chat/completions

OpenAI Responses

/v1/responses
/v1/responses

Anthropic Messages

/v1/messages
/v1/messages

Google Gemini

/v1beta/models/{model}:generateContent
/v1beta/models/{model}:generateContent

OpenAI Images

/v1/images/generations
/v1/images/generations
/v1/images/edits
/v1/images/edits

Alibaba Model Studio Video

/api/v1/services/aigc/video-generation/video-synthesis
/api/v1/services/aigc/video-generation/video-synthesis
/api/v1/tasks/{id}
/api/v1/tasks/{id}

MiniMax Video

/v2/video_generation
/v2/video_generation
/v2/query/video_generation/{id}
/v2/query/video_generation/{id}

Amux Tasks

/v1/tasks
/v1/tasks
/v1/tasks/{id}
/v1/tasks/{id}
/v1/tasks/{id}/cancel
/v1/tasks/{id}/cancel

Features

Streaming
Streaming
Reasoning
Reasoning
Function calling
Function calling
Structured output
Structured output
Prompt caching
Prompt caching
Vision
Vision
Image output
Image output
Video output
Video output