Amux

Model comparison

Compare models side by side on price, context, performance, endpoints and features. Each column is one model on one provider — put the same model twice to compare providers.

Chat

Overview

DeveloperQwen
Context1M token
Max output131K token
ReasoningNot supported
Input
Output

Pricing

Input$0.5$0.4/1M tokens
Output$3$2.4/1M tokens
Readexplicit$0.05$0.04/1M tokensimplicit$0.1$0.08/1M tokens
Write$0.625$0.5/1M tokens
Web search$0.01$0.008/request

Performance

Avg. time to first token
Avg. throughput
Availability

Averages over the recent window, weighted by request volume — not percentiles. Hours with no successful samples are left out rather than counted as zero.

Endpoints

OpenAI Chat Completions

/v1/chat/completions

OpenAI Responses

/v1/responses

Anthropic Messages

/v1/messages

Google Gemini

/v1beta/models/{model}:generateContent

OpenAI Images

/v1/images/generations
/v1/images/edits

Alibaba Model Studio Video

/api/v1/services/aigc/video-generation/video-synthesis
/api/v1/tasks/{id}

MiniMax Video

/v2/video_generation
/v2/query/video_generation/{id}

Alibaba Model Studio Image

/api/v1/services/aigc/multimodal-generation/generation
/api/v1/services/aigc/image-generation/generation

Amux Tasks

/v1/tasks
/v1/tasks/{id}
/v1/tasks/{id}/cancel

Features

Streaming
Reasoning
Function calling
Structured output
Prompt caching
Vision
Image output
Video output