Model comparison
Compare models side by side on price, context, performance, endpoints and features. Each column is one model on one provider — put the same model twice to compare providers.
Overview
DeveloperDeepSeek
DeveloperDeepSeek
Context1M token
Context1M token
Max output393K token
Max output393K token
ReasoningSupported
ReasoningSupported
Input
Input
Output
Output
Pricing
InputPeak$1.32/1M tokensOff-Peak$0.66/1M tokens
InputOff-Peak$0.66$0.561/1M tokensPeak$1.32$1.122/1M tokens
OutputPeak$3.96/1M tokensOff-Peak$1.98/1M tokens
OutputOff-Peak$1.98$1.683/1M tokensPeak$3.96$3.366/1M tokens
ReadPeak$0.044/1M tokensOff-Peak$0.022/1M tokens
ReadOff-Peak$0.066$0.0561/1M tokensPeak$0.132$0.1122/1M tokens
Performance
Avg. time to first token1.64 s
Avg. time to first token3.42 s
Avg. throughput63.0 tok/s
Avg. throughput52.4 tok/s
Availability100.0%
Availability100.0%
Averages over the recent window, weighted by request volume — not percentiles. Hours with no successful samples are left out rather than counted as zero.
Endpoints
OpenAI Chat Completions
/v1/chat/completions/v1/chat/completionsOpenAI Responses
/v1/responses/v1/responsesAnthropic Messages
/v1/messages/v1/messagesGoogle Gemini
/v1beta/models/{model}:generateContent/v1beta/models/{model}:generateContentOpenAI Images
/v1/images/generations/v1/images/generations/v1/images/edits/v1/images/editsAlibaba Model Studio Video
/api/v1/services/aigc/video-generation/video-synthesis/api/v1/services/aigc/video-generation/video-synthesis/api/v1/tasks/{id}/api/v1/tasks/{id}MiniMax Video
/v2/video_generation/v2/video_generation/v2/query/video_generation/{id}/v2/query/video_generation/{id}Amux Tasks
/v1/tasks/v1/tasks/v1/tasks/{id}/v1/tasks/{id}/v1/tasks/{id}/cancel/v1/tasks/{id}/cancelFeatures
Streaming
Streaming
Reasoning
Reasoning
Function calling
Function calling
Structured output
Structured output
Prompt caching
Prompt caching
Vision
Vision
Image output
Image output
Video output
Video output