Model comparison
Compare models side by side on price, context, performance, endpoints and features. Each column is one model on one provider — put the same model twice to compare providers.
Overview
DeveloperGoogle
Context1.05M token
Max output66K token
ReasoningSupported
Input
Output
Pricing
Input≤200K$2$1.7/1M tokens>200K$4$3.4/1M tokens
Output≤200K$12$10.2/1M tokens>200K$18$15.3/1M tokens
Read≤200K$0.2$0.17/1M tokens>200K$0.4$0.34/1M tokens
Web search$0.014$0.0119/request
Performance
Avg. time to first token—
Avg. throughput—
Availability—
Averages over the recent window, weighted by request volume — not percentiles. Hours with no successful samples are left out rather than counted as zero.
Endpoints
OpenAI Chat Completions
/v1/chat/completionsOpenAI Responses
/v1/responsesAnthropic Messages
/v1/messagesGoogle Gemini
/v1beta/models/{model}:generateContentOpenAI Images
/v1/images/generations/v1/images/editsAlibaba Model Studio Video
/api/v1/services/aigc/video-generation/video-synthesis/api/v1/tasks/{id}MiniMax Video
/v2/video_generation/v2/query/video_generation/{id}Amux Tasks
/v1/tasks/v1/tasks/{id}/v1/tasks/{id}/cancelFeatures
Streaming
Reasoning
Function calling
Structured output
Prompt caching
Vision
Image output
Video output