Amux
All models

Microsoft Azure

Models served by Microsoft Azure

Models · 11

  • OpenAI: GPT-4o
    20% offChat

    GPT-4o ("o" for "omni") is OpenAI's latest AI model, supporting both text and image inputs with text outputs. It maintains the intelligence level of GPT-4 Turbo while being twice as fast and 50% more cost-effective. GPT-4o also offers improved performance in processing non-English languages and enhanced visual capabilities.

    Input:
    Output:
    Input:
    $2.5$2/1M
    Output:
    $10$8/1M
    Context length:
    128K
    Max output:
    16K
    1 provider
  • OpenAI: GPT-5.4
    50% offChat

    GPT-5.4 is OpenAI’s latest frontier model, unifying the Codex and GPT lines into a single system. It features a 1M+ token context window (922K input, 128K output) with support for text and image inputs, enabling high-context reasoning, coding, and multimodal analysis within the same workflow.

    Input:
    Output:
    Input:
    $2.5$1.25/1M
    Output:
    $15$7.5/1M
    Context length:
    1.05M
    Max output:
    128K
    2 providers
  • OpenAI: GPT-5.4 Mini
    50% offChat

    GPT-5.4 mini brings the core capabilities of GPT-5.4 to a faster, more efficient model optimized for high-throughput workloads. It supports text and image inputs with strong performance across reasoning, coding, and tool use, while reducing latency and cost for large-scale deployments.

    Input:
    Output:
    Input:
    $0.75$0.375/1M
    Output:
    $4.5$2.25/1M
    Context length:
    400K
    Max output:
    128K
    2 providers
  • OpenAI: GPT-5.4 Nano
    20% offChat

    GPT-5.4 nano is the most lightweight and cost-efficient variant of the GPT-5.4 family, optimized for speed-critical and high-volume tasks. It supports text and image inputs and is designed for low-latency use cases such as classification, data extraction, ranking, and sub-agent execution.

    Input:
    Output:
    Input:
    $0.2$0.16/1M
    Output:
    $1.25$1/1M
    Context length:
    400K
    Max output:
    128K
    1 provider
  • OpenAI: GPT-5.5
    95% offChat

    GPT-5.5 is OpenAI’s frontier model designed for complex professional workloads, building on GPT-5.4 with stronger reasoning, higher reliability, and improved token efficiency on hard tasks. It features a 1M+ token context window (922K input, 128K output) with support for text and image inputs, enabling large-scale reasoning, coding, and multimodal workflows within a single system.

    Input:
    Output:
    Input:
    $5$0.25/1M
    Output:
    $30$1.5/1M
    Context length:
    1.05M
    Max output:
    128K
    2 providers
  • OpenAI: GPT-5.6 Luna
    50% offChat

    GPT-5.6 Luna is a fast, cost-efficient model in OpenAI's GPT-5.6 series. It is suited for high-volume, latency-sensitive tasks such as chat, classification, and lightweight agentic workflows, providing capable reasoning for its price tier.

    Input:
    Output:
    Input:
    $0.2$0.1/1M
    Output:
    $1.2$0.6/1M
    Context length:
    1.05M
    Max output:
    128K
    2 providers
  • OpenAI: GPT-5.6 Sol
    95% offChat

    GPT-5.6 Sol is the flagship model in OpenAI's GPT-5.6 series. It is suited for complex reasoning, coding, and agentic workflows, and is particularly strong at command-line and multi-step coding tasks and long-horizon problem solving.

    Input:
    Output:
    Input:
    $5$0.25/1M
    Output:
    $30$1.5/1M
    Context length:
    1.05M
    Max output:
    128K
    2 providers
  • OpenAI: GPT-5.6 Terra
    95% offChat

    GPT-5.6 Terra is a balanced model in OpenAI's GPT-5.6 series, positioned between the flagship Sol tier and the cost-efficient Luna tier. It is suited for everyday coding, reasoning, and agentic tasks where capability and cost need to be balanced, offering strong performance at roughly half the cost of Sol.

    Input:
    Output:
    Input:
    $2$0.1/1M
    Output:
    $12$0.6/1M
    Context length:
    1.05M
    Max output:
    128K
    2 providers
  • OpenAI: GPT-6 Astra
    95% offChat

    GPT-6 Astra is OpenAI's flagship model for demanding end-to-end work. It is suited for advanced analysis, software engineering, deep research, scientific work, and document creation, with particular strengths in long-horizon agentic tasks that involve computer and browser use.

    Input:
    Output:
    Input:
    $10$0.5/1M
    Output:
    $50$2.5/1M
    Context length:
    1.05M
    Max output:
    128K
    2 providers
  • OpenAI: GPT-Image-1.5
    20% offChat

    OpenAI’s GPT-Image-1.5 is the latest evolution of its AI image generation, offering superior command compliance, fidelity, text rendering and editing control, making it ideal for detailed creative and production work at a faster and lower cost than previous generations such as DALL-E 3. It excels at handling complex requests, maintains character/style consistency, renders clear text visually, and understands subtle cues through built-in reasoning. It has been integrated into ChatGPT and is available through the API.

    Input:
    Output:
    Input:
    $5$4/1M
    Output:
    $10$8/1M
    Context length:
    10K
    Max output:
    1 provider
  • OpenAI: GPT-Image-2
    20% offChat

    OpenAI's latest image generation model. Supports high-fidelity image generation and editing via the dedicated Images API.

    Input:
    Output:
    Input:
    $5$4/1M
    Output:
    Context length:
    10K
    Max output:
    1 provider