GPT 4o 2024 05-13
OpenAI logo

GPT 4o 2024 05-13

gpt-4o-2024-05-13llms.txt
OpenAI

Pricing

PricingCache ReadImage GenerationWeb Search
$5.000$15.000
$5/M tokens-$0.025/request

Input Modalities

    Providers

    Azure gpt-4o-2024-05-13
    Pricing$5.000$15.000
    Cache Read$5/M tokens
    Web Search$0.025/request
    Context128K
    Max output4K
    Latency1.2S
    Throughput198.9TPS
    Uptime
    100.00% uptime 2 days ago
    0.00% uptime yesterday
    0.00% uptime today
    OpenAI gpt-4o-2024-05-13
    Pricing$5.000$15.000
    Cache Read$5/M tokens
    Web Search$0.025/request
    Context128K
    Max output4K
    Latency0.5S
    Throughput78.2TPS
    Uptime
    0.00% uptime 2 days ago
    0.00% uptime yesterday
    0.00% uptime today

    Performance for gpt-4o-2024-05-13

    Uptime is the percentage of requests that succeeded over the past 72 hours. AIHubMix continuously monitors every provider and automatically retries with the next-best provider when one returns an error or responds too slowly; Latency is total round-trip time (lower is better); Throughput is how fast the model writes (tokens per second, higher is better).

    Uptime
    Loading...
    Latency
    Loading...
    Throughput
    Loading...

    Try this model

    Python
    import os
    from openai import OpenAI
    
    client = OpenAI(
        api_key=os.environ["AIHUBMIX_API_KEY"],
        base_url="https://aihubmix.com/v1",
    )
    
    response = client.chat.completions.create(
        model="gpt-4o-2024-05-13",
        messages=[
          {
            "role": "user",
            "content": "Hello, how are you?"
          }
        ],
        max_tokens=1024,
        stream=False,
    )
    
    print(response.choices[0].message.content)

    Frequently asked questions

    What is the context length of GPT 4o 2024 05-13?

    GPT 4o 2024 05-13 has a 128,000 token context window. It supports up to 4,096 output tokens.