GPT Chat Latest points to OpenAI's stable API alias chat-latest that always resolves to the latest Instant chat model used in ChatGPT. As OpenAI rolls out new Instant model updates in the future, they are routed behind this slug automatically.
Pricing
Input Modalities
- Text
- Vision
Output Modalities
- Text
Context length
- 400K tokens
Max output
- 128K tokens
Capabilities
- Thinking
- Streaming
- Tool calling
- Web search
- URL context
- Code interpreter
- Computer use
- File search
- Memory tool
- Structured outputs
- Citations
- Prompt caching
- Background mode
- Server-side sessions
Providers
Azure gpt-chat-latest
Pricing$5.000$30.000
Cache Read$0.5/M tokens
Web Search$0.01/request
Pricing$10.000$45.000
Cache Read$1/M tokens
Web Search$0.01/request
Context1M
Max output128K
Latency2.3S
Throughput50.5TPS
Uptime
100.00% uptime 2 days ago
100.00% uptime yesterday
100.00% uptime today
Performance for gpt-chat-latest
Uptime is the percentage of requests that succeeded over the past 72 hours. AIHubMix continuously monitors every provider and automatically retries with the next-best provider when one returns an error or responds too slowly; Latency is total round-trip time (lower is better); Throughput is how fast the model writes (tokens per second, higher is better).
Uptime
Loading...
Latency
Loading...
Throughput
Loading...
Try this model
Python
