GPT 4.1 Nano
OpenAI · text, image → text
Ultra-lightweight model with million-token context, optimized for speed and low latency, costing only $0.10 per million input tokens. It is suitable for edge computing and real-time interaction. The automatic caching mechanism offers a 75% cost reduction on cache hits.
Input$0.10 /M
Output$0.40 /M
