gemini-3.1-pro-preview-customtools
For users who build applications mixing bash and custom tools, the Gemini 3.1 Pro preview provides a separate endpoint accessible via the API call gemini-3.1-pro-preview-customtools. This endpoint is better at prioritizing your custom tools (for example, view_file or search_code).
Please note that while gemini-3.1-pro-preview-customtools is optimized for agent workflows that use custom tools and Bash, you may experience quality fluctuations in some use cases that cannot benefit from these tools.
Pricing
Input Modalities
- Text
- Vision
- Audio
- Video
Output Modalities
- Text
Context length
- 1.05M tokens
Max output
- 65.5K tokens
Capabilities
- Thinking
- Streaming
- Tool calling
- Web search
- URL context
- Code interpreter
- Computer use
- File search
- Memory tool
- Structured outputs
- Citations
- Prompt caching
- Background mode
- Server-side sessions
Providers
Google AI Studio gemini-3.1-pro-preview-customtools
Pricing$2.000$12.000
Cache Read$0.2/M tokens
Web Search$0.014/request
Cache Storage$4.5/h/M tokens
Pricing$4.000$18.000
Cache Read$0.4/M tokens
Web Search$0.014/request
Cache Storage$4.5/h/M tokens
Context1M
Max output64K
Latency193.9S
Throughput14.8TPS
Uptime
1.00% uptime 2 days ago
0.00% uptime yesterday
4.35% uptime today
Performance for gemini-3.1-pro-preview-customtools
Uptime is the percentage of requests that succeeded over the past 72 hours. AIHubMix continuously monitors every provider and automatically retries with the next-best provider when one returns an error or responds too slowly; Latency is total round-trip time (lower is better); Throughput is how fast the model writes (tokens per second, higher is better).
Uptime
Loading...
Latency
Loading...
Throughput
Loading...
Try this model
Python
