Gemini 3.1 Pro Preview Customtools
Google logo

Gemini 3.1 Pro Preview Customtools

gemini-3.1-pro-preview-customtoolsllms.txt
Google
New
gemini-3.1-pro-preview-customtools For users who build applications mixing bash and custom tools, the Gemini 3.1 Pro preview provides a separate endpoint accessible via the API call gemini-3.1-pro-preview-customtools. This endpoint is better at prioritizing your custom tools (for example, view_file or search_code). Please note that while gemini-3.1-pro-preview-customtools is optimized for agent workflows that use custom tools and Bash, you may experience quality fluctuations in some use cases that cannot benefit from these tools.

Pricing

TierPricingCache ReadWeb SearchCache Storage
Input<=200K
$2.000$12.000
$0.2/M tokens$0.014/request$4.5/h/M tokens
200K<Input
$4.000$18.000
$0.4/M tokens$0.014/request$4.5/h/M tokens

Input Modalities

  • Text
  • Vision
  • Audio
  • Video
  • PDF

Output Modalities

  • Text

Context length

  • 1.05M tokens

Max output

  • 65.5K tokens

Capabilities

  • Thinking
  • Streaming
  • Tool calling
  • Web search
  • URL context
  • Code interpreter
  • Computer use
  • File search
  • Memory tool
  • Structured outputs
  • Citations
  • Prompt caching
  • Background mode
  • Server-side sessions

Providers

Google AI Studio gemini-3.1-pro-preview-customtools
Pricing$2.000$12.000
Cache Read$0.2/M tokens
Web Search$0.014/request
Cache Storage$4.5/h/M tokens
Pricing$4.000$18.000
Cache Read$0.4/M tokens
Web Search$0.014/request
Cache Storage$4.5/h/M tokens
Context1M
Max output64K
Latency193.9S
Throughput14.8TPS
Uptime
1.00% uptime 2 days ago
0.00% uptime yesterday
4.35% uptime today

Performance for gemini-3.1-pro-preview-customtools

Uptime is the percentage of requests that succeeded over the past 72 hours. AIHubMix continuously monitors every provider and automatically retries with the next-best provider when one returns an error or responds too slowly; Latency is total round-trip time (lower is better); Throughput is how fast the model writes (tokens per second, higher is better).

Uptime
Loading...
Latency
Loading...
Throughput
Loading...

Try this model

Python
import os
from openai import OpenAI

client = OpenAI(
    api_key=os.environ["AIHUBMIX_API_KEY"],
    base_url="https://aihubmix.com/v1",
)

response = client.chat.completions.create(
    model="gemini-3.1-pro-preview-customtools",
    messages=[
      {
        "role": "user",
        "content": "Hello, how are you?"
      }
    ],
    max_tokens=1024,
    stream=False,
)

print(response.choices[0].message.content)

Frequently asked questions

What is Gemini 3.1 Pro Preview Customtools?

gemini-3.1-pro-preview-customtools For users who build applications mixing bash and custom tools, the Gemini 3.1 Pro preview provides a separate endpoint accessible via the API call gemini-3.1-pro-preview-customtools. This endpoint is better at prioritizing your custom tools (for example, view_file or search_code). Please note that while gemini-3.1-pro-preview-customtools is optimized for agent workflows that use custom tools and Bash, you may experience quality fluctuations in some use cases that cannot benefit from these tools.