Gemini 3.1 Pro Preview Custom Tools

Compare with another model

Google's tool-calling-focused Gemini 3.1 Pro preview, built for multi-step agents with a 1M-token context and $2/$12 per million pricing.

Key specifications

Capability
88.5
Context window
1.0M
Max output
65.5K
Input $/1M
$2.00
Output $/1M
$12.00
License
Proprietary

Inputs and outputs

Recorded modality support shows what this model can accept and produce. An undocumented modality is shown as unknown rather than assumed to be unsupported.

Accepts

  • Text input
  • Image input
  • Video input
  • Audio input
  • PDF input

Produces

  • Text output

Capabilities

ReasoningFunction callingVisionAudioJSON mode

What is Gemini 3.1 Pro Preview Custom Tools?

gemini 3.1 pro preview custom tools is the tool-calling-oriented release in Google's 3.1 Pro line — same reasoning core as the standard Pro preview, but positioned for agents that spend most of a conversation calling out to external functions rather than talking. If you're wiring up a multi-step agent that hits a database, a search API, and a code sandbox in sequence, this is the variant built with that loop in mind.

What "custom tools" changes

The underlying specs match Gemini 3.1 Pro Preview: a 1,048,576 token context window, up to 65,536 tokens of output, and full support for function calling, vision, and reasoning. What the naming signals is intent — this build is meant for agent frameworks and orchestration layers that define their own tool schemas and expect the model to reason across several tool calls in one turn, rather than for single-shot chat or document Q&A.

Pricing and how the context window plays into it

Input runs $2.00 per million tokens, output $12.00 per million, and cached input drops to $0.20 per million. That cached rate matters more here than in most Gemini tiers: a tool-calling agent typically resends the same system prompt, tool definitions, and conversation history on every step, and those are exactly the tokens caching discounts. A long-running agent session that would otherwise re-bill the same 5,000-token tool manifest on every step gets a 10x break on that portion once it's cached.

Where it fits next to the rest of the family

  • Multi-step agent loops with several tool calls per turn
  • Workflows built on custom function/tool schemas rather than a default toolset
  • Tasks that need the reasoning depth of Pro, not the throughput pricing of Flash Lite
  • Long sessions where cached system prompts and tool definitions offset the $2.00/$12.00 rate

Against Gemini 3.1 Flash Lite Preview's $0.25/$1.50 pricing, this model costs 8x more per token in both directions — pay it when the task genuinely needs Pro-level reasoning across tool calls, not for high-volume simple extraction.

The honest limitation

It's a preview build, so treat pricing and rate limits as subject to change before general availability. There's no published benchmark data attached to this release, so any claim about how it performs against other tool-calling models should come from your own eval on your own tool schema, not from a spec sheet.

How to evaluate Gemini 3.1 Pro Preview Custom Tools

Use the recorded facts as a shortlist, then validate the model against representative inputs and production constraints.

Workload fit

Gemini 3.1 Pro Preview Custom Tools is categorized for Audio-text-to-text, Automatic speech recognition, Image-text-to-text, Image-to-text, Text generation, Text-to-text AI models, Video-text-to-text. Its current record accepts text, image, video, audio, pdf and produces text. Confirm file formats, preprocessing, and provider-specific request schemas before implementation.

Capacity and cost

The directory records 1.0M context and 65.5K maximum output. Listed token prices are $2.00 input and $12.00 output per 1M tokens. Treat missing values as unknown and recheck current commercial terms.

Operational behavior

Recorded output speed is — and time to first token is —. Hosting route, region, prompt length, concurrency, and provider load can materially change both measurements.

Evidence boundary

This page separates sourced model facts from incomplete fields. Benchmark evidence is displayed only when its source is linked and verified. Before choosing Gemini 3.1 Pro Preview Custom Tools, test task quality, tool reliability, safety behavior, data controls, rate limits, and total cost on the exact route you intend to use.

Strengths

  • + Registry-backed capabilities
  • + Provider: Google

Limitations

  • Draft record — editorial review required
  • Verify pricing and independent evaluations before approval

Frequently asked questions

Input tokens are $2.00 per million, output tokens are $12.00 per million, and cached input is $0.20 per million — a 10x discount that helps most when an agent resends the same tool definitions and system prompt every step.

Related approved models in Automatic speech recognition. Compare specifications and verify fit for your workload.

Google

Gemini 3.5 Flash Lite

Google's low-cost Gemini 3 model with a 1,048,576-token context window, vision support, and $0.30/$2.50 per million token pricing for volume workloads.

1.0M contextCompare

Google

Gemini 3.6 Flash

Google's mid-tier Gemini 3 model: 1,048,576-token context, multimodal input, and $1.50/$7.50 per million token pricing for production agent workloads.

1.0M contextCompare

Google

Gemini 3.1 Flash Lite

Google's cheapest Gemini 3 model, with a 1,048,576-token context window and $0.25/$1.50 per million token pricing for high-volume, cost-sensitive work.

1.0M contextCompare

Google

Gemini 3.5 Flash

Google's full Gemini 3 Flash model, with a 1,048,576-token context window and $1.50/$9.00 per million token pricing for production-grade agent work.

1.0M contextCompare

Google

Gemini Flash Latest

Google's fast, rolling-updated Gemini model with a 1M-token context window, reasoning, and tool calling at $1.50/$9.00 per million tokens.

1.0M contextCompare

Google

Gemini Flash-Lite Latest

Google's low-cost, high-volume Gemini model at $0.25/$1.50 per million tokens, with a 1M-token context window and reasoning support.

1.0M contextCompare