Skip to main content
Google Gemini provides multimodal AI capabilities:
  • Advanced reasoning with Gemini 2.5 and 3.0
  • Multimodal understanding (text, image, audio)
  • Prompt caching for cost optimization
  • Cost-effective options with Flash variants

Available Models

Gemini 3 Pro

Next-generation model with enhanced capabilities and prompt caching.
Gemini 3 Pro supports prompt caching - reuse context at 90% discount.

Gemini 2.5 Pro

Most capable current-gen models for complex tasks.

Gemini 2.5 Flash

Fast, cost-efficient models for high-volume tasks.

Best Practices

Model Selection

Use Gemini 3 for

  • Cutting-edge features
  • Latest capabilities
  • Prompt caching needs
  • Advanced reasoning

Use 2.5 Pro for

  • Complex reasoning tasks
  • Production applications
  • High-quality outputs
  • When accuracy matters

Use Flash for

  • High-volume tasks
  • Fast responses needed
  • Cost-sensitive workloads
  • Simple queries

Prompt Caching (Gemini 3)

Optimize costs with prompt caching on Gemini 3 Pro:
  • Cache Read: $0.26/M (90% cheaper than input)
  • Use case: Repeated system prompts, documentation, knowledge bases
  • Strategy: Place cacheable content at the start of your prompt
Example savings:
  • 100K context without cache: $260/1M requests
  • With cache: $26/1M requests = 90% savings

Context Windows

Gemini models support large context:
  • Gemini 3 Pro: Up to 2M tokens
  • Gemini 2.5 Pro: Up to 2M tokens
  • Gemini 2.5 Flash: Up to 1M tokens

Support

Need help with Gemini integration?

Google AI Documentation

Official Gemini documentation

Google AI Pricing

Official pricing details

Splox Docs

Browse our guides

Community

Get help from the community