Google Gemini provides multimodal AI capabilities:
- Advanced reasoning with Gemini 2.5 and 3.0
- Multimodal understanding (text, image, audio)
- Prompt caching for cost optimization
- Cost-effective options with Flash variants
Available Models
Gemini 3 Pro
Next-generation model with enhanced capabilities and prompt caching.Gemini 3 Pro supports prompt caching - reuse context at 90% discount.
Gemini 2.5 Pro
Most capable current-gen models for complex tasks.Gemini 2.5 Flash
Fast, cost-efficient models for high-volume tasks.Best Practices
Model Selection
Use Gemini 3 for
- Cutting-edge features
- Latest capabilities
- Prompt caching needs
- Advanced reasoning
Use 2.5 Pro for
- Complex reasoning tasks
- Production applications
- High-quality outputs
- When accuracy matters
Use Flash for
- High-volume tasks
- Fast responses needed
- Cost-sensitive workloads
- Simple queries
Prompt Caching (Gemini 3)
Optimize costs with prompt caching on Gemini 3 Pro:
- Cache Read: $0.26/M (90% cheaper than input)
- Use case: Repeated system prompts, documentation, knowledge bases
- Strategy: Place cacheable content at the start of your prompt
- 100K context without cache: $260/1M requests
- With cache: $26/1M requests = 90% savings
Context Windows
Gemini models support large context:- Gemini 3 Pro: Up to 2M tokens
- Gemini 2.5 Pro: Up to 2M tokens
- Gemini 2.5 Flash: Up to 1M tokens
Support
Need help with Gemini integration?Google AI Documentation
Official Gemini documentation
Google AI Pricing
Official pricing details
Splox Docs
Browse our guides
Community
Get help from the community

