oxyy.ai
HomeModelsProvidersPricingBlogDocs
Sign inGet API key
  1. Home/
  2. Providers/
  3. Google

Google models

Access 25 Google models through the Oxyy unified API including Gemini 3 Flash, Gemini 3.8 Flash, Gemini 3.1 Flash-Lite. Compare pricing, context windows and capabilities between Google models.

Google tokens processed on Oxyy· daily, UTC

Models 25

Google: Gemini 3 FlashTextImageVideoAudioFile→Text-60%3.9B tokens

Legacy first-generation Gemini 3 Flash model providing baseline speed and intelligence with frontier-class multimodal understanding.

by GoogleJan 1, 20261.05M context$0.500/M input$3.00/M output12.5s latency885 t/s
Google: Gemini 3.8 FlashTextImageVideoAudioFile→Text-60%320.9M tokens

Google's most intelligent Flash model, engineered for long-horizon software engineering, autonomous agents, and complex enterprise workflows.

by GoogleSep 2, 20261.05M context$0.750/M input$3.75/M output7019ms latency734 t/s
Google: Gemini 3.1 Flash-LiteTextImageVideoAudioFile→Text-60%226.4M tokens

Cost-efficient multimodal model for high-volume agentic tasks, translation, and simple data extraction where budget and latency are the primary constraints.

by Google1.05M context$0.250/M input$1.50/M output1772ms latency264 t/s
Google: Gemini 2.5 Flash-LiteTextImageVideoAudioFile→Text-60%177.6M tokens

Smallest and most cost-effective multimodal model in the 2.5 family, built for at-scale usage.

by GoogleJul 22, 20251.05M context$0.100/M input$0.400/M output2068ms latency218 t/s
Google: Nano Banana 2TextImageVideo→TextImage-60%176.7M tokens

High-efficiency production-scale image generation and editing, balancing speed with 4K generation, world knowledge and reliable text rendering. The generalist workhorse of the Nano Banana family.

by GoogleFeb 26, 2026— context$0.500/M input$3.00/M output9949ms latency
Google: Gemini 3.7 FlashTextImageVideoAudioFile→Text-60%151.2M tokens

High-speed, efficient Flash model built for everyday coding, agentic tool use, and reliable multi-step execution.

by Google1.05M context$0.750/M input$3.75/M output11.6s latency515 t/s
Google: Nano Banana 2 Lite (Gemini 3.1 Flash Lite Image)ImageText→ImageText-60%127.0M tokens

Nano Banana 2 Lite (Gemini 3.1 Flash Lite Image) is Google's fastest, most cost-efficient Gemini image model, built for high-velocity developer pipelines and rapid-fire visual exploration. It delivers text-to-image gene…

by GoogleJun 30, 202666K context$0.250/M input$1.50/M output4557ms latency
Google: Gemini 3.1 ProTextImageVideoAudioFile→Text-60%124.1M tokens

Third-generation Pro model built for multimodal understanding, agentic capability, and vibe-coding; improved thinking, token efficiency, and factual grounding over Gemini 3 Pro.

by GoogleFeb 1, 20261.05M context$2.00/M input$12.00/M output13.8s latency257 t/s
Google: Gemini 2.5 FlashTextImageVideoAudioFile→Text-60%121.3M tokens

Google's first hybrid reasoning model with configurable thinking budgets; best price-performance for low-latency, high-volume tasks that require reasoning.

by GoogleJun 17, 20251.05M context$0.300/M input$2.50/M output2097ms latency230 t/s
Google: Gemini 3.5 Flash-LiteTextImageVideoAudioFile→Text-60%59.1M tokens

Fastest, most cost-effective model in the 3.5 family, optimized for high-throughput agentic tasks, translation, and simple data processing.

by Google1.05M context$0.300/M input$2.50/M output2355ms latency268 t/s
Google: Gemma 4 26B A4B Instruct (MoE)TextImageVideo→Text-60%32.8M tokens
open-weightsmoereasoningcodingfast+1 more

Gemma 4's Mixture-of-Experts model. 25.2B total parameters but only 3.8B active per token, so it runs almost as fast as a 4B model while scoring close to the dense 31B.

by GoogleApr 2, 2026262K context$0.130/M input$0.400/M output8572ms latency60 t/s
Google: Gemini 3.6 FlashTextImageVideoAudioFile→Text-60%18.0M tokens

Previous-generation Flash model balancing speed and multimodal capability across general agentic and everyday tasks; strong at code generation, agentic execution, and spatial reasoning.

by Google1.05M context$0.750/M input$3.75/M output3269ms latency742 t/s
Google: Gemini 3.5 FlashTextImageVideoAudioFile→Text-60%14.1M tokens

Legacy Flash model providing sustained frontier-level intelligence for real-world tasks; effective for sub-agent deployment, multi-step workflows, and long-horizon tasks at scale.

by Google1.05M context$1.50/M input$9.00/M output12.3s latency317 t/s
Google: Gemma 4 31B InstructTextImageVideo→Text-60%2.2M tokens
open-weightsreasoningcodingagenticmultimodal+1 more

The largest Gemma 4 model: a 30.7B-parameter dense multimodal model for reasoning, agentic workflows, coding and multimodal understanding, deployable on consumer GPUs and workstations.

by GoogleApr 2, 2026262K context$0.270/M input$0.760/M output48.5s latency36 t/s
Google: Gemini 2.5 ProTextImageVideoAudioFile→Text-60%1.5M tokens

Most advanced model of the 2.5 family, with deep reasoning and coding capability for complex tasks.

by GoogleJun 17, 20251.05M context$1.25/M input$10.00/M output30.2s latency140 t/s
Google: Gemini 3.1 Flash TTSText→Audio-60%16.5K tokens

Powerful, low-latency speech generation with natural outputs, steerable prompts and expressive inline audio tags for precise narration control across 70+ languages.

by GoogleApr 15, 202633K context$1.00/M input—/M output5434ms latency
Google: Gemini 2.5 Flash TTSText→Audio-60%45 tokens

Fast and controllable text-to-speech for low-latency, cost-efficient applications and real-time assistants, with fine control over style and pacing.

by GoogleMay 20, 20258K context$0.500/M input—/M output4241ms latency
Google: Gemini 3.5 TranscribeAudioText→Text-60%29d ago

High-accuracy, low-latency non-streaming speech-to-text with utterance-based language detection across 85+ languages, speaker diarization, word-level timestamps and custom vocabulary biasing.

by GoogleAug 26, 202696K context$2.00/M input$12.00/M output
Google: Gemini 3.5 Transcribe LiveAudio→Text-60%29d ago

Low-latency bidirectional streaming speech-to-text over WebSockets using the Live API, with interim and finalized transcription events, Smart transcription mode and multiple voice-activity-detection strategies.

by GoogleAug 26, 202696K context$3.50/M input$21.00/M output
Google: Gemma 4 12B UnifiedTextImageAudioVideo→Text-60%5mo ago
open-weightsmultimodalaudioencoder-freelocal-inference

Encoder-free multimodal Gemma 4 model. Instead of separate vision and audio encoders, it projects raw image patches and audio waveforms straight into the LLM's embedding space through lightweight linear layers, so every…

by GoogleApr 2, 2026262K context$0.100/M input$0.300/M output
Google: Gemma 4 E2BTextImageAudioVideo→Text-60%5mo ago
open-weightson-devicemobileaudioedge+1 more

The smallest Gemma 4 model, built for efficient on-device execution on phones and laptops, with native audio input.

by GoogleApr 2, 2026131K context$0.040/M input$0.040/M output
Google: Gemma 4 E4BTextImageAudioVideo→Text-60%5mo ago
open-weightson-devicemobileaudioedge

On-device Gemma 4 model for laptops and mobile devices, with native audio input. 'E' stands for effective parameters.

by GoogleApr 2, 2026131K context$0.200/M input$0.200/M output
Google: Gemini 2.5 Pro TTSText→Audio-60%1y ago

High-fidelity speech synthesis optimized for quality in structured workflows such as podcasts and audiobooks, with more natural outputs and easier-to-steer prompts.

by GoogleMay 20, 20258K context$1.00/M input—/M output
Google: Gemini Omni FlashTextImageVideoAudio→TextVideoAudio-60%—

Fast conversational video generation and editing with native audio, keyframe interpolation and clip extension. Turn text and images into video and refine results through natural language.

by Google— context$1.50/M input$9.00/M output
Google: Gemini Embedding 2TextImageAudioVideoFile→Embeddings-60%—
embeddingmultimodalragsemantic-search

Google's first multimodal embedding model, mapping text, images, video, audio, and PDFs into a unified embedding space for semantic search and RAG.

by Google— context$0.200/M input—/M output
oxyy.ai

One endpoint for every major model. The OpenAI, Anthropic and Gemini SDKs work unchanged.

Product

  • Models
  • Providers
  • Pricing
  • Status
  • Startup program

Company

  • About
  • Blog
  • Contact
  • Terms of Service
  • Privacy Policy
  • Refund Policy
  • Cookie Policy

Developer

  • Documentation
  • Quickstart
  • SDKs
  • Model catalog
  • AI providers

Connect

  • Telegram
  • Discord
  • WhatsApp
© 2026 Oxyy.ai. All rights reserved.StatusAbout