Gemini
Google DeepMind's multimodal LLM family. Long context windows, vision understanding, and agentic tool use across coding and analysis tasks.
The picks.Gemini models by Google
Gemini 3.7 Flash
Google’s latest Flash workhorse for coding, agents, multimodal understanding, and document-heavy workflows
Gemini 3.5 Flash-Lite
Fast high-throughput Gemini model for low-latency agentic and document-heavy workloads
Gemini 3.6 Flash
More efficient multimodal Flash model for coding, knowledge work, and agentic execution
Gemini Omni Flash
Multimodal video generation and editing with native audio, multi-turn control, and photo-to-video references
Gemini 3.5 Flash
Frontier multimodal reasoning model for agentic and coding workflows
Gemini 3.1 Flash TTS
Expressive text-to-speech with audio tags, multi-speaker dialogue, and 70+ languages
Gemini 3.5 Flash Cyber
Cybersecurity-specialized Gemini model for finding, validating, and patching software vulnerabilities
Common questions
Gemini 3.5 Flash-Lite, released July 2026 per the live catalog. New Gemini releases join this page automatically when they land on Runware.
Gemini 3.1 Flash Lite, from $0.25 / 1M per Input tokens (text, image, video). Prices are read live from the catalog and vary with resolution and configuration — expand any row above to see that model's current pricing.
Production notes
Every price on this page is the model's published rate from the live Runware catalog, using the cheapest listed configuration unless stated otherwise. Prices vary with resolution, duration, quality tier, or token volume, so check the pricing table on each model page before estimating unit economics.
The Runware catalog does not publish per-model latency figures, so this page does not quote end-to-end timings. Where a model's own description commits to speed (for example sub-second generation or realtime streaming), that claim is repeated here. For anything else, benchmark the exact models in the Playground with your own payload sizes before committing to an SLA.
Models are addressed by versioned AIR identifiers, so a workflow pinned to specific model versions keeps producing the same behaviour as new versions ship. Adopt upgrades deliberately by re-running your evaluation set against the new version before switching production traffic.
About this collection
Google's multimodal LLM family. Long context windows, vision understanding, and agentic tool use across coding and analysis tasks. This page lists every Gemini model available through the Runware API, published by Google.
Membership is derived from the live Runware catalog: the 10 models on this page are the catalog's current Gemini releases, ordered by the catalog's own editorial weight and release date. Names, descriptions, pricing, capability chips, samples, and guides are all read live from the catalog — nothing here is hand-curated, and new Gemini releases join automatically when they land on Runware.
Everything on this page re-derives from the catalog on every visit. When Google ships a new Gemini model on Runware it appears here with no editorial lag; deprecated versions remain listed with their live status chip so version pins stay auditable.
Model guides.Learn how to use the stack
Cinematic prompting for Gemini Omni Flash
How to prompt Gemini Omni Flash for cinematic video using Google's five-element structure, camera language, and the less-prescriptive sweet spot.
Read the guide →Editing video with Gemini Omni Flash
How to edit existing footage with Gemini Omni Flash's inputs.video parameter to relight, restyle, swap weather, or add characters while preserving the source's composition and motion.
Read the guide →Reference-driven video with Gemini Omni Flash
How to use Gemini Omni Flash's reference image workflow to lock a visual style, hold a character across scenes, or guide a video through storyboard key beats.
Read the guide →Related collections.More ways to build
Browse all collections→GPT-5
OpenAI's frontier LLM family. Coding, reasoning, structured tool use, and agentic workflows with streaming output by default.
Open →VIDMiniMax · Hailuo
MiniMax video and reasoning families. Hailuo delivers cinematic video; the M-series ships streaming LLMs for coding and agentic workflows.
Open →