Laboratory
Entity that develops or publishes the model and maintains its technical and safety documentation.
Explore the documented models, published evaluations and primary sources associated with this organisation.
Flash model for long-running software engineering, autonomous agents, and complex business flows.
16 SEP 2026↗ MULTIMODAL REASONINGMultimodal reasoning, coding, and complex problems.
16 SEP 2026↗ PREVIOUS FLASH · AGENTSPrevious Flash generation for complex code, agents, and multi-stage execution; useful as a comparison point with Gemini 3.8.
18 SEP 2026↗ EFFICIENT NATIVE IMAGE MODELProduction image model that combines generation, editing, and visual reasoning with a balance of quality, cost, and latency.
18 SEP 2026↗ CINEMATIC VIDEO WITH AUDIOCinematic video model with native audio, scene extension, and control through frames and reference images.
18 SEP 2026↗ MUSIC GENERATIONMusic model for full songs with vocals, lyrics, and control over structure, style, duration, and instrumentation.
18 SEP 2026↗ UNIFIED MULTIMODAL EMBEDDINGEmbedding model that places text, images, video, audio, and PDFs in a shared vector space for search and multimodal RAG.
18 SEP 2026↗ BUILT-IN SPATIAL REASONINGModel specialized in physical understanding, video, spatial reasoning, and tool orchestration for robotic systems.
18 SEP 2026↗ REAL-TIME MULTIMODAL CONVERSATIONLive conversation model integrating voice, video, and tools with low latency.
29 SEP 2026↗ LIVE VOICE WITH EXTENDED REASONINGLive voice variant with background reasoning for complex multi-step tasks.
29 SEP 2026↗Figures are shown with the context reported by their source. A provider result is not an independent comparison and does not replace your own evaluation.
| Model | Benchmark | Result | Metric |
|---|---|---|---|
| Gemini 3.8 Flash | Artificial Analysis Intelligence v4.1 | 50,2 | Index |
| Gemini 3.8 Flash | GDPval-AA v2 | 1.348,8 | Elo |
| Gemini 3.1 Pro | SWE-Bench Pro | 54,2 % | Resolved |
| Gemini 3.1 Pro | Terminal-Bench 2.1 | 70,7 % | Accuracy |
| Gemini 3.7 Flash | Evaluaciones Gemini 3.7 Flash | Published | Coding and agents |
| Nano Banana 2 | Evaluación visual Nano Banana 2 | Recommended model | Quality, cost, and latency |
| Veo 3.1 | MovieGenBench | Published leadership | Global text-to-video preference |
| Lyria 3.5 | Evaluación de Lyria 3.5 | Published | Musicality, vocals, and structural coherence |
| Gemini Embedding 2 | Evaluaciones Gemini Embedding 2 | Published | Multimodal retrieval |
| Gemini Robotics ER 2 | Evaluaciones Robotics ER 2 | Published | Spatial and robotic reasoning |
An organization, a product, and a model are not the same unit. Inferama separates them to avoid attributing capabilities or commercial terms to the wrong item.
Entity that develops or publishes the model and maintains its technical and safety documentation.
Identifiable version with limits, modalities, and behavior that may change between releases.
API, application, associated cloud, or commercial plan; each channel may have different pricing, retention, and limits.
Every claim must retain the official page consulted and the verification date.
Gemini Robotics ER 2 is presented as a vision-language model capable of planning multi-step tasks for robotics. That does not show that it directly controls a robot or that its plans are safe or reliable in every environment. We examine what the available documentation supports and what each team must test.
30 Sep 2026 ↗ ANALISISGoogle’s documentation describes a model that represents text and other types of content in a shared embedding space. That supports a stated capability, not a conclusion that it improves retrieval on any particular corpus. We review what the available sources can establish about performance, access, pricing, and data, and what teams should verify before testing it.
30 Sep 2026 ↗ ANALISISGoogle positions Gemini 3.8 Flash for long-horizon software engineering and autonomous agents. That stated focus does not, by itself, show that the model can reliably complete extended tasks. This review separates the provider’s claims from documented access and Agent Platform pricing conditions, and identifies evidence that is missing from the sources reviewed.
30 Sep 2026 ↗ ANALISISGoogle describes Gemini 3.1 Pro as a model for complex tasks, but that description does not prove that a specific integration can reliably complete a multi-step workflow. We propose a reversible evaluation protocol and explain what the available sources can—and cannot—establish about access, pricing, safety, and performance.
29 Sep 2026 ↗ ANALISISGoogle’s documentation presents Gemini 3.7 Flash as a model for multi-step tasks, code refactoring, and reasoning. Here is what the available sources can—and cannot—verify about access, pricing, safety, and performance.
29 Sep 2026 ↗ COMPARATIVAThere are no results from a controlled comparison here that would support declaring a winner. What we can offer is a reproducible protocol for measuring patch quality, regressions, cost, and time under shared conditions—and for deciding what evidence a team needs before choosing.
28 Sep 2026 ↗