Modèles d’IA Google

Explorez 29 modèles d’IA de Google et comparez contexte, tarifs, modalités et capacités.

29modèles suivis

Répertoire des modèles

29 modèles dans cette vue

Comparer
Actif

A fast multimodal Gemini model listed for responsive agent workflows, coding and multi-step reasoning.

Contexte
1,05M
Entrée
0,375 $US
Sortie
1,88 $US
TexteImageVidéoFichier
Voir le modèle

Google

Chirp 3

Actif

Chirp 3 is Google's latest multilingual speech-to-text model. It offers enhanced transcription accuracy across 24 GA languages and 77+ preview languages, with support for automatic language detection, automatic punctuation, and…

Contexte
0
Entrée
16 000 $US
Sortie
0,00 $US
Audio
Voir le modèle

Gemini 3 Flash Preview is a high speed, high value thinking model designed for agentic workflows, multi turn chat, and coding assistance. It delivers near Pro level reasoning and tool…

Contexte
1,05M
Entrée
0,50 $US
Sortie
3 $US
TexteImageFichierAudio
Voir le modèle

Gemini 3.1 Flash Lite is Google’s GA high-efficiency multimodal model optimized for low-latency, high-volume workloads. It supports text, image, video, audio, and PDF inputs, and is designed for lightweight agentic…

Contexte
1,05M
Entrée
0,25 $US
Sortie
1,5 $US
TexteImageVidéoFichier
Voir le modèle

Gemini 3.1 Flash Lite Preview is Google's high-efficiency model optimized for high-volume use cases. It outperforms Gemini 2.5 Flash Lite on overall quality and approaches Gemini 2.5 Flash performance across…

Contexte
1,05M
Entrée
0,25 $US
Sortie
1,5 $US
TexteImageVidéoFichier
Voir le modèle

Gemini 3.1 Flash TTS Preview is a text-to-speech model from Google, and a substantial generational step up from Gemini 2.5 Flash TTS. It takes text input and produces audio output…

Contexte
33K
Entrée
1 $US
Sortie
20 $US
Texte
Voir le modèle

Gemini 3.1 Pro Preview is Google’s frontier reasoning model, delivering enhanced software engineering performance, improved agentic reliability, and more efficient token usage across complex workflows. Building on the multimodal foundation…

Contexte
1,05M
Entrée
2 $US
Sortie
12 $US
AudioFichierImageTexte
Voir le modèle

Gemini 3.1 Pro Preview Custom Tools is a variant of Gemini 3.1 Pro that improves tool selection behavior by preventing overuse of a general bash tool when more efficient third-party…

Contexte
1,05M
Entrée
2 $US
Sortie
12 $US
TexteAudioImageVidéo
Voir le modèle
Actif

Gemini 3.5 Flash is Google's high-efficiency multimodal model, bringing near-Pro level coding and reasoning at Flash-tier cost and speed. It is highly optimized for coding proficiency and parallel agentic execution…

Contexte
1,05M
Entrée
1,5 $US
Sortie
9 $US
TexteImageVidéoFichier
Voir le modèle

Gemini 3.5 Flash Lite is a high-efficiency model from Google with upgraded agentic capabilities. It is suited for subagents that execute focused tasks within complex, multi-agent workflows.

Contexte
1,05M
Entrée
0,30 $US
Sortie
2,5 $US
TexteImageVidéoFichier
Voir le modèle
Actif

Gemini 3.6 Flash is a high-efficiency model from Google for coding, agentic workflows, and web and app development. It is designed to produce polished outputs with fewer unnecessary edits and…

Contexte
1,05M
Entrée
0,75 $US
Sortie
3,75 $US
TexteImageVidéoFichier
Voir le modèle

gemini-embedding-001 provides a unified cutting edge experience across domains, including science, legal, finance, and coding. This embedding model has consistently held a top spot on the Massive Text Embedding Benchmark…

Contexte
20K
Entrée
0,15 $US
Sortie
0,00 $US
Texte
Voir le modèle
Actif

Gemini Embedding 2 is Google's first multimodal embedding model. We currently support mapping text and images into a unified vector space for semantic search and retrieval-augmented generation (RAG). It supports…

Contexte
8K
Entrée
0,20 $US
Sortie
0,00 $US
TexteImageFichierAudio
Voir le modèle

Gemini Embedding 2 Preview is Google's first multimodal embedding model. We currently support mapping text and images into a unified vector space for semantic search and retrieval-augmented generation (RAG). It…

Contexte
8K
Entrée
0,20 $US
Sortie
0,00 $US
TexteImageFichierAudio
Voir le modèle
Actif

Gemma is a family of open models built by Google DeepMind. Gemma 4 models are multimodal, handling text and image input (with audio supported on E2B, E4B, and 12B) and generating text output. This release includes open-weights models in both pre-trained and instruction-tuned variants. Gemma 4 features a context window of up to 256K tokens and maintains multilingual support in over 140 languages.

Contexte
262K
Entrée
0,07 $US
Sortie
0,34 $US
ImageTexteVidéo
Voir le modèle
Actif

Gemma is a family of open models built by Google DeepMind. Gemma 4 models are multimodal, handling text and image input (with audio supported on E2B, E4B, and 12B) and generating text output. This release includes open-weights models in both pre-trained and instruction-tuned variants. Gemma 4 features a context window of up to 256K tokens and maintains multilingual support in over 140 languages.

Contexte
262K
Entrée
0,10 $US
Sortie
0,34 $US
ImageTexteVidéo
Voir le modèle

This model always redirects to the latest model in the Google Gemini Flash family.

Contexte
1,05M
Entrée
0,375 $US
Sortie
1,88 $US
TexteImageVidéoFichier
Voir le modèle

This model always redirects to the latest model in the Google Gemini Pro family.

Contexte
1,05M
Entrée
2 $US
Sortie
12 $US
AudioFichierImageTexte
Voir le modèle

30 second duration clips are priced at $0.04 per clip. Lyria 3 is Google's family of music generation models, available through the Gemini API. With Lyria 3, you can generate…

Contexte
1,05M
Entrée
0,00 $US
Sortie
0,00 $US
TexteImage
Voir le modèle
Actif

Full-length songs are priced at $0.08 per song. Lyria 3 is Google's family of music generation models, available through the Gemini API. With Lyria 3, you can generate high-quality, 48kHz…

Contexte
1,05M
Entrée
0,00 $US
Sortie
0,00 $US
TexteImage
Voir le modèle

Gemini 2.5 Flash Image, a.k.a. "Nano Banana," is now generally available. It is a state of the art image generation model with contextual understanding. It is capable of image generation,…

Contexte
33K
Entrée
0,30 $US
Sortie
2,5 $US
ImageTexte
Voir le modèle

Gemini 3.1 Flash Image Preview, a.k.a. "Nano Banana 2," is Google’s latest state of the art image generation and editing model, delivering Pro-level visual quality at Flash speed. It combines…

Contexte
66K
Entrée
0,50 $US
Sortie
3 $US
ImageTexte
Voir le modèle

Nano Banana 2 Lite (Gemini 3.1 Flash Lite Image) is Google's fastest, most cost-efficient Gemini image model, built for high-velocity developer pipelines and rapid-fire visual exploration. It delivers text-to-image generation…

Contexte
66K
Entrée
0,25 $US
Sortie
1,5 $US
ImageTexte
Voir le modèle

Nano Banana Pro is Google’s most advanced image-generation and editing model, built on Gemini 3 Pro. It extends the original Nano Banana with significantly improved multimodal reasoning, real-world grounding, and…

Contexte
66K
Entrée
2 $US
Sortie
12 $US
ImageTexte
Voir le modèle

Nano Banana Pro is Google’s most advanced image-generation and editing model, built on Gemini 3 Pro. It extends the original Nano Banana with significantly improved multimodal reasoning, real-world grounding, and…

Contexte
66K
Entrée
2 $US
Sortie
12 $US
ImageTexte
Voir le modèle

Google

Veo 3.1

Actif

Google's state-of-the-art video generation model, built for maximum visual fidelity in final production cuts. Veo 3.1 generates high-quality 1080p video from text or image prompts with native synchronized audio —…

Contexte
0
Entrée
0,00 $US
Sortie
0,00 $US
TexteImage
Voir le modèle
Actif

Google's mid-tier video generation model balancing speed and quality. Veo 3.1 Fast generates high-quality video from text or image prompts with native synchronized audio, offering faster turnaround than Veo 3.1…

Contexte
0
Entrée
0,00 $US
Sortie
0,00 $US
TexteImage
Voir le modèle
Actif

Google's most cost-effective video generation model, designed for high-volume applications and rapid iteration. Veo 3.1 Lite generates 720p and 1080p video from text or image prompts with native synchronized audio…

Contexte
0
Entrée
0,00 $US
Sortie
0,00 $US
TexteImage
Voir le modèle
OUR METHOD

Comment AIToolly traite les données des modèles

Les faits propres au modèle et ceux de l’endpoint fournisseur restent séparés. Les catalogues tiers servent à la découverte et aux instantanés ; la documentation officielle et les fiches vérifiées restent prioritaires pour les faits du modèle.

01Sources et vérification
02Vérifié
03Capacité