Modelos de IA de Google

Explora 29 modelos de IA de Google y compara contexto, precios, modalidades y capacidades.

29modelos registrados

Directorio de modelos

29 modelos en esta vista

Comparar
Activo

A fast multimodal Gemini model listed for responsive agent workflows, coding and multi-step reasoning.

Contexto
1,05M
Entrada
0,375 US$
Salida
1,88 US$
TextoImagenVídeoArchivo
Ver modelo

Google

Chirp 3

Activo

Chirp 3 is Google's latest multilingual speech-to-text model. It offers enhanced transcription accuracy across 24 GA languages and 77+ preview languages, with support for automatic language detection, automatic punctuation, and…

Contexto
0
Entrada
16.000 US$
Salida
0,00 US$
Audio
Ver modelo

Gemini 3 Flash Preview is a high speed, high value thinking model designed for agentic workflows, multi turn chat, and coding assistance. It delivers near Pro level reasoning and tool…

Contexto
1,05M
Entrada
0,50 US$
Salida
3 US$
TextoImagenArchivoAudio
Ver modelo
Activo

Gemini 3.1 Flash Lite is Google’s GA high-efficiency multimodal model optimized for low-latency, high-volume workloads. It supports text, image, video, audio, and PDF inputs, and is designed for lightweight agentic…

Contexto
1,05M
Entrada
0,25 US$
Salida
1,5 US$
TextoImagenVídeoArchivo
Ver modelo

Gemini 3.1 Flash Lite Preview is Google's high-efficiency model optimized for high-volume use cases. It outperforms Gemini 2.5 Flash Lite on overall quality and approaches Gemini 2.5 Flash performance across…

Contexto
1,05M
Entrada
0,25 US$
Salida
1,5 US$
TextoImagenVídeoArchivo
Ver modelo

Gemini 3.1 Flash TTS Preview is a text-to-speech model from Google, and a substantial generational step up from Gemini 2.5 Flash TTS. It takes text input and produces audio output…

Contexto
33K
Entrada
1 US$
Salida
20 US$
Texto
Ver modelo

Gemini 3.1 Pro Preview is Google’s frontier reasoning model, delivering enhanced software engineering performance, improved agentic reliability, and more efficient token usage across complex workflows. Building on the multimodal foundation…

Contexto
1,05M
Entrada
2 US$
Salida
12 US$
AudioArchivoImagenTexto
Ver modelo

Gemini 3.1 Pro Preview Custom Tools is a variant of Gemini 3.1 Pro that improves tool selection behavior by preventing overuse of a general bash tool when more efficient third-party…

Contexto
1,05M
Entrada
2 US$
Salida
12 US$
TextoAudioImagenVídeo
Ver modelo
Activo

Gemini 3.5 Flash is Google's high-efficiency multimodal model, bringing near-Pro level coding and reasoning at Flash-tier cost and speed. It is highly optimized for coding proficiency and parallel agentic execution…

Contexto
1,05M
Entrada
1,5 US$
Salida
9 US$
TextoImagenVídeoArchivo
Ver modelo
Activo

Gemini 3.5 Flash Lite is a high-efficiency model from Google with upgraded agentic capabilities. It is suited for subagents that execute focused tasks within complex, multi-agent workflows.

Contexto
1,05M
Entrada
0,30 US$
Salida
2,5 US$
TextoImagenVídeoArchivo
Ver modelo
Activo

Gemini 3.6 Flash is a high-efficiency model from Google for coding, agentic workflows, and web and app development. It is designed to produce polished outputs with fewer unnecessary edits and…

Contexto
1,05M
Entrada
0,75 US$
Salida
3,75 US$
TextoImagenVídeoArchivo
Ver modelo
Activo

gemini-embedding-001 provides a unified cutting edge experience across domains, including science, legal, finance, and coding. This embedding model has consistently held a top spot on the Massive Text Embedding Benchmark…

Contexto
20K
Entrada
0,15 US$
Salida
0,00 US$
Texto
Ver modelo
Activo

Gemini Embedding 2 is Google's first multimodal embedding model. We currently support mapping text and images into a unified vector space for semantic search and retrieval-augmented generation (RAG). It supports…

Contexto
8K
Entrada
0,20 US$
Salida
0,00 US$
TextoImagenArchivoAudio
Ver modelo

Gemini Embedding 2 Preview is Google's first multimodal embedding model. We currently support mapping text and images into a unified vector space for semantic search and retrieval-augmented generation (RAG). It…

Contexto
8K
Entrada
0,20 US$
Salida
0,00 US$
TextoImagenArchivoAudio
Ver modelo
Activo

Gemma is a family of open models built by Google DeepMind. Gemma 4 models are multimodal, handling text and image input (with audio supported on E2B, E4B, and 12B) and generating text output. This release includes open-weights models in both pre-trained and instruction-tuned variants. Gemma 4 features a context window of up to 256K tokens and maintains multilingual support in over 140 languages.

Contexto
262K
Entrada
0,07 US$
Salida
0,34 US$
ImagenTextoVídeo
Ver modelo
Activo

Gemma is a family of open models built by Google DeepMind. Gemma 4 models are multimodal, handling text and image input (with audio supported on E2B, E4B, and 12B) and generating text output. This release includes open-weights models in both pre-trained and instruction-tuned variants. Gemma 4 features a context window of up to 256K tokens and maintains multilingual support in over 140 languages.

Contexto
262K
Entrada
0,10 US$
Salida
0,34 US$
ImagenTextoVídeo
Ver modelo

This model always redirects to the latest model in the Google Gemini Flash family.

Contexto
1,05M
Entrada
0,375 US$
Salida
1,88 US$
TextoImagenVídeoArchivo
Ver modelo

This model always redirects to the latest model in the Google Gemini Pro family.

Contexto
1,05M
Entrada
2 US$
Salida
12 US$
AudioArchivoImagenTexto
Ver modelo
Activo

30 second duration clips are priced at $0.04 per clip. Lyria 3 is Google's family of music generation models, available through the Gemini API. With Lyria 3, you can generate…

Contexto
1,05M
Entrada
0,00 US$
Salida
0,00 US$
TextoImagen
Ver modelo
Activo

Full-length songs are priced at $0.08 per song. Lyria 3 is Google's family of music generation models, available through the Gemini API. With Lyria 3, you can generate high-quality, 48kHz…

Contexto
1,05M
Entrada
0,00 US$
Salida
0,00 US$
TextoImagen
Ver modelo

Gemini 2.5 Flash Image, a.k.a. "Nano Banana," is now generally available. It is a state of the art image generation model with contextual understanding. It is capable of image generation,…

Contexto
33K
Entrada
0,30 US$
Salida
2,5 US$
ImagenTexto
Ver modelo

Gemini 3.1 Flash Image Preview, a.k.a. "Nano Banana 2," is Google’s latest state of the art image generation and editing model, delivering Pro-level visual quality at Flash speed. It combines…

Contexto
66K
Entrada
0,50 US$
Salida
3 US$
ImagenTexto
Ver modelo

Nano Banana 2 Lite (Gemini 3.1 Flash Lite Image) is Google's fastest, most cost-efficient Gemini image model, built for high-velocity developer pipelines and rapid-fire visual exploration. It delivers text-to-image generation…

Contexto
66K
Entrada
0,25 US$
Salida
1,5 US$
ImagenTexto
Ver modelo

Nano Banana Pro is Google’s most advanced image-generation and editing model, built on Gemini 3 Pro. It extends the original Nano Banana with significantly improved multimodal reasoning, real-world grounding, and…

Contexto
66K
Entrada
2 US$
Salida
12 US$
ImagenTexto
Ver modelo

Nano Banana Pro is Google’s most advanced image-generation and editing model, built on Gemini 3 Pro. It extends the original Nano Banana with significantly improved multimodal reasoning, real-world grounding, and…

Contexto
66K
Entrada
2 US$
Salida
12 US$
ImagenTexto
Ver modelo

Google

Veo 3.1

Activo

Google's state-of-the-art video generation model, built for maximum visual fidelity in final production cuts. Veo 3.1 generates high-quality 1080p video from text or image prompts with native synchronized audio —…

Contexto
0
Entrada
0,00 US$
Salida
0,00 US$
TextoImagen
Ver modelo
Activo

Google's mid-tier video generation model balancing speed and quality. Veo 3.1 Fast generates high-quality video from text or image prompts with native synchronized audio, offering faster turnaround than Veo 3.1…

Contexto
0
Entrada
0,00 US$
Salida
0,00 US$
TextoImagen
Ver modelo
Activo

Google's most cost-effective video generation model, designed for high-volume applications and rapid iteration. Veo 3.1 Lite generates 720p and 1080p video from text or image prompts with native synchronized audio…

Contexto
0
Entrada
0,00 US$
Salida
0,00 US$
TextoImagen
Ver modelo
OUR METHOD

Cómo gestiona AIToolly los datos de modelos

Los datos propios del modelo y los del endpoint del proveedor se mantienen separados. Los catálogos de terceros se usan para descubrimiento y capturas; la documentación oficial y las fichas verificadas tienen prioridad para los datos del modelo.

01Fuentes y verificación
02Verificado
03Capacidad