Modele AI Google

Przeglądaj 29 modeli AI od Google i porównuj kontekst, ceny, modalności i możliwości.

29śledzonych modeli

Katalog modeli

29 modeli w tym widoku

Porównaj
Aktywny

A fast multimodal Gemini model listed for responsive agent workflows, coding and multi-step reasoning.

Kontekst
1,05M
Wejście
0,375 USD
Wyjście
1,88 USD
TekstObrazWideoPlik
Zobacz model

Google

Chirp 3

Aktywny

Chirp 3 is Google's latest multilingual speech-to-text model. It offers enhanced transcription accuracy across 24 GA languages and 77+ preview languages, with support for automatic language detection, automatic punctuation, and…

Kontekst
0
Wejście
16 000 USD
Wyjście
0,00 USD
Dźwięk
Zobacz model
Aktywny

Gemini 3 Flash Preview is a high speed, high value thinking model designed for agentic workflows, multi turn chat, and coding assistance. It delivers near Pro level reasoning and tool…

Kontekst
1,05M
Wejście
0,50 USD
Wyjście
3 USD
TekstObrazPlikDźwięk
Zobacz model
Aktywny

Gemini 3.1 Flash Lite is Google’s GA high-efficiency multimodal model optimized for low-latency, high-volume workloads. It supports text, image, video, audio, and PDF inputs, and is designed for lightweight agentic…

Kontekst
1,05M
Wejście
0,25 USD
Wyjście
1,5 USD
TekstObrazWideoPlik
Zobacz model

Gemini 3.1 Flash Lite Preview is Google's high-efficiency model optimized for high-volume use cases. It outperforms Gemini 2.5 Flash Lite on overall quality and approaches Gemini 2.5 Flash performance across…

Kontekst
1,05M
Wejście
0,25 USD
Wyjście
1,5 USD
TekstObrazWideoPlik
Zobacz model

Gemini 3.1 Flash TTS Preview is a text-to-speech model from Google, and a substantial generational step up from Gemini 2.5 Flash TTS. It takes text input and produces audio output…

Kontekst
33K
Wejście
1 USD
Wyjście
20 USD
Tekst
Zobacz model
Aktywny

Gemini 3.1 Pro Preview is Google’s frontier reasoning model, delivering enhanced software engineering performance, improved agentic reliability, and more efficient token usage across complex workflows. Building on the multimodal foundation…

Kontekst
1,05M
Wejście
2 USD
Wyjście
12 USD
DźwiękPlikObrazTekst
Zobacz model

Gemini 3.1 Pro Preview Custom Tools is a variant of Gemini 3.1 Pro that improves tool selection behavior by preventing overuse of a general bash tool when more efficient third-party…

Kontekst
1,05M
Wejście
2 USD
Wyjście
12 USD
TekstDźwiękObrazWideo
Zobacz model
Aktywny

Gemini 3.5 Flash is Google's high-efficiency multimodal model, bringing near-Pro level coding and reasoning at Flash-tier cost and speed. It is highly optimized for coding proficiency and parallel agentic execution…

Kontekst
1,05M
Wejście
1,5 USD
Wyjście
9 USD
TekstObrazWideoPlik
Zobacz model
Aktywny

Gemini 3.5 Flash Lite is a high-efficiency model from Google with upgraded agentic capabilities. It is suited for subagents that execute focused tasks within complex, multi-agent workflows.

Kontekst
1,05M
Wejście
0,30 USD
Wyjście
2,5 USD
TekstObrazWideoPlik
Zobacz model
Aktywny

Gemini 3.6 Flash is a high-efficiency model from Google for coding, agentic workflows, and web and app development. It is designed to produce polished outputs with fewer unnecessary edits and…

Kontekst
1,05M
Wejście
0,75 USD
Wyjście
3,75 USD
TekstObrazWideoPlik
Zobacz model
Aktywny

gemini-embedding-001 provides a unified cutting edge experience across domains, including science, legal, finance, and coding. This embedding model has consistently held a top spot on the Massive Text Embedding Benchmark…

Kontekst
20K
Wejście
0,15 USD
Wyjście
0,00 USD
Tekst
Zobacz model
Aktywny

Gemini Embedding 2 is Google's first multimodal embedding model. We currently support mapping text and images into a unified vector space for semantic search and retrieval-augmented generation (RAG). It supports…

Kontekst
8K
Wejście
0,20 USD
Wyjście
0,00 USD
TekstObrazPlikDźwięk
Zobacz model

Gemini Embedding 2 Preview is Google's first multimodal embedding model. We currently support mapping text and images into a unified vector space for semantic search and retrieval-augmented generation (RAG). It…

Kontekst
8K
Wejście
0,20 USD
Wyjście
0,00 USD
TekstObrazPlikDźwięk
Zobacz model
Aktywny

Gemma is a family of open models built by Google DeepMind. Gemma 4 models are multimodal, handling text and image input (with audio supported on E2B, E4B, and 12B) and generating text output. This release includes open-weights models in both pre-trained and instruction-tuned variants. Gemma 4 features a context window of up to 256K tokens and maintains multilingual support in over 140 languages.

Kontekst
262K
Wejście
0,07 USD
Wyjście
0,34 USD
ObrazTekstWideo
Zobacz model
Aktywny

Gemma is a family of open models built by Google DeepMind. Gemma 4 models are multimodal, handling text and image input (with audio supported on E2B, E4B, and 12B) and generating text output. This release includes open-weights models in both pre-trained and instruction-tuned variants. Gemma 4 features a context window of up to 256K tokens and maintains multilingual support in over 140 languages.

Kontekst
262K
Wejście
0,10 USD
Wyjście
0,34 USD
ObrazTekstWideo
Zobacz model

This model always redirects to the latest model in the Google Gemini Flash family.

Kontekst
1,05M
Wejście
0,375 USD
Wyjście
1,88 USD
TekstObrazWideoPlik
Zobacz model

This model always redirects to the latest model in the Google Gemini Pro family.

Kontekst
1,05M
Wejście
2 USD
Wyjście
12 USD
DźwiękPlikObrazTekst
Zobacz model
Aktywny

30 second duration clips are priced at $0.04 per clip. Lyria 3 is Google's family of music generation models, available through the Gemini API. With Lyria 3, you can generate…

Kontekst
1,05M
Wejście
0,00 USD
Wyjście
0,00 USD
TekstObraz
Zobacz model
Aktywny

Full-length songs are priced at $0.08 per song. Lyria 3 is Google's family of music generation models, available through the Gemini API. With Lyria 3, you can generate high-quality, 48kHz…

Kontekst
1,05M
Wejście
0,00 USD
Wyjście
0,00 USD
TekstObraz
Zobacz model

Gemini 2.5 Flash Image, a.k.a. "Nano Banana," is now generally available. It is a state of the art image generation model with contextual understanding. It is capable of image generation,…

Kontekst
33K
Wejście
0,30 USD
Wyjście
2,5 USD
ObrazTekst
Zobacz model

Gemini 3.1 Flash Image Preview, a.k.a. "Nano Banana 2," is Google’s latest state of the art image generation and editing model, delivering Pro-level visual quality at Flash speed. It combines…

Kontekst
66K
Wejście
0,50 USD
Wyjście
3 USD
ObrazTekst
Zobacz model

Nano Banana 2 Lite (Gemini 3.1 Flash Lite Image) is Google's fastest, most cost-efficient Gemini image model, built for high-velocity developer pipelines and rapid-fire visual exploration. It delivers text-to-image generation…

Kontekst
66K
Wejście
0,25 USD
Wyjście
1,5 USD
ObrazTekst
Zobacz model

Nano Banana Pro is Google’s most advanced image-generation and editing model, built on Gemini 3 Pro. It extends the original Nano Banana with significantly improved multimodal reasoning, real-world grounding, and…

Kontekst
66K
Wejście
2 USD
Wyjście
12 USD
ObrazTekst
Zobacz model

Nano Banana Pro is Google’s most advanced image-generation and editing model, built on Gemini 3 Pro. It extends the original Nano Banana with significantly improved multimodal reasoning, real-world grounding, and…

Kontekst
66K
Wejście
2 USD
Wyjście
12 USD
ObrazTekst
Zobacz model

Google

Veo 3.1

Aktywny

Google's state-of-the-art video generation model, built for maximum visual fidelity in final production cuts. Veo 3.1 generates high-quality 1080p video from text or image prompts with native synchronized audio —…

Kontekst
0
Wejście
0,00 USD
Wyjście
0,00 USD
TekstObraz
Zobacz model
Aktywny

Google's mid-tier video generation model balancing speed and quality. Veo 3.1 Fast generates high-quality video from text or image prompts with native synchronized audio, offering faster turnaround than Veo 3.1…

Kontekst
0
Wejście
0,00 USD
Wyjście
0,00 USD
TekstObraz
Zobacz model
Aktywny

Google's most cost-effective video generation model, designed for high-volume applications and rapid iteration. Veo 3.1 Lite generates 720p and 1080p video from text or image prompts with native synchronized audio…

Kontekst
0
Wejście
0,00 USD
Wyjście
0,00 USD
TekstObraz
Zobacz model
OUR METHOD

Jak AIToolly przetwarza dane modeli

Fakty dotyczące modelu i endpointu dostawcy są rozdzielone. Katalogi zewnętrzne służą do odkrywania i migawek, a oficjalna dokumentacja i zweryfikowane karty modeli mają pierwszeństwo dla danych modelu.

01Źródła i weryfikacja
02Zweryfikowano
03Możliwość