Google KI-Modelle

Entdecken Sie 29 kuratierte KI-Modelle von Google und vergleichen Sie Kontext, Preise, Modalitäten und Fähigkeiten.

29erfasste Modelle

Modellverzeichnis

29 Modelle in dieser Ansicht

Vergleichen
Aktiv

A fast multimodal Gemini model listed for responsive agent workflows, coding and multi-step reasoning.

Kontext
1,05M
Eingabe
0,375 $
Ausgabe
1,88 $
TextBildVideoDatei
Modell ansehen

Google

Chirp 3

Aktiv

Chirp 3 is Google's latest multilingual speech-to-text model. It offers enhanced transcription accuracy across 24 GA languages and 77+ preview languages, with support for automatic language detection, automatic punctuation, and…

Kontext
0
Eingabe
16.000 $
Ausgabe
0,00 $
Audio
Modell ansehen

Gemini 3 Flash Preview is a high speed, high value thinking model designed for agentic workflows, multi turn chat, and coding assistance. It delivers near Pro level reasoning and tool…

Kontext
1,05M
Eingabe
0,50 $
Ausgabe
3 $
TextBildDateiAudio
Modell ansehen

Gemini 3.1 Flash Lite is Google’s GA high-efficiency multimodal model optimized for low-latency, high-volume workloads. It supports text, image, video, audio, and PDF inputs, and is designed for lightweight agentic…

Kontext
1,05M
Eingabe
0,25 $
Ausgabe
1,5 $
TextBildVideoDatei
Modell ansehen

Gemini 3.1 Flash Lite Preview is Google's high-efficiency model optimized for high-volume use cases. It outperforms Gemini 2.5 Flash Lite on overall quality and approaches Gemini 2.5 Flash performance across…

Kontext
1,05M
Eingabe
0,25 $
Ausgabe
1,5 $
TextBildVideoDatei
Modell ansehen

Gemini 3.1 Flash TTS Preview is a text-to-speech model from Google, and a substantial generational step up from Gemini 2.5 Flash TTS. It takes text input and produces audio output…

Kontext
33K
Eingabe
1 $
Ausgabe
20 $
Text
Modell ansehen

Gemini 3.1 Pro Preview is Google’s frontier reasoning model, delivering enhanced software engineering performance, improved agentic reliability, and more efficient token usage across complex workflows. Building on the multimodal foundation…

Kontext
1,05M
Eingabe
2 $
Ausgabe
12 $
AudioDateiBildText
Modell ansehen

Gemini 3.1 Pro Preview Custom Tools is a variant of Gemini 3.1 Pro that improves tool selection behavior by preventing overuse of a general bash tool when more efficient third-party…

Kontext
1,05M
Eingabe
2 $
Ausgabe
12 $
TextAudioBildVideo
Modell ansehen
Aktiv

Gemini 3.5 Flash is Google's high-efficiency multimodal model, bringing near-Pro level coding and reasoning at Flash-tier cost and speed. It is highly optimized for coding proficiency and parallel agentic execution…

Kontext
1,05M
Eingabe
1,5 $
Ausgabe
9 $
TextBildVideoDatei
Modell ansehen

Gemini 3.5 Flash Lite is a high-efficiency model from Google with upgraded agentic capabilities. It is suited for subagents that execute focused tasks within complex, multi-agent workflows.

Kontext
1,05M
Eingabe
0,30 $
Ausgabe
2,5 $
TextBildVideoDatei
Modell ansehen
Aktiv

Gemini 3.6 Flash is a high-efficiency model from Google for coding, agentic workflows, and web and app development. It is designed to produce polished outputs with fewer unnecessary edits and…

Kontext
1,05M
Eingabe
0,75 $
Ausgabe
3,75 $
TextBildVideoDatei
Modell ansehen

gemini-embedding-001 provides a unified cutting edge experience across domains, including science, legal, finance, and coding. This embedding model has consistently held a top spot on the Massive Text Embedding Benchmark…

Kontext
20K
Eingabe
0,15 $
Ausgabe
0,00 $
Text
Modell ansehen
Aktiv

Gemini Embedding 2 is Google's first multimodal embedding model. We currently support mapping text and images into a unified vector space for semantic search and retrieval-augmented generation (RAG). It supports…

Kontext
8K
Eingabe
0,20 $
Ausgabe
0,00 $
TextBildDateiAudio
Modell ansehen

Gemini Embedding 2 Preview is Google's first multimodal embedding model. We currently support mapping text and images into a unified vector space for semantic search and retrieval-augmented generation (RAG). It…

Kontext
8K
Eingabe
0,20 $
Ausgabe
0,00 $
TextBildDateiAudio
Modell ansehen
Aktiv

Gemma is a family of open models built by Google DeepMind. Gemma 4 models are multimodal, handling text and image input (with audio supported on E2B, E4B, and 12B) and generating text output. This release includes open-weights models in both pre-trained and instruction-tuned variants. Gemma 4 features a context window of up to 256K tokens and maintains multilingual support in over 140 languages.

Kontext
262K
Eingabe
0,07 $
Ausgabe
0,34 $
BildTextVideo
Modell ansehen
Aktiv

Gemma is a family of open models built by Google DeepMind. Gemma 4 models are multimodal, handling text and image input (with audio supported on E2B, E4B, and 12B) and generating text output. This release includes open-weights models in both pre-trained and instruction-tuned variants. Gemma 4 features a context window of up to 256K tokens and maintains multilingual support in over 140 languages.

Kontext
262K
Eingabe
0,10 $
Ausgabe
0,34 $
BildTextVideo
Modell ansehen

This model always redirects to the latest model in the Google Gemini Flash family.

Kontext
1,05M
Eingabe
0,375 $
Ausgabe
1,88 $
TextBildVideoDatei
Modell ansehen

This model always redirects to the latest model in the Google Gemini Pro family.

Kontext
1,05M
Eingabe
2 $
Ausgabe
12 $
AudioDateiBildText
Modell ansehen

30 second duration clips are priced at $0.04 per clip. Lyria 3 is Google's family of music generation models, available through the Gemini API. With Lyria 3, you can generate…

Kontext
1,05M
Eingabe
0,00 $
Ausgabe
0,00 $
TextBild
Modell ansehen
Aktiv

Full-length songs are priced at $0.08 per song. Lyria 3 is Google's family of music generation models, available through the Gemini API. With Lyria 3, you can generate high-quality, 48kHz…

Kontext
1,05M
Eingabe
0,00 $
Ausgabe
0,00 $
TextBild
Modell ansehen

Gemini 2.5 Flash Image, a.k.a. "Nano Banana," is now generally available. It is a state of the art image generation model with contextual understanding. It is capable of image generation,…

Kontext
33K
Eingabe
0,30 $
Ausgabe
2,5 $
BildText
Modell ansehen

Gemini 3.1 Flash Image Preview, a.k.a. "Nano Banana 2," is Google’s latest state of the art image generation and editing model, delivering Pro-level visual quality at Flash speed. It combines…

Kontext
66K
Eingabe
0,50 $
Ausgabe
3 $
BildText
Modell ansehen

Nano Banana 2 Lite (Gemini 3.1 Flash Lite Image) is Google's fastest, most cost-efficient Gemini image model, built for high-velocity developer pipelines and rapid-fire visual exploration. It delivers text-to-image generation…

Kontext
66K
Eingabe
0,25 $
Ausgabe
1,5 $
BildText
Modell ansehen

Nano Banana Pro is Google’s most advanced image-generation and editing model, built on Gemini 3 Pro. It extends the original Nano Banana with significantly improved multimodal reasoning, real-world grounding, and…

Kontext
66K
Eingabe
2 $
Ausgabe
12 $
BildText
Modell ansehen

Nano Banana Pro is Google’s most advanced image-generation and editing model, built on Gemini 3 Pro. It extends the original Nano Banana with significantly improved multimodal reasoning, real-world grounding, and…

Kontext
66K
Eingabe
2 $
Ausgabe
12 $
BildText
Modell ansehen

Google

Veo 3.1

Aktiv

Google's state-of-the-art video generation model, built for maximum visual fidelity in final production cuts. Veo 3.1 generates high-quality 1080p video from text or image prompts with native synchronized audio —…

Kontext
0
Eingabe
0,00 $
Ausgabe
0,00 $
TextBild
Modell ansehen
Aktiv

Google's mid-tier video generation model balancing speed and quality. Veo 3.1 Fast generates high-quality video from text or image prompts with native synchronized audio, offering faster turnaround than Veo 3.1…

Kontext
0
Eingabe
0,00 $
Ausgabe
0,00 $
TextBild
Modell ansehen
Aktiv

Google's most cost-effective video generation model, designed for high-volume applications and rapid iteration. Veo 3.1 Lite generates 720p and 1080p video from text or image prompts with native synchronized audio…

Kontext
0
Eingabe
0,00 $
Ausgabe
0,00 $
TextBild
Modell ansehen
OUR METHOD

So behandelt AIToolly Modelldaten

Modelleigene Fakten und Fakten zu Anbieter-Endpunkten bleiben getrennt. Drittanbieter-Kataloge dienen der Ermittlung und Anbieter-Snapshots; für Modelldaten werden offizielle Dokumentationen und verifizierte Modellkarten bevorzugt.

01Quellen und Verifizierung
02Verifiziert
03Fähigkeit