Google AIモデル

Googleの選定AIモデル29件のコンテキスト、API価格、モダリティ、機能を比較します。

29件の収録モデル

モデル一覧

この表示に 29 モデル

比較
利用可能

A fast multimodal Gemini model listed for responsive agent workflows, coding and multi-step reasoning.

コンテキスト
1.05M
入力
$0.375
出力
$1.88
テキスト画像動画ファイル
モデルを見る

Google

Chirp 3

利用可能

Chirp 3 is Google's latest multilingual speech-to-text model. It offers enhanced transcription accuracy across 24 GA languages and 77+ preview languages, with support for automatic language detection, automatic punctuation, and…

コンテキスト
0
入力
$16,000
出力
$0.00
音声
モデルを見る
利用可能

Gemini 3 Flash Preview is a high speed, high value thinking model designed for agentic workflows, multi turn chat, and coding assistance. It delivers near Pro level reasoning and tool…

コンテキスト
1.05M
入力
$0.50
出力
$3
テキスト画像ファイル音声
モデルを見る
利用可能

Gemini 3.1 Flash Lite is Google’s GA high-efficiency multimodal model optimized for low-latency, high-volume workloads. It supports text, image, video, audio, and PDF inputs, and is designed for lightweight agentic…

コンテキスト
1.05M
入力
$0.25
出力
$1.5
テキスト画像動画ファイル
モデルを見る
利用可能

Gemini 3.1 Flash Lite Preview is Google's high-efficiency model optimized for high-volume use cases. It outperforms Gemini 2.5 Flash Lite on overall quality and approaches Gemini 2.5 Flash performance across…

コンテキスト
1.05M
入力
$0.25
出力
$1.5
テキスト画像動画ファイル
モデルを見る
利用可能

Gemini 3.1 Flash TTS Preview is a text-to-speech model from Google, and a substantial generational step up from Gemini 2.5 Flash TTS. It takes text input and produces audio output…

コンテキスト
33K
入力
$1
出力
$20
テキスト
モデルを見る
利用可能

Gemini 3.1 Pro Preview is Google’s frontier reasoning model, delivering enhanced software engineering performance, improved agentic reliability, and more efficient token usage across complex workflows. Building on the multimodal foundation…

コンテキスト
1.05M
入力
$2
出力
$12
音声ファイル画像テキスト
モデルを見る

Gemini 3.1 Pro Preview Custom Tools is a variant of Gemini 3.1 Pro that improves tool selection behavior by preventing overuse of a general bash tool when more efficient third-party…

コンテキスト
1.05M
入力
$2
出力
$12
テキスト音声画像動画
モデルを見る
利用可能

Gemini 3.5 Flash is Google's high-efficiency multimodal model, bringing near-Pro level coding and reasoning at Flash-tier cost and speed. It is highly optimized for coding proficiency and parallel agentic execution…

コンテキスト
1.05M
入力
$1.5
出力
$9
テキスト画像動画ファイル
モデルを見る
利用可能

Gemini 3.5 Flash Lite is a high-efficiency model from Google with upgraded agentic capabilities. It is suited for subagents that execute focused tasks within complex, multi-agent workflows.

コンテキスト
1.05M
入力
$0.30
出力
$2.5
テキスト画像動画ファイル
モデルを見る
利用可能

Gemini 3.6 Flash is a high-efficiency model from Google for coding, agentic workflows, and web and app development. It is designed to produce polished outputs with fewer unnecessary edits and…

コンテキスト
1.05M
入力
$0.75
出力
$3.75
テキスト画像動画ファイル
モデルを見る
利用可能

gemini-embedding-001 provides a unified cutting edge experience across domains, including science, legal, finance, and coding. This embedding model has consistently held a top spot on the Massive Text Embedding Benchmark…

コンテキスト
20K
入力
$0.15
出力
$0.00
テキスト
モデルを見る
利用可能

Gemini Embedding 2 is Google's first multimodal embedding model. We currently support mapping text and images into a unified vector space for semantic search and retrieval-augmented generation (RAG). It supports…

コンテキスト
8K
入力
$0.20
出力
$0.00
テキスト画像ファイル音声
モデルを見る
利用可能

Gemini Embedding 2 Preview is Google's first multimodal embedding model. We currently support mapping text and images into a unified vector space for semantic search and retrieval-augmented generation (RAG). It…

コンテキスト
8K
入力
$0.20
出力
$0.00
テキスト画像ファイル音声
モデルを見る
利用可能

Gemma is a family of open models built by Google DeepMind. Gemma 4 models are multimodal, handling text and image input (with audio supported on E2B, E4B, and 12B) and generating text output. This release includes open-weights models in both pre-trained and instruction-tuned variants. Gemma 4 features a context window of up to 256K tokens and maintains multilingual support in over 140 languages.

コンテキスト
262K
入力
$0.07
出力
$0.34
画像テキスト動画
モデルを見る
利用可能

Gemma is a family of open models built by Google DeepMind. Gemma 4 models are multimodal, handling text and image input (with audio supported on E2B, E4B, and 12B) and generating text output. This release includes open-weights models in both pre-trained and instruction-tuned variants. Gemma 4 features a context window of up to 256K tokens and maintains multilingual support in over 140 languages.

コンテキスト
262K
入力
$0.10
出力
$0.34
画像テキスト動画
モデルを見る
利用可能

This model always redirects to the latest model in the Google Gemini Flash family.

コンテキスト
1.05M
入力
$0.375
出力
$1.88
テキスト画像動画ファイル
モデルを見る
利用可能

This model always redirects to the latest model in the Google Gemini Pro family.

コンテキスト
1.05M
入力
$2
出力
$12
音声ファイル画像テキスト
モデルを見る
利用可能

30 second duration clips are priced at $0.04 per clip. Lyria 3 is Google's family of music generation models, available through the Gemini API. With Lyria 3, you can generate…

コンテキスト
1.05M
入力
$0.00
出力
$0.00
テキスト画像
モデルを見る
利用可能

Full-length songs are priced at $0.08 per song. Lyria 3 is Google's family of music generation models, available through the Gemini API. With Lyria 3, you can generate high-quality, 48kHz…

コンテキスト
1.05M
入力
$0.00
出力
$0.00
テキスト画像
モデルを見る

Gemini 2.5 Flash Image, a.k.a. "Nano Banana," is now generally available. It is a state of the art image generation model with contextual understanding. It is capable of image generation,…

コンテキスト
33K
入力
$0.30
出力
$2.5
画像テキスト
モデルを見る

Gemini 3.1 Flash Image Preview, a.k.a. "Nano Banana 2," is Google’s latest state of the art image generation and editing model, delivering Pro-level visual quality at Flash speed. It combines…

コンテキスト
66K
入力
$0.50
出力
$3
画像テキスト
モデルを見る

Nano Banana 2 Lite (Gemini 3.1 Flash Lite Image) is Google's fastest, most cost-efficient Gemini image model, built for high-velocity developer pipelines and rapid-fire visual exploration. It delivers text-to-image generation…

コンテキスト
66K
入力
$0.25
出力
$1.5
画像テキスト
モデルを見る

Nano Banana Pro is Google’s most advanced image-generation and editing model, built on Gemini 3 Pro. It extends the original Nano Banana with significantly improved multimodal reasoning, real-world grounding, and…

コンテキスト
66K
入力
$2
出力
$12
画像テキスト
モデルを見る

Nano Banana Pro is Google’s most advanced image-generation and editing model, built on Gemini 3 Pro. It extends the original Nano Banana with significantly improved multimodal reasoning, real-world grounding, and…

コンテキスト
66K
入力
$2
出力
$12
画像テキスト
モデルを見る

Google

Veo 3.1

利用可能

Google's state-of-the-art video generation model, built for maximum visual fidelity in final production cuts. Veo 3.1 generates high-quality 1080p video from text or image prompts with native synchronized audio —…

コンテキスト
0
入力
$0.00
出力
$0.00
テキスト画像
モデルを見る
利用可能

Google's mid-tier video generation model balancing speed and quality. Veo 3.1 Fast generates high-quality video from text or image prompts with native synchronized audio, offering faster turnaround than Veo 3.1…

コンテキスト
0
入力
$0.00
出力
$0.00
テキスト画像
モデルを見る
利用可能

Google's most cost-effective video generation model, designed for high-volume applications and rapid iteration. Veo 3.1 Lite generates 720p and 1080p video from text or image prompts with native synchronized audio…

コンテキスト
0
入力
$0.00
出力
$0.00
テキスト画像
モデルを見る
OUR METHOD

AIToollyのモデルデータについて

モデル固有の事実とプロバイダーエンドポイントの事実は分離して管理します。第三者カタログは検索とスナップショットに使い、モデル情報には公式文書と検証済みモデルカードを優先します。

01出典と確認
02確認日
03機能