Google AI 모델

Google의 선별 AI 모델 29개를 컨텍스트, API 가격, 모달리티, 기능으로 비교합니다.

29개의 수집 모델

모델 디렉토리

현재 29개 모델

비교
사용 가능

A fast multimodal Gemini model listed for responsive agent workflows, coding and multi-step reasoning.

컨텍스트
1.05M
입력
US$0.375
출력
US$1.88
텍스트이미지비디오파일
모델 보기

Google

Chirp 3

사용 가능

Chirp 3 is Google's latest multilingual speech-to-text model. It offers enhanced transcription accuracy across 24 GA languages and 77+ preview languages, with support for automatic language detection, automatic punctuation, and…

컨텍스트
0
입력
US$16,000
출력
US$0.00
오디오
모델 보기
사용 가능

Gemini 3 Flash Preview is a high speed, high value thinking model designed for agentic workflows, multi turn chat, and coding assistance. It delivers near Pro level reasoning and tool…

컨텍스트
1.05M
입력
US$0.50
출력
US$3
텍스트이미지파일오디오
모델 보기
사용 가능

Gemini 3.1 Flash Lite is Google’s GA high-efficiency multimodal model optimized for low-latency, high-volume workloads. It supports text, image, video, audio, and PDF inputs, and is designed for lightweight agentic…

컨텍스트
1.05M
입력
US$0.25
출력
US$1.5
텍스트이미지비디오파일
모델 보기
사용 가능

Gemini 3.1 Flash Lite Preview is Google's high-efficiency model optimized for high-volume use cases. It outperforms Gemini 2.5 Flash Lite on overall quality and approaches Gemini 2.5 Flash performance across…

컨텍스트
1.05M
입력
US$0.25
출력
US$1.5
텍스트이미지비디오파일
모델 보기
사용 가능

Gemini 3.1 Flash TTS Preview is a text-to-speech model from Google, and a substantial generational step up from Gemini 2.5 Flash TTS. It takes text input and produces audio output…

컨텍스트
33K
입력
US$1
출력
US$20
텍스트
모델 보기
사용 가능

Gemini 3.1 Pro Preview is Google’s frontier reasoning model, delivering enhanced software engineering performance, improved agentic reliability, and more efficient token usage across complex workflows. Building on the multimodal foundation…

컨텍스트
1.05M
입력
US$2
출력
US$12
오디오파일이미지텍스트
모델 보기

Gemini 3.1 Pro Preview Custom Tools is a variant of Gemini 3.1 Pro that improves tool selection behavior by preventing overuse of a general bash tool when more efficient third-party…

컨텍스트
1.05M
입력
US$2
출력
US$12
텍스트오디오이미지비디오
모델 보기
사용 가능

Gemini 3.5 Flash is Google's high-efficiency multimodal model, bringing near-Pro level coding and reasoning at Flash-tier cost and speed. It is highly optimized for coding proficiency and parallel agentic execution…

컨텍스트
1.05M
입력
US$1.5
출력
US$9
텍스트이미지비디오파일
모델 보기
사용 가능

Gemini 3.5 Flash Lite is a high-efficiency model from Google with upgraded agentic capabilities. It is suited for subagents that execute focused tasks within complex, multi-agent workflows.

컨텍스트
1.05M
입력
US$0.30
출력
US$2.5
텍스트이미지비디오파일
모델 보기
사용 가능

Gemini 3.6 Flash is a high-efficiency model from Google for coding, agentic workflows, and web and app development. It is designed to produce polished outputs with fewer unnecessary edits and…

컨텍스트
1.05M
입력
US$0.75
출력
US$3.75
텍스트이미지비디오파일
모델 보기
사용 가능

gemini-embedding-001 provides a unified cutting edge experience across domains, including science, legal, finance, and coding. This embedding model has consistently held a top spot on the Massive Text Embedding Benchmark…

컨텍스트
20K
입력
US$0.15
출력
US$0.00
텍스트
모델 보기
사용 가능

Gemini Embedding 2 is Google's first multimodal embedding model. We currently support mapping text and images into a unified vector space for semantic search and retrieval-augmented generation (RAG). It supports…

컨텍스트
8K
입력
US$0.20
출력
US$0.00
텍스트이미지파일오디오
모델 보기
사용 가능

Gemini Embedding 2 Preview is Google's first multimodal embedding model. We currently support mapping text and images into a unified vector space for semantic search and retrieval-augmented generation (RAG). It…

컨텍스트
8K
입력
US$0.20
출력
US$0.00
텍스트이미지파일오디오
모델 보기
사용 가능

Gemma is a family of open models built by Google DeepMind. Gemma 4 models are multimodal, handling text and image input (with audio supported on E2B, E4B, and 12B) and generating text output. This release includes open-weights models in both pre-trained and instruction-tuned variants. Gemma 4 features a context window of up to 256K tokens and maintains multilingual support in over 140 languages.

컨텍스트
262K
입력
US$0.07
출력
US$0.34
이미지텍스트비디오
모델 보기
사용 가능

Gemma is a family of open models built by Google DeepMind. Gemma 4 models are multimodal, handling text and image input (with audio supported on E2B, E4B, and 12B) and generating text output. This release includes open-weights models in both pre-trained and instruction-tuned variants. Gemma 4 features a context window of up to 256K tokens and maintains multilingual support in over 140 languages.

컨텍스트
262K
입력
US$0.10
출력
US$0.34
이미지텍스트비디오
모델 보기
사용 가능

This model always redirects to the latest model in the Google Gemini Flash family.

컨텍스트
1.05M
입력
US$0.375
출력
US$1.88
텍스트이미지비디오파일
모델 보기
사용 가능

This model always redirects to the latest model in the Google Gemini Pro family.

컨텍스트
1.05M
입력
US$2
출력
US$12
오디오파일이미지텍스트
모델 보기
사용 가능

30 second duration clips are priced at $0.04 per clip. Lyria 3 is Google's family of music generation models, available through the Gemini API. With Lyria 3, you can generate…

컨텍스트
1.05M
입력
US$0.00
출력
US$0.00
텍스트이미지
모델 보기
사용 가능

Full-length songs are priced at $0.08 per song. Lyria 3 is Google's family of music generation models, available through the Gemini API. With Lyria 3, you can generate high-quality, 48kHz…

컨텍스트
1.05M
입력
US$0.00
출력
US$0.00
텍스트이미지
모델 보기

Gemini 2.5 Flash Image, a.k.a. "Nano Banana," is now generally available. It is a state of the art image generation model with contextual understanding. It is capable of image generation,…

컨텍스트
33K
입력
US$0.30
출력
US$2.5
이미지텍스트
모델 보기

Gemini 3.1 Flash Image Preview, a.k.a. "Nano Banana 2," is Google’s latest state of the art image generation and editing model, delivering Pro-level visual quality at Flash speed. It combines…

컨텍스트
66K
입력
US$0.50
출력
US$3
이미지텍스트
모델 보기

Nano Banana 2 Lite (Gemini 3.1 Flash Lite Image) is Google's fastest, most cost-efficient Gemini image model, built for high-velocity developer pipelines and rapid-fire visual exploration. It delivers text-to-image generation…

컨텍스트
66K
입력
US$0.25
출력
US$1.5
이미지텍스트
모델 보기

Nano Banana Pro is Google’s most advanced image-generation and editing model, built on Gemini 3 Pro. It extends the original Nano Banana with significantly improved multimodal reasoning, real-world grounding, and…

컨텍스트
66K
입력
US$2
출력
US$12
이미지텍스트
모델 보기

Nano Banana Pro is Google’s most advanced image-generation and editing model, built on Gemini 3 Pro. It extends the original Nano Banana with significantly improved multimodal reasoning, real-world grounding, and…

컨텍스트
66K
입력
US$2
출력
US$12
이미지텍스트
모델 보기

Google

Veo 3.1

사용 가능

Google's state-of-the-art video generation model, built for maximum visual fidelity in final production cuts. Veo 3.1 generates high-quality 1080p video from text or image prompts with native synchronized audio —…

컨텍스트
0
입력
US$0.00
출력
US$0.00
텍스트이미지
모델 보기
사용 가능

Google's mid-tier video generation model balancing speed and quality. Veo 3.1 Fast generates high-quality video from text or image prompts with native synchronized audio, offering faster turnaround than Veo 3.1…

컨텍스트
0
입력
US$0.00
출력
US$0.00
텍스트이미지
모델 보기
사용 가능

Google's most cost-effective video generation model, designed for high-volume applications and rapid iteration. Veo 3.1 Lite generates 720p and 1080p video from text or image prompts with native synchronized audio…

컨텍스트
0
입력
US$0.00
출력
US$0.00
텍스트이미지
모델 보기
OUR METHOD

AIToolly의 모델 데이터 처리 방식

모델 고유 정보와 제공사 엔드포인트 정보를 분리합니다. 서드파티 카탈로그는 탐색과 스냅샷에 사용하며 모델 정보는 공식 문서와 검증된 모델 카드를 우선합니다.

01출처 및 검증
02검증일
03기능