長文コンテキスト向けAIモデル

長文コンテキスト向けの選定AIモデル209件を、追跡可能なプロバイダー情報で比較します。

209件の収録モデル

モデル一覧

この表示に 209 モデル

比較
利用可能

A flagship GPT-5.6 model listed for complex reasoning, coding and multi-step agent workflows.

コンテキスト
1.05M
入力
$2.5
出力
$15
ファイル画像テキスト
モデルを見る
利用可能

A lower-cost GPT-5.6 model for high-volume chat, classification and lightweight agent workflows.

コンテキスト
1.05M
入力
$0.20
出力
$1.2
ファイル画像テキスト
モデルを見る
利用可能

A Sonnet-class model listed for coding, agents and professional knowledge work with adaptive reasoning.

コンテキスト
1M
入力
$2
出力
$10
テキスト画像ファイル
モデルを見る
利用可能

A fast multimodal Gemini model listed for responsive agent workflows, coding and multi-step reasoning.

コンテキスト
1.05M
入力
$0.375
出力
$1.88
テキスト画像動画ファイル
モデルを見る
利用可能

A Grok model listed for coding, knowledge work and STEM tasks with text, image and file input.

コンテキスト
500K
入力
$2
出力
$6
テキスト画像ファイル
モデルを見る
利用可能

A text model listed as a large mixture-of-experts release with a 1,048,576-token provider context window.

コンテキスト
1.05M
入力
$1.32
出力
$3.96
テキスト
モデルを見る

AionLabs

Aion-2.0

利用可能

Aion-2.0 is a variant of DeepSeek V3.2 optimized for immersive roleplaying and storytelling. It is particularly strong at introducing tension, crises, and conflict into stories, making narratives feel more engaging.…

コンテキスト
131K
入力
$0.80
出力
$1.6
テキスト
モデルを見る

AionLabs

Aion-3.0

利用可能

Aion-3.0 is a multi-model roleplaying and storytelling system from AionLabs, built on the GLM family of models. It uses a collaborative generation process in which multiple specialized models each contribute…

コンテキスト
131K
入力
$3
出力
$6
テキスト
モデルを見る
利用可能

Aion-3.0 Mini is a multi-model roleplaying and storytelling system from AionLabs, built on the DeepSeek family of models. It uses a collaborative generation process in which multiple specialized models each…

コンテキスト
131K
入力
$0.70
出力
$1.4
テキスト
モデルを見る
利用可能

This model always redirects to the latest model in the Anthropic Claude Haiku family.

コンテキスト
200K
入力
$1
出力
$5
テキスト画像ファイル
モデルを見る
利用可能

Auto Router (Beta) is a task-aware router from the provider catalog. It classifies each request, then routes it the most popular model for that task based on aggregate spend, filtered by your…

コンテキスト
2M
入力
不明
出力
不明
テキスト画像音声ファイル
モデルを見る
利用可能

Transform your natural language requests into structured the provider catalog API request objects. Describe what you want to accomplish with AI models, and Body Builder will construct the appropriate API calls. Example:…

コンテキスト
128K
入力
不明
出力
不明
テキスト
モデルを見る
利用可能

Claude Fable 5 is a Mythos-class model from Anthropic, built for autonomous knowledge work and coding. It supports text, image, and file inputs with text output, with reasoning support and…

コンテキスト
1M
入力
$10
出力
$50
テキスト画像ファイル
モデルを見る
利用可能

This model always redirects to the latest model in the Claude Fable family.

コンテキスト
1M
入力
$10
出力
$50
テキスト画像ファイル
モデルを見る
利用可能

Claude Haiku 4.5 is Anthropic’s fastest and most efficient model, delivering near-frontier intelligence at a fraction of the cost and latency of larger Claude models. Matching Claude Sonnet 4’s performance…

コンテキスト
200K
入力
$1
出力
$5
テキスト画像ファイル
モデルを見る
利用可能

Claude Opus 4.5 is Anthropic’s frontier reasoning model optimized for complex software engineering, agentic workflows, and long-horizon computer use. It offers strong multimodal capabilities, competitive performance across real-world coding and…

コンテキスト
200K
入力
$5
出力
$25
ファイル画像テキスト
モデルを見る
利用可能

Opus 4.6 is Anthropic’s strongest model for coding and long-running professional tasks. It is built for agents that operate across entire workflows rather than single prompts, making it especially effective…

コンテキスト
1M
入力
$5
出力
$25
テキスト画像ファイル
モデルを見る
利用可能

Opus 4.7 is the next generation of Anthropic's Opus family, built for long-running, asynchronous agents. Building on the coding and agentic strengths of Opus 4.6, it delivers stronger performance on…

コンテキスト
1M
入力
$5
出力
$25
テキスト画像ファイル
モデルを見る
利用可能

Fast-mode variant of Opus 4.7 - identical capabilities with higher output speed at premium 6x pricing.

コンテキスト
1M
入力
$30
出力
$150
テキスト画像ファイル
モデルを見る
利用可能

Claude Opus 4.8 is Anthropic's most capable generally available model in the Opus family. It supports text, image, and file inputs with text output, with reasoning support and a 1M-token…

コンテキスト
1M
入力
$5
出力
$25
テキスト画像ファイル
モデルを見る
利用可能

Fast-mode variant of Opus 4.8 - identical capabilities with higher output speed at 2x pricing relative to regular Opus 4.8.

コンテキスト
1M
入力
$10
出力
$50
テキスト画像ファイル
モデルを見る
利用可能

Claude Opus 5 is Anthropic’s flagship model for demanding reasoning, coding, and long-horizon agentic work. It is particularly strong at end-to-end software tasks, code review and bug finding, visual analysis…

コンテキスト
1M
入力
$5
出力
$25
テキスト画像ファイル
モデルを見る
利用可能

Fast-mode variant of Opus 5 - identical capabilities with higher output speed at 2x pricing relative to regular Opus 5.

コンテキスト
1M
入力
$10
出力
$50
テキスト画像ファイル
モデルを見る
利用可能

This model always redirects to the latest model in the Claude Opus family.

コンテキスト
1M
入力
$5
出力
$25
テキスト画像ファイル
モデルを見る
利用可能

Claude Sonnet 4.5 is Anthropic’s most advanced Sonnet model to date, optimized for real-world agents and coding workflows. It delivers state-of-the-art performance on coding benchmarks such as SWE-bench Verified, with…

コンテキスト
1M
入力
$3
出力
$15
テキスト画像ファイル
モデルを見る
利用可能

Sonnet 4.6 is Anthropic's most capable Sonnet-class model yet, with frontier performance across coding, agents, and professional work. It excels at iterative development, complex codebase navigation, end-to-end project management with…

コンテキスト
1M
入力
$3
出力
$15
テキスト画像ファイル
モデルを見る

Deep Cogito

Cogito v2.1 671B

利用可能

Cogito v2.1 671B MoE represents one of the strongest open models globally, matching performance of frontier closed and open models. This model is trained using self play with reinforcement learning…

コンテキスト
128K
入力
$1.25
出力
$1.25
テキスト
モデルを見る
利用可能

> I have to praise this model for good focus. I said earlier that it still remembers it at 12K. I think my personal evaluation of it has already beaten the rest.

コンテキスト
131K
入力
$0.30
出力
$0.50
テキスト
モデルを見る
利用可能

DeepSeek-V3.1 is a hybrid model that supports both thinking mode and non-thinking mode. Compared to the previous version, this upgrade brings improvements in multiple aspects:

コンテキスト
164K
入力
$0.25
出力
$0.95
テキスト
モデルを見る
利用可能

This update maintains the model's original capabilities while addressing issues reported by users, including:

コンテキスト
164K
入力
$0.27
出力
$1
テキスト
モデルを見る
利用可能

We introduce **DeepSeek-V3.2**, a model that harmonizes high computational efficiency with superior reasoning and agent performance. Our approach is built upon three key technical breakthroughs:

コンテキスト
164K
入力
$0.269
出力
$0.40
テキスト
モデルを見る
利用可能

We are excited to announce the official release of DeepSeek-V3.2-Exp, an experimental version of our model. As an intermediate step toward our next-generation architecture, V3.2-Exp builds upon V3.1-Terminus by introducing DeepSeek Sparse Attention—a sparse attention mechanism designed to explore and validate optimizations for training and inference efficiency in long-context scenarios.

コンテキスト
164K
入力
$0.27
出力
$0.41
テキスト
モデルを見る
利用可能

We present a preview version of **DeepSeek-V4** series, including two strong Mixture-of-Experts (MoE) language models — **DeepSeek-V4-Pro** with 1.6T parameters (49B activated) and **DeepSeek-V4-Flash** with 284B parameters (13B activated) — both supporting a context length of **one million tokens**.

コンテキスト
1.02M
入力
$0.0826
出力
$0.1652
テキスト
モデルを見る
利用可能

**DeepSeek-V4-Flash-0731** is the official release of **DeepSeek-V4-Flash**, superseding the preview version, with substantially enhanced agentic capabilities. It has the same model structure as DeepSeek-V4-Flash-DSpark, i.e. it comes with a speculative decoding module attached.

コンテキスト
1.05M
入力
$0.14
出力
$0.28
テキスト
モデルを見る
利用可能

This model always redirects to the latest model in the DeepSeek V4 Flash family.

コンテキスト
1.02M
入力
$0.0786
出力
$0.1572
テキスト
モデルを見る
利用可能

We present a preview version of **DeepSeek-V4** series, including two strong Mixture-of-Experts (MoE) language models — **DeepSeek-V4-Pro** with 1.6T parameters (49B activated) and **DeepSeek-V4-Flash** with 284B parameters (13B activated) — both supporting a context length of **one million tokens**.

コンテキスト
1.05M
入力
$1.32
出力
$3.96
テキスト
モデルを見る
利用可能

Dots3-Note Preview is an open-weight mixture-of-experts model from Dots Studio, with 16B active parameters out of 280B total. It is the lightest model in the Dots 3 family and is…

コンテキスト
512K
入力
$0.00
出力
$0.00
テキスト画像
モデルを見る
利用可能

The simplest way to get free inference. the provider catalog/free is a router that selects free models at random from the models available on the provider catalog. The router smartly filters for models that…

コンテキスト
200K
入力
$0.00
出力
$0.00
テキスト画像
モデルを見る
利用可能

Fugu Ultra is the higher-performance model in Sakana AI's Fugu family. Rather than a single monolithic model, Fugu is a learned multi-agent orchestration system: a language model trained to route…

コンテキスト
1M
入力
$5
出力
$30
テキスト画像
モデルを見る

OpenRouter

Fusion

利用可能

Fusion turns your prompt into a small multi-model deliberation. A panel of expert models (see below) analyzes your prompt in parallel with web search and web fetch enabled, then a…

コンテキスト
1M
入力
不明
出力
不明
テキスト
モデルを見る
利用可能

Gemini 3 Flash Preview is a high speed, high value thinking model designed for agentic workflows, multi turn chat, and coding assistance. It delivers near Pro level reasoning and tool…

コンテキスト
1.05M
入力
$0.50
出力
$3
テキスト画像ファイル音声
モデルを見る
利用可能

Gemini 3.1 Flash Lite is Google’s GA high-efficiency multimodal model optimized for low-latency, high-volume workloads. It supports text, image, video, audio, and PDF inputs, and is designed for lightweight agentic…

コンテキスト
1.05M
入力
$0.25
出力
$1.5
テキスト画像動画ファイル
モデルを見る
利用可能

Gemini 3.1 Flash Lite Preview is Google's high-efficiency model optimized for high-volume use cases. It outperforms Gemini 2.5 Flash Lite on overall quality and approaches Gemini 2.5 Flash performance across…

コンテキスト
1.05M
入力
$0.25
出力
$1.5
テキスト画像動画ファイル
モデルを見る
利用可能

Gemini 3.1 Pro Preview is Google’s frontier reasoning model, delivering enhanced software engineering performance, improved agentic reliability, and more efficient token usage across complex workflows. Building on the multimodal foundation…

コンテキスト
1.05M
入力
$2
出力
$12
音声ファイル画像テキスト
モデルを見る

Gemini 3.1 Pro Preview Custom Tools is a variant of Gemini 3.1 Pro that improves tool selection behavior by preventing overuse of a general bash tool when more efficient third-party…

コンテキスト
1.05M
入力
$2
出力
$12
テキスト音声画像動画
モデルを見る
利用可能

Gemini 3.5 Flash is Google's high-efficiency multimodal model, bringing near-Pro level coding and reasoning at Flash-tier cost and speed. It is highly optimized for coding proficiency and parallel agentic execution…

コンテキスト
1.05M
入力
$1.5
出力
$9
テキスト画像動画ファイル
モデルを見る
利用可能

Gemini 3.5 Flash Lite is a high-efficiency model from Google with upgraded agentic capabilities. It is suited for subagents that execute focused tasks within complex, multi-agent workflows.

コンテキスト
1.05M
入力
$0.30
出力
$2.5
テキスト画像動画ファイル
モデルを見る
利用可能

Gemini 3.6 Flash is a high-efficiency model from Google for coding, agentic workflows, and web and app development. It is designed to produce polished outputs with fewer unnecessary edits and…

コンテキスト
1.05M
入力
$0.75
出力
$3.75
テキスト画像動画ファイル
モデルを見る
利用可能

Gemma is a family of open models built by Google DeepMind. Gemma 4 models are multimodal, handling text and image input (with audio supported on E2B, E4B, and 12B) and generating text output. This release includes open-weights models in both pre-trained and instruction-tuned variants. Gemma 4 features a context window of up to 256K tokens and maintains multilingual support in over 140 languages.

コンテキスト
262K
入力
$0.07
出力
$0.34
画像テキスト動画
モデルを見る
利用可能

Gemma is a family of open models built by Google DeepMind. Gemma 4 models are multimodal, handling text and image input (with audio supported on E2B, E4B, and 12B) and generating text output. This release includes open-weights models in both pre-trained and instruction-tuned variants. Gemma 4 features a context window of up to 256K tokens and maintains multilingual support in over 140 languages.

コンテキスト
262K
入力
$0.10
出力
$0.34
画像テキスト動画
モデルを見る
利用可能

📖 Check out the GLM-4.6 technical blog, technical report(GLM-4.5), and Zhipu AI technical documentation.

コンテキスト
203K
入力
$0.50
出力
$2
テキスト
モデルを見る
利用可能

This model is part of the GLM-V family of models, introduced in the paper GLM-4.1V-Thinking and GLM-4.5V: Towards Versatile Multimodal Reasoning with Scalable Reinforcement Learning.

コンテキスト
131K
入力
$0.30
出力
$0.90
画像テキスト動画
モデルを見る
利用可能

You can also see significant improvements in many other scenarios such as chat, creative writing, and role-play scenario.

コンテキスト
203K
入力
$0.40
出力
$1.75
テキスト
モデルを見る
利用可能

GLM-4.7-Flash is a 30B-A3B MoE model. As the strongest model in the 30B class, GLM-4.7-Flash offers a new option for lightweight deployment that balances performance and efficiency.

コンテキスト
203K
入力
$0.06
出力
$0.40
テキスト
モデルを見る

Z.ai

GLM 5

利用可能

We are launching GLM-5, targeting complex systems engineering and long-horizon agentic tasks. Scaling is still one of the most important ways to improve the intelligence efficiency of Artificial General Intelligence (AGI). Compared to GLM-4.5, GLM-5 scales from 355B parameters (32B active) to 744B parameters (40B active), and increases pre-training data from 23T to 28.5T tokens. GLM-5 also integrates DeepSeek Sparse Attention (DSA), largely reducing deployment cost while preserving long-context capacity.

コンテキスト
198K
入力
$0.60
出力
$1.92
テキスト
モデルを見る
利用可能

GLM-5 Turbo is a new model from Z.ai designed for fast inference and strong performance in agent-driven environments such as OpenClaw scenarios. It is deeply optimized for real-world agent workflows…

コンテキスト
203K
入力
$1.2
出力
$4
テキスト
モデルを見る
利用可能

GLM-5.1 is our next-generation flagship model for agentic engineering, with significantly stronger coding capabilities than its predecessor. It achieves state-of-the-art performance on SWE-Bench Pro and leads GLM-5 by a wide margin on NL2Repo (repo generation) and Terminal-Bench 2.0 (real-world terminal tasks).

コンテキスト
200K
入力
$0.966
出力
$3.04
テキスト
モデルを見る
利用可能

We're introducing GLM-5.2, our latest flagship model for long-horizon tasks. It marks a substantial leap in long-horizon task capability over its predecessor GLM-5.1 and, for the first time, delivers that capability on a **solid 1M-token context**. GLM-5.2's new capabilities include:

コンテキスト
1.05M
入力
$0.50
出力
$3.15
テキスト
モデルを見る
利用可能

GLM-5V-Turbo is Z.ai’s first native multimodal agent foundation model, built for vision-based coding and agent-driven tasks. It natively handles image, video, and text inputs, excels at long-horizon planning, complex coding,…

コンテキスト
203K
入力
$1.2
出力
$4
画像テキスト動画
モデルを見る
利用可能

This model always redirects to the latest model in the Google Gemini Flash family.

コンテキスト
1.05M
入力
$0.375
出力
$1.88
テキスト画像動画ファイル
モデルを見る
利用可能

This model always redirects to the latest model in the Google Gemini Pro family.

コンテキスト
1.05M
入力
$2
出力
$12
音声ファイル画像テキスト
モデルを見る

OpenAI

GPT Audio

利用可能

The gpt-audio model is OpenAI's first generally available audio model. The new snapshot features an upgraded decoder for more natural sounding voices and maintains better voice consistency. Audio is priced…

コンテキスト
128K
入力
$2.5
出力
$10
テキスト音声
モデルを見る
利用可能

A cost-efficient version of GPT Audio. The new snapshot features an upgraded decoder for more natural sounding voices and maintains better voice consistency. Input is priced at $0.60 per million…

コンテキスト
128K
入力
$0.60
出力
$2.4
テキスト音声
モデルを見る
利用可能

GPT Chat Latest points to OpenAI's stable API alias chat-latest that always resolves to the latest Instant chat model used in ChatGPT. As OpenAI rolls out new Instant model updates…

コンテキスト
400K
入力
$5
出力
$30
テキスト画像ファイル
モデルを見る
利用可能

OpenAI's GPT Image 1 generates and edits images via the dedicated Images API. Features accurate text rendering, transparent backgrounds, and up to 16 reference images for edits.

コンテキスト
400K
入力
$10
出力
$10
テキスト画像
モデルを見る
利用可能

A cost-efficient variant of GPT Image 1 for high-quality image generation at reduced latency and cost via OpenAI's dedicated Images API.

コンテキスト
400K
入力
$2.5
出力
$2.5
テキスト画像
モデルを見る
利用可能

OpenAI's latest image generation model. Supports high-fidelity image generation and editing via the dedicated Images API.

コンテキスト
400K
入力
$8
出力
$8
テキスト画像
モデルを見る
利用可能

GPT-4o Mini Transcribe is OpenAI's smaller, cost-efficient speech-to-text model built on GPT-4o Mini audio capabilities. It's priced per token (input and output), making it suitable for high-volume transcription workflows that…

コンテキスト
128K
入力
$1.25
出力
$5
音声
モデルを見る
利用可能

GPT-4o Transcribe is OpenAI's high-quality speech-to-text model built on GPT-4o audio capabilities. It's priced per token (input and output), making it suitable for workflows that benefit from token-level billing transparency.

コンテキスト
128K
入力
$2.5
出力
$10
音声
モデルを見る
利用可能

GPT-5-Codex is a specialized version of GPT-5 optimized for software engineering and coding workflows. It is designed for both interactive development sessions and long, independent execution of complex engineering tasks.…

コンテキスト
400K
入力
$0.625
出力
$5
テキスト画像
モデルを見る
利用可能

GPT-5 Image combines OpenAI's GPT-5 model with state-of-the-art image generation capabilities. It offers major improvements in reasoning, code quality, and user experience while incorporating GPT Image 1's superior instruction following,…

コンテキスト
400K
入力
$10
出力
$10
画像テキストファイル
モデルを見る
利用可能

GPT-5 Image Mini combines OpenAI's advanced language capabilities, powered by GPT-5 Mini, with GPT Image 1 Mini for efficient image generation. This natively multimodal model features superior instruction following, text…

コンテキスト
400K
入力
$2.5
出力
$2
ファイル画像テキスト
モデルを見る

OpenAI

GPT-5 Pro

利用可能

GPT-5 Pro is OpenAI’s most advanced model, offering major improvements in reasoning, code quality, and user experience. It is optimized for complex tasks that require step-by-step reasoning, instruction following, and…

コンテキスト
400K
入力
$15
出力
$120
画像テキストファイル
モデルを見る

OpenAI

GPT-5.1

利用可能

GPT-5.1 is the latest frontier-grade model in the GPT-5 series, offering stronger general-purpose reasoning, improved instruction adherence, and a more natural conversational style compared to GPT-5. It uses adaptive reasoning…

コンテキスト
400K
入力
$1.25
出力
$10
画像テキストファイル
モデルを見る
利用可能

GPT-5.1-Codex is a specialized version of GPT-5.1 optimized for software engineering and coding workflows. It is designed for both interactive development sessions and long, independent execution of complex engineering tasks.…

コンテキスト
400K
入力
$1.25
出力
$10
テキスト画像
モデルを見る
利用可能

GPT-5.1-Codex-Max is OpenAI’s latest agentic coding model, designed for long-running, high-context software development tasks. It is based on an updated version of the 5.1 reasoning stack and trained on agentic…

コンテキスト
400K
入力
$1.25
出力
$10
テキスト画像
モデルを見る
利用可能

GPT-5.1-Codex-Mini is a smaller and faster version of GPT-5.1-Codex

コンテキスト
400K
入力
$0.25
出力
$2
画像テキスト
モデルを見る

OpenAI

GPT-5.2

利用可能

GPT-5.2 is the latest frontier-grade model in the GPT-5 series, offering stronger agentic and long context perfomance compared to GPT-5.1. It uses adaptive reasoning to allocate computation dynamically, responding quickly…

コンテキスト
400K
入力
$1.75
出力
$14
ファイル画像テキスト
モデルを見る
利用可能

GPT-5.2 Chat (AKA Instant) is the fast, lightweight member of the 5.2 family, optimized for low-latency chat while retaining strong general intelligence. It uses adaptive reasoning to selectively “think” on…

コンテキスト
128K
入力
$1.75
出力
$14
ファイル画像テキスト
モデルを見る
利用可能

GPT-5.2 Pro is OpenAI’s most advanced model, offering major improvements in agentic coding and long context performance over GPT-5 Pro. It is optimized for complex tasks that require step-by-step reasoning,…

コンテキスト
400K
入力
$21
出力
$168
画像テキストファイル
モデルを見る
利用可能

GPT-5.2-Codex is an upgraded version of GPT-5.1-Codex optimized for software engineering and coding workflows. It is designed for both interactive development sessions and long, independent execution of complex engineering tasks.…

コンテキスト
400K
入力
$1.75
出力
$14
テキスト画像
モデルを見る
利用可能

GPT-5.3-Codex is OpenAI’s most advanced agentic coding model, combining the frontier software engineering performance of GPT-5.2-Codex with the broader reasoning and professional knowledge capabilities of GPT-5.2. It achieves state-of-the-art results…

コンテキスト
400K
入力
$1.75
出力
$14
テキスト画像ファイル
モデルを見る

OpenAI

GPT-5.4

利用可能

GPT-5.4 is OpenAI’s latest frontier model, unifying the Codex and GPT lines into a single system. It features a 1M+ token context window (922K input, 128K output) with support for…

コンテキスト
1.05M
入力
$2.5
出力
$15
テキスト画像ファイル
モデルを見る
利用可能

GPT-5.4 mini brings the core capabilities of GPT-5.4 to a faster, more efficient model optimized for high-throughput workloads. It supports text and image inputs with strong performance across reasoning, coding,…

コンテキスト
400K
入力
$0.75
出力
$4.5
ファイル画像テキスト
モデルを見る
利用可能

GPT-5.4 nano is the most lightweight and cost-efficient variant of the GPT-5.4 family, optimized for speed-critical and high-volume tasks. It supports text and image inputs and is designed for low-latency…

コンテキスト
400K
入力
$0.20
出力
$1.25
ファイル画像テキスト
モデルを見る
利用可能

GPT-5.4 Pro is OpenAI's most advanced model, building on GPT-5.4's unified architecture with enhanced reasoning capabilities for complex, high-stakes tasks. It features a 1M+ token context window (922K input, 128K…

コンテキスト
1.05M
入力
$30
出力
$180
テキスト画像ファイル
モデルを見る

OpenAI

GPT-5.5

利用可能

GPT-5.5 is OpenAI’s frontier model designed for complex professional workloads, building on GPT-5.4 with stronger reasoning, higher reliability, and improved token efficiency on hard tasks. It features a 1M+ token…

コンテキスト
1.05M
入力
$5
出力
$30
ファイル画像テキスト
モデルを見る
利用可能

GPT-5.5 Pro is OpenAI’s high-capability model optimized for deep reasoning and accuracy on complex, high-stakes workloads. It features a 1M+ token context window (922K input, 128K output) with support for…

コンテキスト
1.05M
入力
$30
出力
$180
ファイル画像テキスト
モデルを見る
利用可能

GPT-5.6 Luna Pro is the same underlying model as GPT-5.6 Luna, served with reasoning.mode set to pro for higher-quality responses on complex tasks.

コンテキスト
1.05M
入力
$0.20
出力
$1.2
ファイル画像テキスト
モデルを見る
利用可能

GPT-5.6 Sol Pro is the same underlying model as GPT-5.6 Sol, served with reasoning.mode set to pro for higher-quality responses on complex tasks.

コンテキスト
1.05M
入力
$2.5
出力
$15
ファイル画像テキスト
モデルを見る
利用可能

GPT-5.6 Terra is a balanced model in OpenAI's GPT-5.6 series, positioned between the flagship Sol tier and the cost-efficient Luna tier. It is suited for everyday coding, reasoning, and agentic…

コンテキスト
1.05M
入力
$2
出力
$12
ファイル画像テキスト
モデルを見る
利用可能

GPT-5.6 Terra Pro is the same underlying model as GPT-5.6 Terra, served with reasoning.mode set to pro for higher-quality responses on complex tasks.

コンテキスト
1.05M
入力
$2
出力
$12
ファイル画像テキスト
モデルを見る
利用可能

gpt-oss-safeguard-120b and gpt-oss-safeguard-20b are safety reasoning models built-upon gpt-oss. With these models, you can classify text content based on safety policies that you provide and perform a suite of foundational safety tasks. These models are intended for safety use cases. For other applications, we recommend using gpt-oss models.

コンテキスト
131K
入力
$0.075
出力
$0.30
テキスト
モデルを見る
利用可能

📣 **Update [10-07-2025]:** Added a *default system prompt* to the chat template to guide the model towards more *professional, accurate, and safe* responses.

コンテキスト
131K
入力
$0.017
出力
$0.112
テキスト
モデルを見る
利用可能

**Model Summary:** Granite-4.1-8B is a 8B parameter long-context instruct model finetuned from *Granite-4.1-8B-Base* using a combination of open source instruction datasets with permissive license and internally collected synthetic datasets. Granite 4.1 models have gone through an improved post-training pipeline, including supervised finetuning and reinforcement learning alignment, resulting in enhanced tool calling, instruction following, and chat capabilities.

コンテキスト
131K
入力
$0.05
出力
$0.10
テキスト
モデルを見る

SpaceXAI

Grok 4.20

利用可能

Grok 4.20 is a reasoning model from SpaceXAI with industry-leading speed and agentic tool calling capabilities. It combines the lowest hallucination rate on the market with strict prompt adherance, delivering…

コンテキスト
2M
入力
$1.25
出力
$2.5
テキスト画像ファイル
モデルを見る
利用可能

Grok 4.20 Multi-Agent is a variant of SpaceXAI’s Grok 4.20 designed for collaborative, agent-based workflows. Multiple agents operate in parallel to conduct deep research, coordinate tool use, and synthesize information…

コンテキスト
2M
入力
$1.25
出力
$2.5
テキスト画像ファイル
モデルを見る

SpaceXAI

Grok 4.3

利用可能

Grok 4.3 is a reasoning model from SpaceXAI. It accepts text and image inputs with text output, and is suited for agentic workflows, instruction-following tasks, and applications requiring high factual…

コンテキスト
1M
入力
$1.25
出力
$2.5
テキスト画像ファイル
モデルを見る

SpaceXAI

Grok 4.5

利用可能

Grok 4.5 is a model from SpaceXAI with frontier performance on coding, knowledge work, and STEM.

コンテキスト
500K
入力
$2
出力
$6
テキスト画像ファイル
モデルを見る
利用可能

Grok Build 0.1 is SpaceXAI’s fast coding model trained specifically for agentic software engineering workflows. It supports text and image inputs with text output, and is optimized for interactive coding…

コンテキスト
256K
入力
$1
出力
$2
テキスト画像ファイル
モデルを見る
利用可能

This model always redirects to the latest Grok model from xAI.

コンテキスト
500K
入力
$2
出力
$6
テキスト画像ファイル
モデルを見る
利用可能

Hermes 4 405B is a frontier, hybrid-mode **reasoning** model based on Llama-3.1-405B by Nous Research that is aligned to **you**.

コンテキスト
131K
入力
$1
出力
$3
テキスト
モデルを見る
利用可能

Hermes 4 70B is a frontier, hybrid-mode **reasoning** model based on Llama-3.1-70B by Nous Research that is aligned to **you**.

コンテキスト
131K
入力
$0.13
出力
$0.40
テキスト
モデルを見る

Tencent

Hy3

利用可能

**Hy3** is a 295B-parameter Mixture-of-Experts (MoE) model with 21B active parameters and 3.8B MTP layer parameters, developed by the Tencent Hy Team. Following the Hy3 Preview launch in late April, we gathered feedback from 50+ products and scaled up post-training with higher quality data. Today, we introduce Hy3, which outperforms similar-size models and rivals flagship open-source models with 2-5x parameters. It also shows significant gains in utility across various products and productivity tasks.

コンテキスト
262K
入力
$0.132
出力
$0.528
テキスト
モデルを見る
利用可能

**Hy3 preview** is a 295B-parameter Mixture-of-Experts (MoE) model with 21B active parameters and 3.8B MTP layer parameters, developed by the Tencent Hy Team. Hy3 preview is the first model trained on our rebuilt infrastructure, and the strongest we've shipped so far. It improves significantly on complex reasoning, instruction following, context learning, coding, and agent tasks.

コンテキスト
262K
入力
$0.18
出力
$0.60
テキスト
モデルを見る

Thinking Machines

Inkling

利用可能

Inkling is a general-purpose multimodal model that accepts text, image and audio inputs and generates text outputs. It is intended for use in English and other languages, and across multiple coding languages. The model is designed to be used by developers building AI-powered applications, including agentic and tool-use systems, coding assistants, chatbots, and retrieval-augmented generation systems, and is suitable for general-purpose conversational use, instruction-following, and other natural language and multimodal tasks. It is released with open weights to support research, fine-tuning and integration into third-party products by downstream developers.

コンテキスト
524K
入力
$0.95
出力
$4.05
テキスト画像音声
モデルを見る

Thinking Machines

Inkling Small

利用可能

Inkling-Small is a general-purpose multimodal model that accepts text, image and audio inputs and generates text outputs. It is intended for use in English and other languages, and across multiple coding languages. The model is designed to be used by developers building AI-powered applications, including agentic and tool-use systems, coding assistants, chatbots, and retrieval-augmented generation systems, and is suitable for general-purpose conversational use, instruction-following, and other natural language and multimodal tasks. It is released with open weights to support research, fine-tuning and integration into third-party products by downstream developers.

コンテキスト
524K
入力
$0.45
出力
$1.2
テキスト画像音声
モデルを見る
利用可能

KAT-Coder-Air V2.5 is a flagship-level Agentic Coding model that can directly hand over an entire issue or an entire business workflow to it, allowing it to autonomously locate and make…

コンテキスト
256K
入力
$0.15
出力
$0.60
テキスト
モデルを見る
利用可能

KAT-Coder-Pro V2 is the latest high-performance model in KwaiKAT’s KAT-Coder series, designed for complex enterprise-grade software engineering and SaaS integration. It builds on the agentic coding strengths of earlier versions,…

コンテキスト
256K
入力
$0.30
出力
$1.2
テキスト
モデルを見る
利用可能

KAT-Coder-Pro V2.5 is a flagship-level Agentic Coding model that can directly hand over an entire issue or an entire business workflow to it, allowing it to autonomously locate and make…

コンテキスト
256K
入力
$0.74
出力
$2.96
テキスト
モデルを見る

MoonshotAI

Kimi K2 0905

利用可能

📰   Tech Blog     |     📄   Paper

コンテキスト
262K
入力
$0.60
出力
$2.5
テキスト
モデルを見る
利用可能

Kimi K2 Thinking is the latest, most capable version of open-source thinking model. Starting with Kimi K2, we built it as a thinking agent that reasons step-by-step while dynamically invoking tools. It sets a new state-of-the-art on Humanity's Last Exam (HLE), BrowseComp, and other benchmarks by dramatically scaling multi-step reasoning depth and maintaining stable tool-use across 200–300 sequential calls. At the same time, K2 Thinking is a native INT4 quantization model with 256k context window, achieving lossless reductions in inference latency and GPU memory usage.

コンテキスト
262K
入力
$0.60
出力
$2.5
テキスト
モデルを見る

MoonshotAI

Kimi K2.5

利用可能

📰   Tech Blog     |     📄   Paper

コンテキスト
262K
入力
$0.57
出力
$2.85
テキスト画像
モデルを見る

MoonshotAI

Kimi K2.6

利用可能

Kimi K2.6 is an open-source, native multimodal agentic model that advances practical capabilities in long-horizon coding, coding-driven design, proactive autonomous execution, and swarm-based task orchestration.

コンテキスト
262K
入力
$0.95
出力
$4
テキスト画像
モデルを見る

MoonshotAI

Kimi K2.7 Code

利用可能

Kimi K2.7 Code is a coding-focused agentic model built upon Kimi K2.6. With substantial improvements on real-world long-horizon coding tasks, it strengthens end-to-end task completion across complex software engineering workflows while improving token efficiency, reducing thinking-token usage by approximately 30% compared with Kimi K2.6.

コンテキスト
262K
入力
$0.71
出力
$3.5
テキスト画像
モデルを見る

MoonshotAI

Kimi K3

利用可能

Kimi K3 is an open-weight, native multimodal agentic model and our most capable model to date. It is a 2.8T-parameter model built on Kimi Delta Attention (KDA) and Attention Residuals (AttnRes), with native vision capabilities and a 1-million-token context window. It is the world's first open 3T-class model, designed for frontier intelligence across long-horizon coding, knowledge work, and reasoning.

コンテキスト
1.05M
入力
$3
出力
$15
テキスト画像動画
モデルを見る

Poolside

Laguna S 2.1

利用可能

Laguna S 2.1 is a 118B total parameter Mixture-of-Experts model with 8B activated parameters per token, designed for agentic coding and long-horizon work. It sits between Laguna XS 2.1 (33B-A3B) and Laguna M.1 (225B-A23B) in the Laguna series and shares the family recipe: a token-choice router with softplus gating over 256 routed experts plus one shared expert, grouped-query attention, and interleaved full/sliding-window attention.

コンテキスト
1.05M
入力
$0.09
出力
$0.18
テキスト
モデルを見る
利用可能

Laguna XS 2.1 is a 33B total parameter Mixture-of-Experts model with 3B activated parameters per token designed for agentic coding and long-horizon work on a local machine. This model is an upgraded version of our Laguna XS.2 model with a +5.4% jump on SWE-bench Multilingual as well as stronger performance on terminal-style tasks.

コンテキスト
262K
入力
$0.06
出力
$0.12
テキスト
モデルを見る
利用可能

LFM2.5-2.6B is a compact reasoning model from Liquid AI. It is suited for agent workflows, data extraction, RAG, and long-context processing. Liquid advises against using it for agentic coding or…

コンテキスト
128K
入力
$0.00
出力
$0.00
テキスト
モデルを見る

inclusionAI

Ling-2.6-1T

利用可能

Ling-2.6-1T is an instant (instruct) model from inclusionAI and the company’s trillion-parameter flagship, designed for real-world agents that require fast execution and high efficiency at scale. It uses a “fast…

コンテキスト
262K
入力
$0.075
出力
$0.625
テキスト
モデルを見る

inclusionAI

Ling-2.6-flash

利用可能

Ling-2.6-flash is an instant (instruct) model from inclusionAI with 104B total parameters and 7.4B active parameters, designed for real-world agents that require fast responses, strong execution, and high token efficiency.…

コンテキスト
262K
入力
$0.01
出力
$0.03
テキスト
モデルを見る
利用可能

🤗 Hugging Face    |   🤖 ModelScope    |   🐙 the provider catalog   

コンテキスト
262K
入力
$0.021
出力
$0.063
テキスト
モデルを見る
利用可能

LongCat 2.0 is a sparse mixture-of-experts language model from Meituan, with 48B active parameters out of 1.6T total. It is suited for coding, repository-level changes, long-horizon problem solving, and agentic…

コンテキスト
1.05M
入力
$0.30
出力
$1.2
テキスト
モデルを見る
利用可能

30 second duration clips are priced at $0.04 per clip. Lyria 3 is Google's family of music generation models, available through the Gemini API. With Lyria 3, you can generate…

コンテキスト
1.05M
入力
$0.00
出力
$0.00
テキスト画像
モデルを見る
利用可能

Full-length songs are priced at $0.08 per song. Lyria 3 is Google's family of music generation models, available through the Gemini API. With Lyria 3, you can generate high-quality, 48kHz…

コンテキスト
1.05M
入力
$0.00
出力
$0.00
テキスト画像
モデルを見る

Inception

Mercury 2

利用可能

Mercury 2 is an extremely fast reasoning LLM, and the first reasoning diffusion LLM (dLLM). Instead of generating tokens sequentially, Mercury 2 produces and refines multiple tokens in parallel, achieving…

コンテキスト
128K
入力
$0.25
出力
$0.75
テキスト
モデルを見る

Xiaomi

MiMo-V2.5

利用可能

🤗 HuggingFace  | 📰 Blog  | 🎨 Xiaomi MiMo API Platform  | 🗨️ Xiaomi MiMo Studio  |

コンテキスト
1.05M
入力
$0.14
出力
$0.28
テキスト音声画像動画
モデルを見る
利用可能

🤗 HuggingFace  | 📰 Blog  | 🎨 Xiaomi MiMo API Platform  | 🗨️ Xiaomi MiMo Studio  |

コンテキスト
1.05M
入力
$0.435
出力
$0.87
テキスト
モデルを見る

MiniMax

MiniMax M2

利用可能

Today, we release and open source MiniMax-M2, a **Mini** model built for **Max** coding & agentic workflows.

コンテキスト
205K
入力
$0.255
出力
$1.02
テキスト
モデルを見る
利用可能

Today, we are handing **MiniMax-M2.1** over to the open-source community. This release is more than just a parameter update; it is a significant step toward democratizing top-tier agentic capabilities.

コンテキスト
205K
入力
$0.30
出力
$1.2
テキスト
モデルを見る
利用可能

Extensively trained with reinforcement learning in hundreds of thousands of complex real-world environments, M2.5 is **SOTA in coding, agentic tool use and search, office work, and a range of other economically valuable tasks**, boasting scores of **80.2% in SWE-Bench Verified, 51.3% in Multi-SWE-Bench, and 76.3% in BrowseComp** (with context management).

コンテキスト
197K
入力
$0.22
出力
$0.90
テキスト
モデルを見る
利用可能

**MiniMax-M2.7** is our first model deeply participating in its own evolution. M2.7 is capable of building complex agent harnesses and completing highly elaborate productivity tasks, leveraging Agent Teams, complex Skills, and dynamic tool search. For more details, see our blog post.

コンテキスト
205K
入力
$0.30
出力
$1.2
テキスト
モデルを見る

MiniMax

MiniMax M3

利用可能

MiniMax-M3 is a native multimodal model with 1M context. It has ~428B parameters and ~23B activated parameters.

コンテキスト
524K
入力
$0.30
出力
$1.2
テキスト画像動画
モデルを見る
利用可能

The largest model in the Ministral 3 family, **Ministral 3 14B** offers frontier capabilities and performance comparable to its larger Mistral Small 3.2 24B counterpart. A powerful and efficient language model with vision capabilities.

コンテキスト
262K
入力
$0.20
出力
$0.20
テキスト画像
モデルを見る
利用可能

The smallest model in the Ministral 3 family, **Ministral 3 3B** is a powerful, efficient tiny language model with vision capabilities.

コンテキスト
131K
入力
$0.10
出力
$0.10
テキスト画像
モデルを見る
利用可能

A balanced model in the Ministral 3 family, **Ministral 3 8B** is a powerful, efficient tiny language model with vision capabilities.

コンテキスト
262K
入力
$0.15
出力
$0.15
テキスト画像
モデルを見る
利用可能

Mistral Large 3 2512 is Mistral’s most capable model to date, featuring a sparse mixture-of-experts architecture with 41B active parameters (675B total), and released under the Apache 2.0 license.

コンテキスト
262K
入力
$0.50
出力
$1.5
テキスト画像ファイル
モデルを見る
利用可能

Mistral Small 4 is a powerful hybrid model capable of acting as both a general instruction model and a reasoning model. It unifies the capabilities of three different model families—**Instruct**, **Reasoning** (previously called Magistral), and **Devstral**—into a single, unified model.

コンテキスト
262K
入力
$0.15
出力
$0.60
テキスト画像
モデルを見る
利用可能

This model always redirects to the latest model in the MoonshotAI Kimi family.

コンテキスト
975K
入力
$2.6
出力
$13
テキスト画像動画
モデルを見る
利用可能

Muse Glimmer 30B is a dense, open-weight multimodal model from Meta Superintelligence Labs, distilled from Muse Spark and optimized for autonomous agents on consumer hardware. It is suited for long-horizon…

コンテキスト
131K
入力
$0.35
出力
$1.5
テキスト画像
モデルを見る
利用可能

Muse Spark 1.1 is a multimodal reasoning model from Meta, built for agentic tasks. It accepts text, images, video, audio, and PDF documents and returns text, with a 1M-token context…

コンテキスト
1.05M
入力
$1.25
出力
$4.25
テキスト画像動画ファイル
モデルを見る
利用可能

Muse Spark 1.2 is a reasoning model from Meta, designed for complex agentic tasks. It accepts text, images, video, audio, and PDF documents, returns text, and offers a 1M-token context…

コンテキスト
1.05M
入力
$1.25
出力
$4.25
テキスト画像動画ファイル
モデルを見る
利用可能

Nemotron-3-Nano-30B-A3B-BF16 is a large language model (LLM) trained from scratch by NVIDIA, and designed as a unified model for both reasoning and non-reasoning tasks. It responds to user queries and tasks by first generating a reasoning trace and then concluding with a final response. The model's reasoning capabilities can be configured through a flag in the chat template. If the user prefers the model to provide its final answer without intermediate reasoning traces, it can be configured to do so, albeit with a slight decrease in accuracy for harder prompts that require reasoning. Conversely, allowing the model to generate reasoning traces first generally results in higher-quality final solutions to queries and tasks.

コンテキスト
262K
入力
$0.05
出力
$0.20
テキスト
モデルを見る
利用可能

NVIDIA Nemotron 3 Nano Omni is a multimodal large language model that unifies video, audio, image, and text understanding to support enterprise-grade Q&A, summarization, transcription, and document intelligence workflows. It extends the Nemotron Nano family with integrated video+speech comprehension, Graphical User Interface (GUI), Optical Character Recognition (OCR), and speech transcription capabilities, enabling end-to-end processing of rich enterprise content such as meeting recordings, M&E assets, training videos, and complex business documents. NVIDIA Nemotron 3 Nano Omni was developed by NVIDIA as part of the Nemotron model family.

コンテキスト
256K
入力
$0.00
出力
$0.00
テキスト音声画像動画
モデルを見る
利用可能

> Use temperature=1.0 and top_p=0.95 across **all tasks and serving backends** — reasoning, tool calling, and general chat alike.

コンテキスト
262K
入力
$0.085
出力
$0.40
テキスト
モデルを見る
利用可能

For more details on how to deploy and use the model - see the Quick Start Guide below!

コンテキスト
512K
入力
$0.60
出力
$3.6
テキスト
モデルを見る
利用可能

NVIDIA Nemotron™ is a family of open models with open weights, training data, and recipes, delivering leading efficiency and accuracy for building specialized AI agents.

コンテキスト
262K
入力
$0.08
出力
$0.20
テキスト
モデルを見る
利用可能

NVIDIA-Nemotron-Nano-9B-v2 is a large language model (LLM) trained from scratch by NVIDIA, and designed as a unified model for both reasoning and non-reasoning tasks. It responds to user queries and tasks by first generating a reasoning trace and then concluding with a final response. The model's reasoning capabilities can be controlled via a system prompt. If the user prefers the model to provide its final answer without intermediate reasoning traces, it can be configured to do so, albeit with a slight decrease in accuracy for harder prompts that require reasoning. Conversely, allowing the model to generate reasoning traces first generally results in higher-quality final solutions to queries and tasks.

コンテキスト
128K
入力
$0.00
出力
$0.00
テキスト
モデルを見る
利用可能

🤗 Model &nbsp&nbsp | &nbsp&nbsp 🔀 the provider catalog (Enjoy two weeks free starting June 9!) &nbsp&nbsp | &nbsp&nbsp 💻 Github &nbsp&nbsp | &nbsp&nbsp 🧭 ModelScope &nbsp&nbsp | &nbsp&nbsp 🚀 Nex-AGI

コンテキスト
262K
入力
$0.025
出力
$0.10
テキスト画像
モデルを見る

Nex AGI

Nex-N2-Pro

利用可能

🤗 Model &nbsp&nbsp | &nbsp&nbsp 🔀 the provider catalog (Enjoy two weeks free starting June 9!) &nbsp&nbsp | &nbsp&nbsp 💻 Github &nbsp&nbsp | &nbsp&nbsp 🧭 ModelScope &nbsp&nbsp | &nbsp&nbsp 🚀 Nex-AGI

コンテキスト
262K
入力
$0.25
出力
$1
テキスト画像
モデルを見る
利用可能

North Mini Code is Cohere's first agentic coding model and the debut of its North family. A sparse mixture-of-experts model with 30B total parameters and 3B active, it is optimized…

コンテキスト
256K
入力
$0.00
出力
$0.00
テキスト
モデルを見る
利用可能

Nova 2 Lite is a fast, cost-effective reasoning model for everyday workloads that can process text, images, and videos to generate text. Nova 2 Lite demonstrates standout capabilities in processing…

コンテキスト
1M
入力
$0.30
出力
$2.5
テキスト画像動画ファイル
モデルを見る
利用可能

Amazon Nova Premier is the most capable of Amazon’s multimodal models for complex reasoning tasks and for use as the best teacher for distilling custom models.

コンテキスト
1M
入力
$2.5
出力
$12.5
テキスト画像
モデルを見る
利用可能

This model always redirects to the latest model in the OpenAI GPT family.

コンテキスト
1.05M
入力
$2.5
出力
$15
ファイル画像テキスト
モデルを見る
利用可能

This model always redirects to the latest model in the OpenAI GPT Mini family.

コンテキスト
400K
入力
$0.75
出力
$4.5
ファイル画像テキスト
モデルを見る
利用可能

Palmyra X5 is Writer's most advanced model, purpose-built for building and scaling AI agents across the enterprise. It delivers industry-leading speed and efficiency on context windows up to 1 million…

コンテキスト
1.04M
入力
$0.60
出力
$6
テキスト
モデルを見る
利用可能

The Pareto Router maintains a tiered shortlist of strong coding models, ranked by Artificial Analysis coding percentiles. Set min_coding_score between 0 and 1 on the pareto-router plugin to control how…

コンテキスト
2M
入力
不明
出力
不明
テキスト
モデルを見る
利用可能

Qwen Plus 0728, based on the Qwen3 foundation model, is a 1 million context hybrid reasoning model with a balanced performance, speed, and cost combination.

コンテキスト
1M
入力
$0.26
出力
$0.78
テキスト
モデルを見る
利用可能

Qwen3 Coder Flash is Alibaba's fast and cost efficient version of their proprietary Qwen3 Coder Plus. It is a powerful coding agent model specializing in autonomous programming via tool calling…

コンテキスト
1M
入力
$0.195
出力
$0.975
テキスト
モデルを見る
利用可能

Qwen3 Coder Plus is Alibaba's proprietary version of the Open Source Qwen3 Coder 480B A35B. It is a powerful coding agent model specializing in autonomous programming via tool calling and…

コンテキスト
1M
入力
$0.65
出力
$3.25
テキスト
モデルを見る
利用可能

Qwen3-Max is an updated release built on the Qwen3 series, offering major improvements in reasoning, instruction following, multilingual support, and long-tail knowledge coverage compared to the January 2025 version. It…

コンテキスト
262K
入力
$0.78
出力
$3.9
テキスト
モデルを見る
利用可能

Qwen3-Max-Thinking is the flagship reasoning model in the Qwen3 series, designed for high-stakes cognitive tasks that require deep, multi-step reasoning. By significantly scaling model capacity and reinforcement learning compute, it…

コンテキスト
262K
入力
$0.78
出力
$3.9
テキスト
モデルを見る
利用可能

Over the past few months, we have observed increasingly clear trends toward scaling both total parameters and context lengths in the pursuit of more powerful and agentic artificial intelligence (AI). We are excited to share our latest advancements in addressing these demands, centered on improving scaling efficiency through innovative model architecture. We call this next-generation foundation models **Qwen3-Next**.

コンテキスト
262K
入力
$0.10
出力
$1.1
テキスト
モデルを見る
利用可能

Over the past few months, we have observed increasingly clear trends toward scaling both total parameters and context lengths in the pursuit of more powerful and agentic artificial intelligence (AI). We are excited to share our latest advancements in addressing these demands, centered on improving scaling efficiency through innovative model architecture. We call this next-generation foundation models **Qwen3-Next**.

コンテキスト
131K
入力
$0.15
出力
$1.2
テキスト
モデルを見る
利用可能

Meet Qwen3-VL — the most powerful vision-language model in the Qwen series to date.

コンテキスト
131K
入力
$0.13
出力
$0.52
テキスト画像
モデルを見る
利用可能

Meet Qwen3-VL — the most powerful vision-language model in the Qwen series to date.

コンテキスト
131K
入力
$0.20
出力
$2.4
テキスト画像
モデルを見る
利用可能

Meet Qwen3-VL — the most powerful vision-language model in the Qwen series to date.

コンテキスト
131K
入力
$0.104
出力
$0.416
テキスト画像
モデルを見る
利用可能

Meet Qwen3-VL — the most powerful vision-language model in the Qwen series to date.

コンテキスト
131K
入力
$0.117
出力
$0.455
画像テキスト
モデルを見る
利用可能

Meet Qwen3-VL — the most powerful vision-language model in the Qwen series to date.

コンテキスト
131K
入力
$0.18
出力
$2.1
画像テキスト
モデルを見る
利用可能

> [!Note] > This repository contains model weights and configuration files for the post-trained model in the Hugging Face Transformers format. > > These artifacts are compatible with Hugging Face Transformers, vLLM, SGLang, KTransformers, etc.

コンテキスト
262K
入力
$0.39
出力
$2.34
テキスト画像動画
モデルを見る
利用可能

The Qwen3.5 native vision-language series Plus models are built on a hybrid architecture that integrates linear attention mechanisms with sparse mixture-of-experts models, achieving higher inference efficiency. In a variety of…

コンテキスト
1M
入力
$0.26
出力
$1.56
テキスト画像動画
モデルを見る
利用可能

Qwen3.5 Plus (April 2026) is a large-scale multimodal language model from Alibaba. It accepts text, image, and video input and produces text output, with a 1M token context window. This…

コンテキスト
1M
入力
$0.30
出力
$1.8
テキスト画像動画
モデルを見る
利用可能

> [!Note] > This repository contains model weights and configuration files for the post-trained model in the Hugging Face Transformers format. > > These artifacts are compatible with Hugging Face Transformers, vLLM, SGLang, KTransformers, etc.

コンテキスト
262K
入力
$0.26
出力
$2.08
テキスト画像動画
モデルを見る
利用可能

> [!Note] > This repository contains model weights and configuration files for the post-trained model in the Hugging Face Transformers format. > > These artifacts are compatible with Hugging Face Transformers, vLLM, SGLang, KTransformers, etc.

コンテキスト
262K
入力
$0.195
出力
$1.56
テキスト画像動画
モデルを見る
利用可能

> [!Note] > This repository contains model weights and configuration files for the post-trained model in the Hugging Face Transformers format. > > These artifacts are compatible with Hugging Face Transformers, vLLM, SGLang, KTransformers, etc.

コンテキスト
262K
入力
$0.225
出力
$1.8
テキスト画像動画
モデルを見る
利用可能

> [!Note] > This repository contains model weights and configuration files for the post-trained model in the Hugging Face Transformers format. > > These artifacts are compatible with Hugging Face Transformers, vLLM, SGLang, KTransformers, etc.

コンテキスト
262K
入力
$0.10
出力
$0.15
テキスト画像動画
モデルを見る
利用可能

The Qwen3.5 native vision-language Flash models are built on a hybrid architecture that integrates a linear attention mechanism with a sparse mixture-of-experts model, achieving higher inference efficiency. Compared to the…

コンテキスト
1M
入力
$0.065
出力
$0.26
テキスト画像動画
モデルを見る
利用可能

> [!Note] > This repository contains model weights and configuration files for the post-trained model in the Hugging Face Transformers format. > > These artifacts are compatible with Hugging Face Transformers, vLLM, SGLang, KTransformers, etc.

コンテキスト
262K
入力
$0.60
出力
$3.6
テキスト画像動画
モデルを見る
利用可能

> [!Note] > This repository contains model weights and configuration files for the post-trained model in the Hugging Face Transformers format. > > These artifacts are compatible with Hugging Face Transformers, vLLM, SGLang, KTransformers, etc.

コンテキスト
262K
入力
$0.14
出力
$1
テキスト画像動画
モデルを見る
利用可能

Qwen3.6 Flash is a fast, efficient language model from Alibaba's Qwen 3.6 series. It supports text, image, and video input with a 1M token context window. Tiered pricing kicks in…

コンテキスト
1M
入力
$0.1875
出力
$1.13
テキスト画像動画
モデルを見る
利用可能

Qwen3.6-Max-Preview is a proprietary frontier model from Alibaba Cloud built on a sparse mixture-of-experts architecture with approximately 1 trillion total parameters. It is optimized for agentic coding, tool use, and…

コンテキスト
262K
入力
$1.03
出力
$6.16
テキスト
モデルを見る
利用可能

Qwen 3.6 Plus builds on a hybrid architecture that combines efficient linear attention with sparse mixture-of-experts routing, enabling strong scalability and high-performance inference. Compared to the 3.5 series, it delivers…

コンテキスト
1M
入力
$0.325
出力
$1.95
テキスト画像動画
モデルを見る
利用可能

Qwen3.7 Flash is a vision-language reasoning model from Alibaba. It is suited for multimodal agents, visual coding, search, and computer interaction, with strengths in object recognition, spatial understanding, and real-world…

コンテキスト
1M
入力
$0.03
出力
$0.13
テキスト画像動画
モデルを見る
利用可能

Qwen3.7-Max is the flagship model in Alibaba's Qwen3.7 series. It supports text input and output and is designed for agent-centric workloads, with particular strengths in coding, office and productivity tasks,…

コンテキスト
1M
入力
$1.48
出力
$4.43
テキスト
モデルを見る
利用可能

Qwen3.7-Plus is a cost-effective model in Alibaba's Qwen3.7 series. It supports text and image input with text output, building on the series' text capabilities with a comprehensive upgrade to its…

コンテキスト
1M
入力
$0.32
出力
$1.28
テキスト画像
モデルを見る
利用可能

> [!Note] > This repository contains model weights and configuration files for the post-trained model in the Hugging Face Transformers format. > > These artifacts are compatible with vLLM, SGLang, TokenSpeed, etc.

コンテキスト
1M
入力
$2
出力
$6
テキスト
モデルを見る
利用可能

Qwen3.8 Max is the flagship model in Alibaba's Qwen3.8 series, the general-availability successor to the Qwen3.8 Max Preview. It is a multimodal reasoning model intended for complex reasoning, visual understanding,…

コンテキスト
1M
入力
$2
出力
$6
テキスト画像動画
モデルを見る
利用可能

Relace Apply 3 is a specialized code-patching LLM that merges AI-suggested edits straight into your source files. It can apply updates from GPT-4o, Claude, and others into your files at…

コンテキスト
256K
入力
$0.85
出力
$1.25
テキスト
モデルを見る
利用可能

The relace-search model uses 4-12 view_file and grep tools in parallel to explore a codebase and return relevant files to the user request. In contrast to RAG, relace-search performs agentic…

コンテキスト
256K
入力
$1
出力
$3
テキスト
モデルを見る

inclusionAI

Ring-2.6-1T

利用可能

Ring-2.6-1T is a 1T-parameter-scale thinking model with 63B active parameters, built for real-world agent workflows that require both strong capability and operational efficiency. It is optimized for coding agents, tool…

コンテキスト
262K
入力
$0.075
出力
$0.625
テキスト
モデルを見る
利用可能

Sakana Namazu is a Japanese-specialized reasoning model from Sakana AI, based on Kimi K2.6 with additional training for Japanese language and business contexts. It is suited for Japanese instruction following,…

コンテキスト
262K
入力
$0.95
出力
$4
テキスト画像ファイル
モデルを見る

ByteDance Seed

Seed 1.6

利用可能

Seed 1.6 is a general-purpose model released by the ByteDance Seed team. It incorporates multimodal capabilities and adaptive deep thinking with a 256K context window.

コンテキスト
262K
入力
$0.25
出力
$2
画像テキスト動画
モデルを見る

ByteDance Seed

Seed 1.6 Flash

利用可能

Seed 1.6 Flash is an ultra-fast multimodal deep thinking model by ByteDance Seed, supporting both text and visual understanding. It features a 256k context window and can generate outputs of…

コンテキスト
262K
入力
$0.075
出力
$0.30
画像テキスト動画
モデルを見る

ByteDance Seed

Seed 2.1 Turbo

利用可能

Seed 2.1 Turbo is a multimodal model from ByteDance Seed for coding and long-horizon agent workflows. It is suited for end-to-end software delivery, multi-step task execution, and understanding visual and…

コンテキスト
262K
入力
$0.50
出力
$2.5
テキスト画像動画
モデルを見る

ByteDance Seed

Seed-2.0-Code

利用可能

Seed 2.0 Code is a model from ByteDance Seed optimized for agentic coding. It is suited for frontend development, multilingual programming tasks, and coding-agent workflows in tools such as Claude…

コンテキスト
262K
入力
$0.50
出力
$3
テキスト画像動画
モデルを見る

ByteDance Seed

Seed-2.0-Lite

利用可能

Seed-2.0-Lite is a versatile, cost‑efficient enterprise workhorse that delivers strong multimodal and agent capabilities while offering noticeably lower latency, making it a practical default choice for most production workloads across…

コンテキスト
262K
入力
$0.25
出力
$2
テキスト画像動画
モデルを見る

ByteDance Seed

Seed-2.0-Mini

利用可能

Seed-2.0-mini targets latency-sensitive, high-concurrency, and cost-sensitive scenarios, emphasizing fast response and flexible inference deployment. It delivers performance comparable to ByteDance-Seed-1.6, supports 256k context, four reasoning effort modes (minimal/low/medium/high), multimodal understanding,…

コンテキスト
262K
入力
$0.10
出力
$0.40
テキスト画像動画
モデルを見る
利用可能

Solar Pro 3 is Upstage's powerful Mixture-of-Experts (MoE) language model. With 102B total parameters and 12B active parameters per forward pass, it delivers exceptional performance while maintaining computational efficiency. Optimized…

コンテキスト
131K
入力
$0.15
出力
$0.60
テキスト
モデルを見る
利用可能

Solar Pro 4 is Upstage's cost-efficient large language model, featuring a 524K context window. It is built for long-horizon tasks and agentic workflows, with strong capabilities in office productivity, document-intensive…

コンテキスト
524K
入力
$0.03
出力
$0.12
テキスト
モデルを見る
利用可能

Exclusively available on the the provider catalog API, Sonar Pro's new Pro Search mode is Perplexity's most advanced agentic search system. It is designed for deeper reasoning and analysis. Pricing is based…

コンテキスト
200K
入力
$3
出力
$15
テキスト画像
モデルを見る
利用可能

Step 3.5 Flash is StepFun's most capable open-source foundation model. Built on a sparse Mixture of Experts (MoE) architecture, it selectively activates only 11B of its 196B parameters per token.…

コンテキスト
262K
入力
$0.10
出力
$0.30
テキスト
モデルを見る
利用可能

Step 3.7 Flash is StepFun's latest high-efficiency multimodal Mixture-of-Experts model. It pairs a 196B-parameter language backbone with a vision encoder for native image and video understanding, activating roughly 11B parameters…

コンテキスト
256K
入力
$0.20
出力
$1.15
テキスト画像動画
モデルを見る
利用可能

Trinity-Large-Thinking is a reasoning-optimized variant of Arcee AI's Trinity-Large family — a 398B-parameter sparse Mixture-of-Experts (MoE) model with approximately 13B active parameters per token. Built on Trinity-Large-Base and post-trained with extended chain-of-thought reasoning and agentic RL, Trinity-Large-Thinking delivers state-of-the-art performance on agentic benchmarks while maintaining strong general capabilities.

コンテキスト
262K
入力
$0.22
出力
$0.85
テキスト
モデルを見る
OUR METHOD

AIToollyのモデルデータについて

モデル固有の事実とプロバイダーエンドポイントの事実は分離して管理します。第三者カタログは検索とスナップショットに使い、モデル情報には公式文書と検証済みモデルカードを優先します。

01出典と確認
02確認日
03機能