AI MODEL INDEX

Przeglądaj i porównuj modele AI

Porównaj możliwości, okna kontekstu, ceny API i zastosowania wiodących modeli AI.

Źródła i weryfikacja Ostatnia weryfikacja
Szukaj modeli AISzukaj modeli AI...
Zastosuj filtry
BROWSE MODELS

Katalog modeli

352 modeli w tym widoku

Porównaj
Aktywny

A flagship GPT-5.6 model listed for complex reasoning, coding and multi-step agent workflows.

Kontekst
1,05M
Wejście
2,5 USD
Wyjście
15 USD
PlikObrazTekst
Zobacz model
Aktywny

A lower-cost GPT-5.6 model for high-volume chat, classification and lightweight agent workflows.

Kontekst
1,05M
Wejście
0,20 USD
Wyjście
1,2 USD
PlikObrazTekst
Zobacz model
Aktywny

A Sonnet-class model listed for coding, agents and professional knowledge work with adaptive reasoning.

Kontekst
1M
Wejście
2 USD
Wyjście
10 USD
TekstObrazPlik
Zobacz model
Aktywny

A fast multimodal Gemini model listed for responsive agent workflows, coding and multi-step reasoning.

Kontekst
1,05M
Wejście
0,375 USD
Wyjście
1,88 USD
TekstObrazWideoPlik
Zobacz model
Aktywny

A Grok model listed for coding, knowledge work and STEM tasks with text, image and file input.

Kontekst
500K
Wejście
2 USD
Wyjście
6 USD
TekstObrazPlik
Zobacz model

Alibaba / Qwen

Qwen3.8 27B

Aktywny

An open-weight vision-language model listed for coding, research, multimodal interaction and agent tasks.

Kontekst
262K
Wejście
0,45 USD
Wyjście
3,2 USD
TekstObrazWideo
Zobacz model
Aktywny

A text model listed as a large mixture-of-experts release with a 1,048,576-token provider context window.

Kontekst
1,05M
Wejście
1,32 USD
Wyjście
3,96 USD
Tekst
Zobacz model
Aktywny

A dense instruction model listed for agent workflows, coding and complex professional tasks.

Kontekst
262K
Wejście
1,5 USD
Wyjście
7,5 USD
TekstObrazPlik
Zobacz model

Alibaba / Qwen

Qwen3 Coder Next

Aktywny

An open-weight coding model listed for coding agents and local development workflows.

Kontekst
262K
Wejście
0,12 USD
Wyjście
0,80 USD
Tekst
Zobacz model
Aktywny

A multimodal endpoint listed for workflows that combine reasoning, text and image generation.

Kontekst
272K
Wejście
8 USD
Wyjście
15 USD
ObrazTekstPlik
Zobacz model

AionLabs

Aion-2.0

Aktywny

Aion-2.0 is a variant of DeepSeek V3.2 optimized for immersive roleplaying and storytelling. It is particularly strong at introducing tension, crises, and conflict into stories, making narratives feel more engaging.…

Kontekst
131K
Wejście
0,80 USD
Wyjście
1,6 USD
Tekst
Zobacz model

AionLabs

Aion-3.0

Aktywny

Aion-3.0 is a multi-model roleplaying and storytelling system from AionLabs, built on the GLM family of models. It uses a collaborative generation process in which multiple specialized models each contribute…

Kontekst
131K
Wejście
3 USD
Wyjście
6 USD
Tekst
Zobacz model
Aktywny

Aion-3.0 Mini is a multi-model roleplaying and storytelling system from AionLabs, built on the DeepSeek family of models. It uses a collaborative generation process in which multiple specialized models each…

Kontekst
131K
Wejście
0,70 USD
Wyjście
1,4 USD
Tekst
Zobacz model

Runway

Aleph 2.0

Aktywny

Runway Aleph 2.0 is an in-context video editing model from Runway. It applies text instructions and keyframe-guided edits across existing footage while preserving details that are not meant to change.…

Kontekst
0
Wejście
0,00 USD
Wyjście
0,00 USD
TekstObrazWideo
Zobacz model

Sentence Transformers

all-MiniLM-L12-v2

Aktywny

Sentence Transformers: all-MiniLM-L12-v2 is included in the AIToolly model catalog.

Kontekst
512
Wejście
0,005 USD
Wyjście
0,00 USD
Tekst
Zobacz model

Sentence Transformers

all-MiniLM-L6-v2

Aktywny

Sentence Transformers: all-MiniLM-L6-v2 is included in the AIToolly model catalog.

Kontekst
512
Wejście
0,005 USD
Wyjście
0,00 USD
Tekst
Zobacz model

Sentence Transformers

all-mpnet-base-v2

Aktywny

Sentence Transformers: all-mpnet-base-v2 is included in the AIToolly model catalog.

Kontekst
512
Wejście
0,005 USD
Wyjście
0,00 USD
Tekst
Zobacz model

This model always redirects to the latest model in the Anthropic Claude Haiku family.

Kontekst
200K
Wejście
1 USD
Wyjście
5 USD
TekstObrazPlik
Zobacz model

This model always redirects to the latest model in the Anthropic Claude Sonnet family.

Kontekst
1M
Wejście
2 USD
Wyjście
10 USD
TekstObrazPlik
Zobacz model

Deepgram

Aura-2

Aktywny

Aura-2 is a multilingual text-to-speech model from Deepgram. It supports Deepgram’s canonical Aura-2 voice catalog for speech synthesis across multiple languages.

Kontekst
0
Wejście
30 USD
Wyjście
0,00 USD
Tekst
Zobacz model
Aktywny

Auto Router (Beta) is a task-aware router from the provider catalog. It classifies each request, then routes it the most popular model for that task based on aggregate spend, filtered by your…

Kontekst
2M
Wejście
Nieznane
Wyjście
Nieznane
TekstObrazDźwiękPlik
Zobacz model
Aktywny

If you are looking for a model that supports more languages, longer texts, and other retrieval methods, you can try using bge-m3.

Kontekst
512
Wejście
0,005 USD
Wyjście
0,00 USD
Tekst
Zobacz model
Aktywny

If you are looking for a model that supports more languages, longer texts, and other retrieval methods, you can try using bge-m3.

Kontekst
512
Wejście
0,01 USD
Wyjście
0,00 USD
Tekst
Zobacz model

BAAI

bge-m3

Aktywny

In this project, we introduce BGE-M3, which is distinguished for its versatility in Multi-Functionality, Multi-Linguality, and Multi-Granularity.

Kontekst
8K
Wejście
0,01 USD
Wyjście
0,00 USD
Tekst
Zobacz model
Aktywny

Transform your natural language requests into structured the provider catalog API request objects. Describe what you want to accomplish with AI models, and Body Builder will construct the appropriate API calls. Example:…

Kontekst
128K
Wejście
Nieznane
Wyjście
Nieznane
Tekst
Zobacz model

Google

Chirp 3

Aktywny

Chirp 3 is Google's latest multilingual speech-to-text model. It offers enhanced transcription accuracy across 24 GA languages and 77+ preview languages, with support for automatic language detection, automatic punctuation, and…

Kontekst
0
Wejście
16 000 USD
Wyjście
0,00 USD
Dźwięk
Zobacz model
Aktywny

Claude Fable 5 is a Mythos-class model from Anthropic, built for autonomous knowledge work and coding. It supports text, image, and file inputs with text output, with reasoning support and…

Kontekst
1M
Wejście
10 USD
Wyjście
50 USD
TekstObrazPlik
Zobacz model
Aktywny

This model always redirects to the latest model in the Claude Fable family.

Kontekst
1M
Wejście
10 USD
Wyjście
50 USD
TekstObrazPlik
Zobacz model
Aktywny

Claude Haiku 4.5 is Anthropic’s fastest and most efficient model, delivering near-frontier intelligence at a fraction of the cost and latency of larger Claude models. Matching Claude Sonnet 4’s performance…

Kontekst
200K
Wejście
1 USD
Wyjście
5 USD
TekstObrazPlik
Zobacz model
Aktywny

Claude Opus 4.5 is Anthropic’s frontier reasoning model optimized for complex software engineering, agentic workflows, and long-horizon computer use. It offers strong multimodal capabilities, competitive performance across real-world coding and…

Kontekst
200K
Wejście
5 USD
Wyjście
25 USD
PlikObrazTekst
Zobacz model
Aktywny

Opus 4.6 is Anthropic’s strongest model for coding and long-running professional tasks. It is built for agents that operate across entire workflows rather than single prompts, making it especially effective…

Kontekst
1M
Wejście
5 USD
Wyjście
25 USD
TekstObrazPlik
Zobacz model
Aktywny

Opus 4.7 is the next generation of Anthropic's Opus family, built for long-running, asynchronous agents. Building on the coding and agentic strengths of Opus 4.6, it delivers stronger performance on…

Kontekst
1M
Wejście
5 USD
Wyjście
25 USD
TekstObrazPlik
Zobacz model
Aktywny

Fast-mode variant of Opus 4.7 - identical capabilities with higher output speed at premium 6x pricing.

Kontekst
1M
Wejście
30 USD
Wyjście
150 USD
TekstObrazPlik
Zobacz model
Aktywny

Claude Opus 4.8 is Anthropic's most capable generally available model in the Opus family. It supports text, image, and file inputs with text output, with reasoning support and a 1M-token…

Kontekst
1M
Wejście
5 USD
Wyjście
25 USD
TekstObrazPlik
Zobacz model
Aktywny

Fast-mode variant of Opus 4.8 - identical capabilities with higher output speed at 2x pricing relative to regular Opus 4.8.

Kontekst
1M
Wejście
10 USD
Wyjście
50 USD
TekstObrazPlik
Zobacz model
Aktywny

Claude Opus 5 is Anthropic’s flagship model for demanding reasoning, coding, and long-horizon agentic work. It is particularly strong at end-to-end software tasks, code review and bug finding, visual analysis…

Kontekst
1M
Wejście
5 USD
Wyjście
25 USD
TekstObrazPlik
Zobacz model
Aktywny

Fast-mode variant of Opus 5 - identical capabilities with higher output speed at 2x pricing relative to regular Opus 5.

Kontekst
1M
Wejście
10 USD
Wyjście
50 USD
TekstObrazPlik
Zobacz model
Aktywny

This model always redirects to the latest model in the Claude Opus family.

Kontekst
1M
Wejście
5 USD
Wyjście
25 USD
TekstObrazPlik
Zobacz model
Aktywny

Claude Sonnet 4.5 is Anthropic’s most advanced Sonnet model to date, optimized for real-world agents and coding workflows. It delivers state-of-the-art performance on coding benchmarks such as SWE-bench Verified, with…

Kontekst
1M
Wejście
3 USD
Wyjście
15 USD
TekstObrazPlik
Zobacz model
Aktywny

Sonnet 4.6 is Anthropic's most capable Sonnet-class model yet, with frontier performance across coding, agents, and professional work. It excels at iterative development, complex codebase navigation, end-to-end project management with…

Kontekst
1M
Wejście
3 USD
Wyjście
15 USD
TekstObrazPlik
Zobacz model
Aktywny

Mistral Codestral Embed is specially designed for code, perfect for embedding code databases, repositories, and powering coding assistants with state-of-the-art retrieval.

Kontekst
8K
Wejście
0,15 USD
Wyjście
0,00 USD
Tekst
Zobacz model

Deep Cogito

Cogito v2.1 671B

Aktywny

Cogito v2.1 671B MoE represents one of the strongest open models globally, matching performance of frontier closed and open models. This model is trained using self play with reinforcement learning…

Kontekst
128K
Wejście
1,25 USD
Wyjście
1,25 USD
Tekst
Zobacz model

Sesame

CSM 1B

Aktywny

CSM 1B is a conversational speech model from Sesame. It accepts text input and produces English speech output, with voice options spanning conversational and read-speech styles. At 1B parameters, it…

Kontekst
4K
Wejście
7 USD
Wyjście
0,00 USD
Tekst
Zobacz model
Aktywny

> I have to praise this model for good focus. I said earlier that it still remembers it at 12K. I think my personal evaluation of it has already beaten the rest.

Kontekst
131K
Wejście
0,30 USD
Wyjście
0,50 USD
Tekst
Zobacz model
Aktywny

DeepSeek-V3.1 is a hybrid model that supports both thinking mode and non-thinking mode. Compared to the previous version, this upgrade brings improvements in multiple aspects:

Kontekst
164K
Wejście
0,25 USD
Wyjście
0,95 USD
Tekst
Zobacz model
Aktywny

This update maintains the model's original capabilities while addressing issues reported by users, including:

Kontekst
164K
Wejście
0,27 USD
Wyjście
1 USD
Tekst
Zobacz model
Aktywny

We introduce **DeepSeek-V3.2**, a model that harmonizes high computational efficiency with superior reasoning and agent performance. Our approach is built upon three key technical breakthroughs:

Kontekst
164K
Wejście
0,269 USD
Wyjście
0,40 USD
Tekst
Zobacz model
Aktywny

We are excited to announce the official release of DeepSeek-V3.2-Exp, an experimental version of our model. As an intermediate step toward our next-generation architecture, V3.2-Exp builds upon V3.1-Terminus by introducing DeepSeek Sparse Attention—a sparse attention mechanism designed to explore and validate optimizations for training and inference efficiency in long-context scenarios.

Kontekst
164K
Wejście
0,27 USD
Wyjście
0,41 USD
Tekst
Zobacz model
Aktywny

We present a preview version of **DeepSeek-V4** series, including two strong Mixture-of-Experts (MoE) language models — **DeepSeek-V4-Pro** with 1.6T parameters (49B activated) and **DeepSeek-V4-Flash** with 284B parameters (13B activated) — both supporting a context length of **one million tokens**.

Kontekst
1,02M
Wejście
0,0826 USD
Wyjście
0,1652 USD
Tekst
Zobacz model
Aktywny

**DeepSeek-V4-Flash-0731** is the official release of **DeepSeek-V4-Flash**, superseding the preview version, with substantially enhanced agentic capabilities. It has the same model structure as DeepSeek-V4-Flash-DSpark, i.e. it comes with a speculative decoding module attached.

Kontekst
1,05M
Wejście
0,14 USD
Wyjście
0,28 USD
Tekst
Zobacz model
Aktywny

This model always redirects to the latest model in the DeepSeek V4 Flash family.

Kontekst
1,02M
Wejście
0,0786 USD
Wyjście
0,1572 USD
Tekst
Zobacz model
Aktywny

We present a preview version of **DeepSeek-V4** series, including two strong Mixture-of-Experts (MoE) language models — **DeepSeek-V4-Pro** with 1.6T parameters (49B activated) and **DeepSeek-V4-Flash** with 284B parameters (13B activated) — both supporting a context length of **one million tokens**.

Kontekst
1,05M
Wejście
1,32 USD
Wyjście
3,96 USD
Tekst
Zobacz model
Aktywny

Dots3-Note Preview is an open-weight mixture-of-experts model from Dots Studio, with 16B active parameters out of 280B total. It is the lightest model in the Dots 3 family and is…

Kontekst
512K
Wejście
0,00 USD
Wyjście
0,00 USD
TekstObraz
Zobacz model

Intfloat

E5-Base-v2

Aktywny

Text Embeddings by Weakly-Supervised Contrastive Pre-training. Liang Wang, Nan Yang, Xiaolong Huang, Binxing Jiao, Linjun Yang, Daxin Jiang, Rangan Majumder, Furu Wei, arXiv 2022

Kontekst
512
Wejście
0,005 USD
Wyjście
0,00 USD
Tekst
Zobacz model

Intfloat

E5-Large-v2

Aktywny

Text Embeddings by Weakly-Supervised Contrastive Pre-training. Liang Wang, Nan Yang, Xiaolong Huang, Binxing Jiao, Linjun Yang, Daxin Jiang, Rangan Majumder, Furu Wei, arXiv 2022

Kontekst
512
Wejście
0,01 USD
Wyjście
0,00 USD
Tekst
Zobacz model

Perplexity

Embed V1 0.6B

Aktywny

pplx-embed-v1-0.6B is one of Perplexity's state-of-the-art text embedding models built for real-world, web-scale retrieval. pplx-embed-v1 is optimized for standard dense text retrieval with the 0.6B parameter model targeting lightweight, low-latency…

Kontekst
32K
Wejście
0,004 USD
Wyjście
0,00 USD
Tekst
Zobacz model

Perplexity

Embed V1 4B

Aktywny

pplx-embed-v1 -4B is one of Perplexity's state-of-the-art text embedding models built for real-world, web-scale retrieval. pplx-embed-v1 is optimized for standard dense text retrieval with the 4B parameter model maximizing retrieval…

Kontekst
32K
Wejście
0,03 USD
Wyjście
0,00 USD
Tekst
Zobacz model
Aktywny

Flux TTS is a text-to-speech model from Deepgram. It is suited for natural, expressive English speech synthesis across Deepgram's Flux voice catalog.

Kontekst
0
Wejście
0,00 USD
Wyjście
0,00 USD
Tekst
Zobacz model

Black Forest Labs

FLUX.2 Flex

Aktywny

FLUX.2 [flex] excels at rendering complex text, typography, and fine details, and supports multi-reference editing in the same unified architecture. Pricing is as follows, per the docs: We charge $0.06…

Kontekst
67K
Wejście
0,00 USD
Wyjście
0,00 USD
TekstObraz
Zobacz model

Black Forest Labs

FLUX.2 Klein 4B

Aktywny

The FLUX.2 [klein] model family are our fastest image models to date. FLUX.2 [klein] unifies generation and editing in a single compact architecture, **delivering state-of-the-art quality with end-to-end inference in as low as under a second**. Built for applications that require real-time image generation without sacrificing quality, and runs on consumer hardware, with as little as 13GB VRAM.

Kontekst
41K
Wejście
0,00 USD
Wyjście
0,00 USD
TekstObraz
Zobacz model

Black Forest Labs

FLUX.2 Max

Aktywny

FLUX.2 [max] is the new top-tier image model from Black Forest Labs, pushing image quality, prompt understanding, and editing consistency to the highest level yet. Pricing is as follows, [per…

Kontekst
47K
Wejście
0,00 USD
Wyjście
0,00 USD
TekstObraz
Zobacz model

Black Forest Labs

FLUX.2 Pro

Aktywny

A high-end image generation and editing model focused on frontier-level visual quality and reliability. It delivers strong prompt adherence, stable lighting, sharp textures, and consistent character/style reproduction across multi-reference inputs.…

Kontekst
47K
Wejście
0,00 USD
Wyjście
0,00 USD
TekstObraz
Zobacz model

Black Forest Labs

FLUX.3 Video

Aktywny

FLUX.3 Video is a video generation model from Black Forest Labs. It supports text-to-video, image-guided generation with opening and closing keyframes, and video continuation workflows, making it suited for controlled…

Kontekst
0
Wejście
0,00 USD
Wyjście
0,00 USD
TekstObrazWideo
Zobacz model
Aktywny

The simplest way to get free inference. the provider catalog/free is a router that selects free models at random from the models available on the provider catalog. The router smartly filters for models that…

Kontekst
200K
Wejście
0,00 USD
Wyjście
0,00 USD
TekstObraz
Zobacz model
Aktywny

Fugu Ultra is the higher-performance model in Sakana AI's Fugu family. Rather than a single monolithic model, Fugu is a learned multi-agent orchestration system: a language model trained to route…

Kontekst
1M
Wejście
5 USD
Wyjście
30 USD
TekstObraz
Zobacz model

OpenRouter

Fusion

Aktywny

Fusion turns your prompt into a small multi-model deliberation. A panel of expert models (see below) analyzes your prompt in parallel with web search and web fetch enabled, then a…

Kontekst
1M
Wejście
Nieznane
Wyjście
Nieznane
Tekst
Zobacz model
Aktywny

Gemini 3 Flash Preview is a high speed, high value thinking model designed for agentic workflows, multi turn chat, and coding assistance. It delivers near Pro level reasoning and tool…

Kontekst
1,05M
Wejście
0,50 USD
Wyjście
3 USD
TekstObrazPlikDźwięk
Zobacz model
Aktywny

Gemini 3.1 Flash Lite is Google’s GA high-efficiency multimodal model optimized for low-latency, high-volume workloads. It supports text, image, video, audio, and PDF inputs, and is designed for lightweight agentic…

Kontekst
1,05M
Wejście
0,25 USD
Wyjście
1,5 USD
TekstObrazWideoPlik
Zobacz model

Gemini 3.1 Flash Lite Preview is Google's high-efficiency model optimized for high-volume use cases. It outperforms Gemini 2.5 Flash Lite on overall quality and approaches Gemini 2.5 Flash performance across…

Kontekst
1,05M
Wejście
0,25 USD
Wyjście
1,5 USD
TekstObrazWideoPlik
Zobacz model

Gemini 3.1 Flash TTS Preview is a text-to-speech model from Google, and a substantial generational step up from Gemini 2.5 Flash TTS. It takes text input and produces audio output…

Kontekst
33K
Wejście
1 USD
Wyjście
20 USD
Tekst
Zobacz model
Aktywny

Gemini 3.1 Pro Preview is Google’s frontier reasoning model, delivering enhanced software engineering performance, improved agentic reliability, and more efficient token usage across complex workflows. Building on the multimodal foundation…

Kontekst
1,05M
Wejście
2 USD
Wyjście
12 USD
DźwiękPlikObrazTekst
Zobacz model

Gemini 3.1 Pro Preview Custom Tools is a variant of Gemini 3.1 Pro that improves tool selection behavior by preventing overuse of a general bash tool when more efficient third-party…

Kontekst
1,05M
Wejście
2 USD
Wyjście
12 USD
TekstDźwiękObrazWideo
Zobacz model
Aktywny

Gemini 3.5 Flash is Google's high-efficiency multimodal model, bringing near-Pro level coding and reasoning at Flash-tier cost and speed. It is highly optimized for coding proficiency and parallel agentic execution…

Kontekst
1,05M
Wejście
1,5 USD
Wyjście
9 USD
TekstObrazWideoPlik
Zobacz model
Aktywny

Gemini 3.5 Flash Lite is a high-efficiency model from Google with upgraded agentic capabilities. It is suited for subagents that execute focused tasks within complex, multi-agent workflows.

Kontekst
1,05M
Wejście
0,30 USD
Wyjście
2,5 USD
TekstObrazWideoPlik
Zobacz model
Aktywny

Gemini 3.6 Flash is a high-efficiency model from Google for coding, agentic workflows, and web and app development. It is designed to produce polished outputs with fewer unnecessary edits and…

Kontekst
1,05M
Wejście
0,75 USD
Wyjście
3,75 USD
TekstObrazWideoPlik
Zobacz model
Aktywny

gemini-embedding-001 provides a unified cutting edge experience across domains, including science, legal, finance, and coding. This embedding model has consistently held a top spot on the Massive Text Embedding Benchmark…

Kontekst
20K
Wejście
0,15 USD
Wyjście
0,00 USD
Tekst
Zobacz model
Aktywny

Gemini Embedding 2 is Google's first multimodal embedding model. We currently support mapping text and images into a unified vector space for semantic search and retrieval-augmented generation (RAG). It supports…

Kontekst
8K
Wejście
0,20 USD
Wyjście
0,00 USD
TekstObrazPlikDźwięk
Zobacz model

Gemini Embedding 2 Preview is Google's first multimodal embedding model. We currently support mapping text and images into a unified vector space for semantic search and retrieval-augmented generation (RAG). It…

Kontekst
8K
Wejście
0,20 USD
Wyjście
0,00 USD
TekstObrazPlikDźwięk
Zobacz model
Aktywny

Gemma is a family of open models built by Google DeepMind. Gemma 4 models are multimodal, handling text and image input (with audio supported on E2B, E4B, and 12B) and generating text output. This release includes open-weights models in both pre-trained and instruction-tuned variants. Gemma 4 features a context window of up to 256K tokens and maintains multilingual support in over 140 languages.

Kontekst
262K
Wejście
0,07 USD
Wyjście
0,34 USD
ObrazTekstWideo
Zobacz model
Aktywny

Gemma is a family of open models built by Google DeepMind. Gemma 4 models are multimodal, handling text and image input (with audio supported on E2B, E4B, and 12B) and generating text output. This release includes open-weights models in both pre-trained and instruction-tuned variants. Gemma 4 features a context window of up to 256K tokens and maintains multilingual support in over 140 languages.

Kontekst
262K
Wejście
0,10 USD
Wyjście
0,34 USD
ObrazTekstWideo
Zobacz model

Runway

Gen-4.5

Aktywny

Runway Gen-4.5 is a video generation model from Runway for text-to-video and image-to-video workflows. It is designed for cinematic scene creation with strong motion quality, visual fidelity, and prompt adherence.…

Kontekst
0
Wejście
0,00 USD
Wyjście
0,00 USD
TekstObraz
Zobacz model
Aktywny

📖 Check out the GLM-4.6 technical blog, technical report(GLM-4.5), and Zhipu AI technical documentation.

Kontekst
203K
Wejście
0,50 USD
Wyjście
2 USD
Tekst
Zobacz model
Aktywny

This model is part of the GLM-V family of models, introduced in the paper GLM-4.1V-Thinking and GLM-4.5V: Towards Versatile Multimodal Reasoning with Scalable Reinforcement Learning.

Kontekst
131K
Wejście
0,30 USD
Wyjście
0,90 USD
ObrazTekstWideo
Zobacz model
Aktywny

You can also see significant improvements in many other scenarios such as chat, creative writing, and role-play scenario.

Kontekst
203K
Wejście
0,40 USD
Wyjście
1,75 USD
Tekst
Zobacz model
Aktywny

GLM-4.7-Flash is a 30B-A3B MoE model. As the strongest model in the 30B class, GLM-4.7-Flash offers a new option for lightweight deployment that balances performance and efficiency.

Kontekst
203K
Wejście
0,06 USD
Wyjście
0,40 USD
Tekst
Zobacz model

Z.ai

GLM 5

Aktywny

We are launching GLM-5, targeting complex systems engineering and long-horizon agentic tasks. Scaling is still one of the most important ways to improve the intelligence efficiency of Artificial General Intelligence (AGI). Compared to GLM-4.5, GLM-5 scales from 355B parameters (32B active) to 744B parameters (40B active), and increases pre-training data from 23T to 28.5T tokens. GLM-5 also integrates DeepSeek Sparse Attention (DSA), largely reducing deployment cost while preserving long-context capacity.

Kontekst
198K
Wejście
0,60 USD
Wyjście
1,92 USD
Tekst
Zobacz model
Aktywny

GLM-5 Turbo is a new model from Z.ai designed for fast inference and strong performance in agent-driven environments such as OpenClaw scenarios. It is deeply optimized for real-world agent workflows…

Kontekst
203K
Wejście
1,2 USD
Wyjście
4 USD
Tekst
Zobacz model
Aktywny

GLM-5.1 is our next-generation flagship model for agentic engineering, with significantly stronger coding capabilities than its predecessor. It achieves state-of-the-art performance on SWE-Bench Pro and leads GLM-5 by a wide margin on NL2Repo (repo generation) and Terminal-Bench 2.0 (real-world terminal tasks).

Kontekst
200K
Wejście
0,966 USD
Wyjście
3,04 USD
Tekst
Zobacz model
Aktywny

We're introducing GLM-5.2, our latest flagship model for long-horizon tasks. It marks a substantial leap in long-horizon task capability over its predecessor GLM-5.1 and, for the first time, delivers that capability on a **solid 1M-token context**. GLM-5.2's new capabilities include:

Kontekst
1,05M
Wejście
0,50 USD
Wyjście
3,15 USD
Tekst
Zobacz model
Aktywny

GLM-5V-Turbo is Z.ai’s first native multimodal agent foundation model, built for vision-based coding and agent-driven tasks. It natively handles image, video, and text inputs, excels at long-horizon planning, complex coding,…

Kontekst
203K
Wejście
1,2 USD
Wyjście
4 USD
ObrazTekstWideo
Zobacz model

This model always redirects to the latest model in the Google Gemini Flash family.

Kontekst
1,05M
Wejście
0,375 USD
Wyjście
1,88 USD
TekstObrazWideoPlik
Zobacz model

This model always redirects to the latest model in the Google Gemini Pro family.

Kontekst
1,05M
Wejście
2 USD
Wyjście
12 USD
DźwiękPlikObrazTekst
Zobacz model

OpenAI

GPT Audio

Aktywny

The gpt-audio model is OpenAI's first generally available audio model. The new snapshot features an upgraded decoder for more natural sounding voices and maintains better voice consistency. Audio is priced…

Kontekst
128K
Wejście
2,5 USD
Wyjście
10 USD
TekstDźwięk
Zobacz model
Aktywny

A cost-efficient version of GPT Audio. The new snapshot features an upgraded decoder for more natural sounding voices and maintains better voice consistency. Input is priced at $0.60 per million…

Kontekst
128K
Wejście
0,60 USD
Wyjście
2,4 USD
TekstDźwięk
Zobacz model
Aktywny

GPT Chat Latest points to OpenAI's stable API alias chat-latest that always resolves to the latest Instant chat model used in ChatGPT. As OpenAI rolls out new Instant model updates…

Kontekst
400K
Wejście
5 USD
Wyjście
30 USD
TekstObrazPlik
Zobacz model
Aktywny

OpenAI's GPT Image 1 generates and edits images via the dedicated Images API. Features accurate text rendering, transparent backgrounds, and up to 16 reference images for edits.

Kontekst
400K
Wejście
10 USD
Wyjście
10 USD
TekstObraz
Zobacz model
Aktywny

A cost-efficient variant of GPT Image 1 for high-quality image generation at reduced latency and cost via OpenAI's dedicated Images API.

Kontekst
400K
Wejście
2,5 USD
Wyjście
2,5 USD
TekstObraz
Zobacz model
Aktywny

OpenAI's latest image generation model. Supports high-fidelity image generation and editing via the dedicated Images API.

Kontekst
400K
Wejście
8 USD
Wyjście
8 USD
TekstObraz
Zobacz model
Aktywny

GPT Transcribe is a high-accuracy speech-to-text model from OpenAI. It is suited for recorded audio, streamed file transcription, and committed Realtime turns, with free-form context, keyword hints, and multiple language…

Kontekst
0
Wejście
4500 USD
Wyjście
0,00 USD
Dźwięk
Zobacz model
Aktywny

GPT-4o Mini Transcribe is OpenAI's smaller, cost-efficient speech-to-text model built on GPT-4o Mini audio capabilities. It's priced per token (input and output), making it suitable for high-volume transcription workflows that…

Kontekst
128K
Wejście
1,25 USD
Wyjście
5 USD
Dźwięk
Zobacz model
Aktywny

GPT-4o Transcribe is OpenAI's high-quality speech-to-text model built on GPT-4o audio capabilities. It's priced per token (input and output), making it suitable for workflows that benefit from token-level billing transparency.

Kontekst
128K
Wejście
2,5 USD
Wyjście
10 USD
Dźwięk
Zobacz model
Aktywny

GPT-5-Codex is a specialized version of GPT-5 optimized for software engineering and coding workflows. It is designed for both interactive development sessions and long, independent execution of complex engineering tasks.…

Kontekst
400K
Wejście
0,625 USD
Wyjście
5 USD
TekstObraz
Zobacz model
Aktywny

GPT-5 Image combines OpenAI's GPT-5 model with state-of-the-art image generation capabilities. It offers major improvements in reasoning, code quality, and user experience while incorporating GPT Image 1's superior instruction following,…

Kontekst
400K
Wejście
10 USD
Wyjście
10 USD
ObrazTekstPlik
Zobacz model
Aktywny

GPT-5 Image Mini combines OpenAI's advanced language capabilities, powered by GPT-5 Mini, with GPT Image 1 Mini for efficient image generation. This natively multimodal model features superior instruction following, text…

Kontekst
400K
Wejście
2,5 USD
Wyjście
2 USD
PlikObrazTekst
Zobacz model

OpenAI

GPT-5 Pro

Aktywny

GPT-5 Pro is OpenAI’s most advanced model, offering major improvements in reasoning, code quality, and user experience. It is optimized for complex tasks that require step-by-step reasoning, instruction following, and…

Kontekst
400K
Wejście
15 USD
Wyjście
120 USD
ObrazTekstPlik
Zobacz model

OpenAI

GPT-5.1

Aktywny

GPT-5.1 is the latest frontier-grade model in the GPT-5 series, offering stronger general-purpose reasoning, improved instruction adherence, and a more natural conversational style compared to GPT-5. It uses adaptive reasoning…

Kontekst
400K
Wejście
1,25 USD
Wyjście
10 USD
ObrazTekstPlik
Zobacz model
Aktywny

GPT-5.1-Codex is a specialized version of GPT-5.1 optimized for software engineering and coding workflows. It is designed for both interactive development sessions and long, independent execution of complex engineering tasks.…

Kontekst
400K
Wejście
1,25 USD
Wyjście
10 USD
TekstObraz
Zobacz model
Aktywny

GPT-5.1-Codex-Max is OpenAI’s latest agentic coding model, designed for long-running, high-context software development tasks. It is based on an updated version of the 5.1 reasoning stack and trained on agentic…

Kontekst
400K
Wejście
1,25 USD
Wyjście
10 USD
TekstObraz
Zobacz model
Aktywny

GPT-5.1-Codex-Mini is a smaller and faster version of GPT-5.1-Codex

Kontekst
400K
Wejście
0,25 USD
Wyjście
2 USD
ObrazTekst
Zobacz model

OpenAI

GPT-5.2

Aktywny

GPT-5.2 is the latest frontier-grade model in the GPT-5 series, offering stronger agentic and long context perfomance compared to GPT-5.1. It uses adaptive reasoning to allocate computation dynamically, responding quickly…

Kontekst
400K
Wejście
1,75 USD
Wyjście
14 USD
PlikObrazTekst
Zobacz model
Aktywny

GPT-5.2 Chat (AKA Instant) is the fast, lightweight member of the 5.2 family, optimized for low-latency chat while retaining strong general intelligence. It uses adaptive reasoning to selectively “think” on…

Kontekst
128K
Wejście
1,75 USD
Wyjście
14 USD
PlikObrazTekst
Zobacz model
Aktywny

GPT-5.2 Pro is OpenAI’s most advanced model, offering major improvements in agentic coding and long context performance over GPT-5 Pro. It is optimized for complex tasks that require step-by-step reasoning,…

Kontekst
400K
Wejście
21 USD
Wyjście
168 USD
ObrazTekstPlik
Zobacz model
Aktywny

GPT-5.2-Codex is an upgraded version of GPT-5.1-Codex optimized for software engineering and coding workflows. It is designed for both interactive development sessions and long, independent execution of complex engineering tasks.…

Kontekst
400K
Wejście
1,75 USD
Wyjście
14 USD
TekstObraz
Zobacz model
Aktywny

GPT-5.3-Codex is OpenAI’s most advanced agentic coding model, combining the frontier software engineering performance of GPT-5.2-Codex with the broader reasoning and professional knowledge capabilities of GPT-5.2. It achieves state-of-the-art results…

Kontekst
400K
Wejście
1,75 USD
Wyjście
14 USD
TekstObrazPlik
Zobacz model

OpenAI

GPT-5.4

Aktywny

GPT-5.4 is OpenAI’s latest frontier model, unifying the Codex and GPT lines into a single system. It features a 1M+ token context window (922K input, 128K output) with support for…

Kontekst
1,05M
Wejście
2,5 USD
Wyjście
15 USD
TekstObrazPlik
Zobacz model
Aktywny

GPT-5.4 mini brings the core capabilities of GPT-5.4 to a faster, more efficient model optimized for high-throughput workloads. It supports text and image inputs with strong performance across reasoning, coding,…

Kontekst
400K
Wejście
0,75 USD
Wyjście
4,5 USD
PlikObrazTekst
Zobacz model
Aktywny

GPT-5.4 nano is the most lightweight and cost-efficient variant of the GPT-5.4 family, optimized for speed-critical and high-volume tasks. It supports text and image inputs and is designed for low-latency…

Kontekst
400K
Wejście
0,20 USD
Wyjście
1,25 USD
PlikObrazTekst
Zobacz model
Aktywny

GPT-5.4 Pro is OpenAI's most advanced model, building on GPT-5.4's unified architecture with enhanced reasoning capabilities for complex, high-stakes tasks. It features a 1M+ token context window (922K input, 128K…

Kontekst
1,05M
Wejście
30 USD
Wyjście
180 USD
TekstObrazPlik
Zobacz model

OpenAI

GPT-5.5

Aktywny

GPT-5.5 is OpenAI’s frontier model designed for complex professional workloads, building on GPT-5.4 with stronger reasoning, higher reliability, and improved token efficiency on hard tasks. It features a 1M+ token…

Kontekst
1,05M
Wejście
5 USD
Wyjście
30 USD
PlikObrazTekst
Zobacz model
Aktywny

GPT-5.5 Pro is OpenAI’s high-capability model optimized for deep reasoning and accuracy on complex, high-stakes workloads. It features a 1M+ token context window (922K input, 128K output) with support for…

Kontekst
1,05M
Wejście
30 USD
Wyjście
180 USD
PlikObrazTekst
Zobacz model
Aktywny

GPT-5.6 Luna Pro is the same underlying model as GPT-5.6 Luna, served with reasoning.mode set to pro for higher-quality responses on complex tasks.

Kontekst
1,05M
Wejście
0,20 USD
Wyjście
1,2 USD
PlikObrazTekst
Zobacz model
Aktywny

GPT-5.6 Sol Pro is the same underlying model as GPT-5.6 Sol, served with reasoning.mode set to pro for higher-quality responses on complex tasks.

Kontekst
1,05M
Wejście
2,5 USD
Wyjście
15 USD
PlikObrazTekst
Zobacz model
Aktywny

GPT-5.6 Terra is a balanced model in OpenAI's GPT-5.6 series, positioned between the flagship Sol tier and the cost-efficient Luna tier. It is suited for everyday coding, reasoning, and agentic…

Kontekst
1,05M
Wejście
2 USD
Wyjście
12 USD
PlikObrazTekst
Zobacz model
Aktywny

GPT-5.6 Terra Pro is the same underlying model as GPT-5.6 Terra, served with reasoning.mode set to pro for higher-quality responses on complex tasks.

Kontekst
1,05M
Wejście
2 USD
Wyjście
12 USD
PlikObrazTekst
Zobacz model
Aktywny

gpt-oss-safeguard-120b and gpt-oss-safeguard-20b are safety reasoning models built-upon gpt-oss. With these models, you can classify text content based on safety policies that you provide and perform a suite of foundational safety tasks. These models are intended for safety use cases. For other applications, we recommend using gpt-oss models.

Kontekst
131K
Wejście
0,075 USD
Wyjście
0,30 USD
Tekst
Zobacz model
Aktywny

📣 **Update [10-07-2025]:** Added a *default system prompt* to the chat template to guide the model towards more *professional, accurate, and safe* responses.

Kontekst
131K
Wejście
0,017 USD
Wyjście
0,112 USD
Tekst
Zobacz model
Aktywny

**Model Summary:** Granite-4.1-8B is a 8B parameter long-context instruct model finetuned from *Granite-4.1-8B-Base* using a combination of open source instruction datasets with permissive license and internally collected synthetic datasets. Granite 4.1 models have gone through an improved post-training pipeline, including supervised finetuning and reinforcement learning alignment, resulting in enhanced tool calling, instruction following, and chat capabilities.

Kontekst
131K
Wejście
0,05 USD
Wyjście
0,10 USD
Tekst
Zobacz model

SpaceXAI

Grok 4.20

Aktywny

Grok 4.20 is a reasoning model from SpaceXAI with industry-leading speed and agentic tool calling capabilities. It combines the lowest hallucination rate on the market with strict prompt adherance, delivering…

Kontekst
2M
Wejście
1,25 USD
Wyjście
2,5 USD
TekstObrazPlik
Zobacz model
Aktywny

Grok 4.20 Multi-Agent is a variant of SpaceXAI’s Grok 4.20 designed for collaborative, agent-based workflows. Multiple agents operate in parallel to conduct deep research, coordinate tool use, and synthesize information…

Kontekst
2M
Wejście
1,25 USD
Wyjście
2,5 USD
TekstObrazPlik
Zobacz model

SpaceXAI

Grok 4.3

Aktywny

Grok 4.3 is a reasoning model from SpaceXAI. It accepts text and image inputs with text output, and is suited for agentic workflows, instruction-following tasks, and applications requiring high factual…

Kontekst
1M
Wejście
1,25 USD
Wyjście
2,5 USD
TekstObrazPlik
Zobacz model

SpaceXAI

Grok 4.5

Aktywny

Grok 4.5 is a model from SpaceXAI with frontier performance on coding, knowledge work, and STEM.

Kontekst
500K
Wejście
2 USD
Wyjście
6 USD
TekstObrazPlik
Zobacz model
Aktywny

Grok Build 0.1 is SpaceXAI’s fast coding model trained specifically for agentic software engineering workflows. It supports text and image inputs with text output, and is optimized for interactive coding…

Kontekst
256K
Wejście
1 USD
Wyjście
2 USD
TekstObrazPlik
Zobacz model

Grok Imagine Image 2.0 is an image generation and editing model from xAI. It is suited for creating images from text prompts and editing images from references, with low and…

Kontekst
66K
Wejście
0,00 USD
Wyjście
0,00 USD
TekstObraz
Zobacz model

Grok Imagine Image Quality is SpaceXAI's fast, high-fidelity image generation and editing model. It accepts text prompts and optional reference images, producing photorealistic outputs at 1K or 2K across a…

Kontekst
66K
Wejście
0,00 USD
Wyjście
0,00 USD
TekstObraz
Zobacz model
Aktywny

Grok Imagine Video is SpaceXAI's fast, text-, image-, and reference-conditioned video generation model. It produces short videos (1–15 seconds, 24 fps) at 480p or 720p across seven aspect ratios -…

Kontekst
0
Wejście
0,00 USD
Wyjście
0,00 USD
TekstObraz
Zobacz model
Aktywny

Grok Imagine Video 1.5 is a video generation model from SpaceXAI. It creates videos from text prompts, with an optional starting image to guide the scene. It can direct subject…

Kontekst
0
Wejście
0,00 USD
Wyjście
0,00 USD
TekstObraz
Zobacz model
Aktywny

This model always redirects to the latest Grok model from xAI.

Kontekst
500K
Wejście
2 USD
Wyjście
6 USD
TekstObrazPlik
Zobacz model

SpaceXAI

Grok STT 1.0

Aktywny

Grok STT is SpaceXAI's speech-to-text model, available via the REST /v1/stt endpoint. It supports transcription with word-level timestamps, optional speaker diarization, and multichannel audio.

Kontekst
0
Wejście
100 000 USD
Wyjście
0,00 USD
Dźwięk
Zobacz model
Aktywny

Grok Voice TTS 1.0 is a text-to-speech model from SpaceXAI. It converts text into spoken audio across 20+ languages with automatic language detection, and offers five built-in voices (Eve, Ara,…

Kontekst
15K
Wejście
15 USD
Wyjście
0,00 USD
Tekst
Zobacz model

Thenlper

GTE-Base

Aktywny

General Text Embeddings (GTE) model. Towards General Text Embeddings with Multi-stage Contrastive Learning

Kontekst
512
Wejście
0,005 USD
Wyjście
0,00 USD
Tekst
Zobacz model

Thenlper

GTE-Large

Aktywny

General Text Embeddings (GTE) model. Towards General Text Embeddings with Multi-stage Contrastive Learning

Kontekst
512
Wejście
0,01 USD
Wyjście
0,00 USD
Tekst
Zobacz model

MiniMax

H3

Aktywny

MiniMax H3 is a lightweight, open-weights video generation model from MiniMax. It is designed for precise multimodal editing and controlled content generation, including instruction-guided edits, text and brand rendering, and…

Kontekst
0
Wejście
0,00 USD
Wyjście
0,00 USD
TekstObrazWideoDźwięk
Zobacz model

MiniMax

Hailuo 2.3

Aktywny

Hailuo 2.3 is a video generation model from MiniMax. It accepts text prompts and reference images as input and generates video output, supporting both text-to-video and image-to-video workflows. It is…

Kontekst
0
Wejście
0,00 USD
Wyjście
0,00 USD
TekstObraz
Zobacz model
Aktywny

HappyHorse 1.0 is a video generation model from Alibaba. It generates short videos from a text prompt, a single starting image, or a set of reference images, with output up…

Kontekst
0
Wejście
0,00 USD
Wyjście
0,00 USD
TekstObraz
Zobacz model
Aktywny

HappyHorse 1.1 is a video generation model from Alibaba. It generates short videos from a text prompt, a single starting image, or a set of reference images, with output up…

Kontekst
0
Wejście
0,00 USD
Wyjście
0,00 USD
TekstObraz
Zobacz model
Aktywny

Hermes 4 405B is a frontier, hybrid-mode **reasoning** model based on Llama-3.1-405B by Nous Research that is aligned to **you**.

Kontekst
131K
Wejście
1 USD
Wyjście
3 USD
Tekst
Zobacz model
Aktywny

Hermes 4 70B is a frontier, hybrid-mode **reasoning** model based on Llama-3.1-70B by Nous Research that is aligned to **you**.

Kontekst
131K
Wejście
0,13 USD
Wyjście
0,40 USD
Tekst
Zobacz model

Tencent

Hy3

Aktywny

**Hy3** is a 295B-parameter Mixture-of-Experts (MoE) model with 21B active parameters and 3.8B MTP layer parameters, developed by the Tencent Hy Team. Following the Hy3 Preview launch in late April, we gathered feedback from 50+ products and scaled up post-training with higher quality data. Today, we introduce Hy3, which outperforms similar-size models and rivals flagship open-source models with 2-5x parameters. It also shows significant gains in utility across various products and productivity tasks.

Kontekst
262K
Wejście
0,132 USD
Wyjście
0,528 USD
Tekst
Zobacz model
Aktywny

**Hy3 preview** is a 295B-parameter Mixture-of-Experts (MoE) model with 21B active parameters and 3.8B MTP layer parameters, developed by the Tencent Hy Team. Hy3 preview is the first model trained on our rebuilt infrastructure, and the strongest we've shipped so far. It improves significantly on complex reasoning, instruction following, context learning, coding, and agent tasks.

Kontekst
262K
Wejście
0,18 USD
Wyjście
0,60 USD
Tekst
Zobacz model

Thinking Machines

Inkling

Aktywny

Inkling is a general-purpose multimodal model that accepts text, image and audio inputs and generates text outputs. It is intended for use in English and other languages, and across multiple coding languages. The model is designed to be used by developers building AI-powered applications, including agentic and tool-use systems, coding assistants, chatbots, and retrieval-augmented generation systems, and is suitable for general-purpose conversational use, instruction-following, and other natural language and multimodal tasks. It is released with open weights to support research, fine-tuning and integration into third-party products by downstream developers.

Kontekst
524K
Wejście
0,95 USD
Wyjście
4,05 USD
TekstObrazDźwięk
Zobacz model

Thinking Machines

Inkling Small

Aktywny

Inkling-Small is a general-purpose multimodal model that accepts text, image and audio inputs and generates text outputs. It is intended for use in English and other languages, and across multiple coding languages. The model is designed to be used by developers building AI-powered applications, including agentic and tool-use systems, coding assistants, chatbots, and retrieval-augmented generation systems, and is suitable for general-purpose conversational use, instruction-following, and other natural language and multimodal tasks. It is released with open weights to support research, fine-tuning and integration into third-party products by downstream developers.

Kontekst
524K
Wejście
0,45 USD
Wyjście
1,2 USD
TekstObrazDźwięk
Zobacz model
Aktywny

KAT-Coder-Air V2.5 is a flagship-level Agentic Coding model that can directly hand over an entire issue or an entire business workflow to it, allowing it to autonomously locate and make…

Kontekst
256K
Wejście
0,15 USD
Wyjście
0,60 USD
Tekst
Zobacz model
Aktywny

KAT-Coder-Pro V2 is the latest high-performance model in KwaiKAT’s KAT-Coder series, designed for complex enterprise-grade software engineering and SaaS integration. It builds on the agentic coding strengths of earlier versions,…

Kontekst
256K
Wejście
0,30 USD
Wyjście
1,2 USD
Tekst
Zobacz model
Aktywny

KAT-Coder-Pro V2.5 is a flagship-level Agentic Coding model that can directly hand over an entire issue or an entire business workflow to it, allowing it to autonomously locate and make…

Kontekst
256K
Wejście
0,74 USD
Wyjście
2,96 USD
Tekst
Zobacz model

MoonshotAI

Kimi K2 0905

Aktywny

📰   Tech Blog     |     📄   Paper

Kontekst
262K
Wejście
0,60 USD
Wyjście
2,5 USD
Tekst
Zobacz model
Aktywny

Kimi K2 Thinking is the latest, most capable version of open-source thinking model. Starting with Kimi K2, we built it as a thinking agent that reasons step-by-step while dynamically invoking tools. It sets a new state-of-the-art on Humanity's Last Exam (HLE), BrowseComp, and other benchmarks by dramatically scaling multi-step reasoning depth and maintaining stable tool-use across 200–300 sequential calls. At the same time, K2 Thinking is a native INT4 quantization model with 256k context window, achieving lossless reductions in inference latency and GPU memory usage.

Kontekst
262K
Wejście
0,60 USD
Wyjście
2,5 USD
Tekst
Zobacz model

MoonshotAI

Kimi K2.5

Aktywny

📰   Tech Blog     |     📄   Paper

Kontekst
262K
Wejście
0,57 USD
Wyjście
2,85 USD
TekstObraz
Zobacz model

MoonshotAI

Kimi K2.6

Aktywny

Kimi K2.6 is an open-source, native multimodal agentic model that advances practical capabilities in long-horizon coding, coding-driven design, proactive autonomous execution, and swarm-based task orchestration.

Kontekst
262K
Wejście
0,95 USD
Wyjście
4 USD
TekstObraz
Zobacz model

MoonshotAI

Kimi K2.7 Code

Aktywny

Kimi K2.7 Code is a coding-focused agentic model built upon Kimi K2.6. With substantial improvements on real-world long-horizon coding tasks, it strengthens end-to-end task completion across complex software engineering workflows while improving token efficiency, reducing thinking-token usage by approximately 30% compared with Kimi K2.6.

Kontekst
262K
Wejście
0,71 USD
Wyjście
3,5 USD
TekstObraz
Zobacz model

MoonshotAI

Kimi K3

Aktywny

Kimi K3 is an open-weight, native multimodal agentic model and our most capable model to date. It is a 2.8T-parameter model built on Kimi Delta Attention (KDA) and Attention Residuals (AttnRes), with native vision capabilities and a 1-million-token context window. It is the world's first open 3T-class model, designed for frontier intelligence across long-horizon coding, knowledge work, and reasoning.

Kontekst
1,05M
Wejście
3 USD
Wyjście
15 USD
TekstObrazWideo
Zobacz model

hexgrad

Kokoro 82M

Aktywny

Kokoro 82M is a lightweight, open-weight text-to-speech model from hexgrad. It converts text to speech across 8 languages (American and British English, Spanish, French, Hindi, Italian, Japanese, Portuguese, and Chinese)…

Kontekst
4K
Wejście
0,62 USD
Wyjście
0,00 USD
Tekst
Zobacz model
Aktywny

Krea 2 Large is Krea's high-capability image generation model, more than twice the size of Krea 2 Medium. Its lighter post-training gives images a rawer, more textured, and flexible character,…

Kontekst
66K
Wejście
0,00 USD
Wyjście
0,00 USD
TekstObraz
Zobacz model
Aktywny

Krea 2 Medium is Krea's balanced, cost-efficient image generation model and a practical starting point for a broad range of use cases. Its extensive post-training supports stable, consistent generations, with…

Kontekst
66K
Wejście
0,00 USD
Wyjście
0,00 USD
TekstObraz
Zobacz model
Aktywny

Krea 2 Medium Turbo is a distilled, speed-focused variant of Krea 2 Medium from Krea. It is designed for rapid iteration and graphic design exploration where fast generation is the…

Kontekst
66K
Wejście
0,00 USD
Wyjście
0,00 USD
TekstObraz
Zobacz model

Poolside

Laguna S 2.1

Aktywny

Laguna S 2.1 is a 118B total parameter Mixture-of-Experts model with 8B activated parameters per token, designed for agentic coding and long-horizon work. It sits between Laguna XS 2.1 (33B-A3B) and Laguna M.1 (225B-A23B) in the Laguna series and shares the family recipe: a token-choice router with softplus gating over 256 routed experts plus one shared expert, grouped-query attention, and interleaved full/sliding-window attention.

Kontekst
1,05M
Wejście
0,09 USD
Wyjście
0,18 USD
Tekst
Zobacz model
Aktywny

Laguna XS 2.1 is a 33B total parameter Mixture-of-Experts model with 3B activated parameters per token designed for agentic coding and long-horizon work on a local machine. This model is an upgraded version of our Laguna XS.2 model with a +5.4% jump on SWE-bench Multilingual as well as stronger performance on terminal-style tasks.

Kontekst
262K
Wejście
0,06 USD
Wyjście
0,12 USD
Tekst
Zobacz model
Aktywny

LFM2.5-2.6B is a compact reasoning model from Liquid AI. It is suited for agent workflows, data extraction, RAG, and long-context processing. Liquid advises against using it for agentic coding or…

Kontekst
128K
Wejście
0,00 USD
Wyjście
0,00 USD
Tekst
Zobacz model

inclusionAI

Ling-2.6-1T

Aktywny

Ling-2.6-1T is an instant (instruct) model from inclusionAI and the company’s trillion-parameter flagship, designed for real-world agents that require fast execution and high efficiency at scale. It uses a “fast…

Kontekst
262K
Wejście
0,075 USD
Wyjście
0,625 USD
Tekst
Zobacz model

inclusionAI

Ling-2.6-flash

Aktywny

Ling-2.6-flash is an instant (instruct) model from inclusionAI with 104B total parameters and 7.4B active parameters, designed for real-world agents that require fast responses, strong execution, and high token efficiency.…

Kontekst
262K
Wejście
0,01 USD
Wyjście
0,03 USD
Tekst
Zobacz model
Aktywny

🤗 Hugging Face    |   🤖 ModelScope    |   🐙 the provider catalog   

Kontekst
262K
Wejście
0,021 USD
Wyjście
0,063 USD
Tekst
Zobacz model

Llama Nemotron Rerank VL 1B V2 is a 1.7B multimodal reranking model from NVIDIA. It evaluates the relevance of document images and text against user queries, designed for vision RAG…

Kontekst
10K
Wejście
0,00 USD
Wyjście
0,00 USD
TekstObraz
Zobacz model
Aktywny

LongCat 2.0 is a sparse mixture-of-experts language model from Meituan, with 48B active parameters out of 1.6T total. It is suited for coding, repository-level changes, long-horizon problem solving, and agentic…

Kontekst
1,05M
Wejście
0,30 USD
Wyjście
1,2 USD
Tekst
Zobacz model
Aktywny

30 second duration clips are priced at $0.04 per clip. Lyria 3 is Google's family of music generation models, available through the Gemini API. With Lyria 3, you can generate…

Kontekst
1,05M
Wejście
0,00 USD
Wyjście
0,00 USD
TekstObraz
Zobacz model
Aktywny

Full-length songs are priced at $0.08 per song. Lyria 3 is Google's family of music generation models, available through the Gemini API. With Lyria 3, you can generate high-quality, 48kHz…

Kontekst
1,05M
Wejście
0,00 USD
Wyjście
0,00 USD
TekstObraz
Zobacz model

Microsoft

MAI-Image-2.5

Aktywny

Microsoft's MAI-Image-2.5 is a high-quality image generation model available via Azure AI Foundry. It produces photorealistic and artistic images from text prompts with support for various aspect ratios.

Kontekst
4K
Wejście
5 USD
Wyjście
0,00 USD
TekstObraz
Zobacz model
Aktywny

Microsoft's MAI-Image-2.5 is a high-quality image generation model available via Azure AI Foundry. It produces photorealistic and artistic images from text prompts with support for various aspect ratios.

Kontekst
4K
Wejście
5 USD
Wyjście
0,00 USD
TekstObraz
Zobacz model
Aktywny

MAI-Transcribe 1.5 is a multilingual speech-to-text model from Microsoft AI. It is suited for captions, call transcription, subtitling, accessibility, and other voice-enabled applications, with reliable transcription across 43 languages, diverse…

Kontekst
0
Wejście
360 000 USD
Wyjście
0,00 USD
Dźwięk
Zobacz model

Microsoft

MAI-Voice-2

Aktywny

MAI-Voice-2 is an expressive text-to-speech model from Microsoft. It is suited for conversational assistants, media narration, accessibility, education, and other long-form voice applications. It supports 15 languages across 18 locales,…

Kontekst
0
Wejście
22 USD
Wyjście
0,00 USD
Tekst
Zobacz model
Aktywny

MAI-Voice-2-Flash is a low-latency text-to-speech model from Microsoft for voice agents, assistants, call centers, accessibility, narration, and other interactive applications. It generates expressive 24 kHz mono speech across 15 languages…

Kontekst
0
Wejście
15 USD
Wyjście
0,00 USD
Tekst
Zobacz model

Inception

Mercury 2

Aktywny

Mercury 2 is an extremely fast reasoning LLM, and the first reasoning diffusion LLM (dLLM). Instead of generating tokens sequentially, Mercury 2 produces and refines multiple tokens in parallel, achieving…

Kontekst
128K
Wejście
0,25 USD
Wyjście
0,75 USD
Tekst
Zobacz model

Xiaomi

MiMo-V2.5

Aktywny

🤗 HuggingFace  | 📰 Blog  | 🎨 Xiaomi MiMo API Platform  | 🗨️ Xiaomi MiMo Studio  |

Kontekst
1,05M
Wejście
0,14 USD
Wyjście
0,28 USD
TekstDźwiękObrazWideo
Zobacz model
Aktywny

🤗 HuggingFace  | 📰 Blog  | 🎨 Xiaomi MiMo API Platform  | 🗨️ Xiaomi MiMo Studio  |

Kontekst
1,05M
Wejście
0,435 USD
Wyjście
0,87 USD
Tekst
Zobacz model

MiniMax

MiniMax M2

Aktywny

Today, we release and open source MiniMax-M2, a **Mini** model built for **Max** coding & agentic workflows.

Kontekst
205K
Wejście
0,255 USD
Wyjście
1,02 USD
Tekst
Zobacz model
Aktywny

MiniMax M2-her is a dialogue-first large language model built for immersive roleplay, character-driven chat, and expressive multi-turn conversations. Designed to stay consistent in tone and personality, it supports rich message…

Kontekst
66K
Wejście
0,30 USD
Wyjście
1,2 USD
Tekst
Zobacz model
Aktywny

Today, we are handing **MiniMax-M2.1** over to the open-source community. This release is more than just a parameter update; it is a significant step toward democratizing top-tier agentic capabilities.

Kontekst
205K
Wejście
0,30 USD
Wyjście
1,2 USD
Tekst
Zobacz model
Aktywny

Extensively trained with reinforcement learning in hundreds of thousands of complex real-world environments, M2.5 is **SOTA in coding, agentic tool use and search, office work, and a range of other economically valuable tasks**, boasting scores of **80.2% in SWE-Bench Verified, 51.3% in Multi-SWE-Bench, and 76.3% in BrowseComp** (with context management).

Kontekst
197K
Wejście
0,22 USD
Wyjście
0,90 USD
Tekst
Zobacz model
Aktywny

**MiniMax-M2.7** is our first model deeply participating in its own evolution. M2.7 is capable of building complex agent harnesses and completing highly elaborate productivity tasks, leveraging Agent Teams, complex Skills, and dynamic tool search. For more details, see our blog post.

Kontekst
205K
Wejście
0,30 USD
Wyjście
1,2 USD
Tekst
Zobacz model

MiniMax

MiniMax M3

Aktywny

MiniMax-M3 is a native multimodal model with 1M context. It has ~428B parameters and ~23B activated parameters.

Kontekst
524K
Wejście
0,30 USD
Wyjście
1,2 USD
TekstObrazWideo
Zobacz model
Aktywny

The largest model in the Ministral 3 family, **Ministral 3 14B** offers frontier capabilities and performance comparable to its larger Mistral Small 3.2 24B counterpart. A powerful and efficient language model with vision capabilities.

Kontekst
262K
Wejście
0,20 USD
Wyjście
0,20 USD
TekstObraz
Zobacz model
Aktywny

The smallest model in the Ministral 3 family, **Ministral 3 3B** is a powerful, efficient tiny language model with vision capabilities.

Kontekst
131K
Wejście
0,10 USD
Wyjście
0,10 USD
TekstObraz
Zobacz model
Aktywny

A balanced model in the Ministral 3 family, **Ministral 3 8B** is a powerful, efficient tiny language model with vision capabilities.

Kontekst
262K
Wejście
0,15 USD
Wyjście
0,15 USD
TekstObraz
Zobacz model
Aktywny

Mistral Embed is a specialized embedding model for text data, optimized for semantic search and RAG applications. Developed by Mistral AI in late 2023, it produces 1024-dimensional vectors that effectively…

Kontekst
8K
Wejście
0,10 USD
Wyjście
0,00 USD
Tekst
Zobacz model
Aktywny

Mistral Large 3 2512 is Mistral’s most capable model to date, featuring a sparse mixture-of-experts architecture with 41B active parameters (675B total), and released under the Apache 2.0 license.

Kontekst
262K
Wejście
0,50 USD
Wyjście
1,5 USD
TekstObrazPlik
Zobacz model
Aktywny

Mistral Small 4 is a powerful hybrid model capable of acting as both a general instruction model and a reasoning model. It unifies the capabilities of three different model families—**Instruct**, **Reasoning** (previously called Magistral), and **Devstral**—into a single, unified model.

Kontekst
262K
Wejście
0,15 USD
Wyjście
0,60 USD
TekstObraz
Zobacz model

This model always redirects to the latest model in the MoonshotAI Kimi family.

Kontekst
975K
Wejście
2,6 USD
Wyjście
13 USD
TekstObrazWideo
Zobacz model

Sentence Transformers

multi-qa-mpnet-base-dot-v1

Aktywny

This is a sentence-transformers model: It maps sentences & paragraphs to a 768 dimensional dense vector space and was designed for **semantic search**. It has been trained on 215M (question, answer) pairs from diverse sources. For an introduction to semantic search, have a look at: SBERT.net - Semantic Search

Kontekst
512
Wejście
0,005 USD
Wyjście
0,00 USD
Tekst
Zobacz model
Aktywny

Multilingual E5 Text Embeddings: A Technical Report. Liang Wang, Nan Yang, Xiaolong Huang, Linjun Yang, Rangan Majumder, Furu Wei, arXiv 2024

Kontekst
512
Wejście
0,01 USD
Wyjście
0,00 USD
Tekst
Zobacz model
Aktywny

Muse Glimmer 30B is a dense, open-weight multimodal model from Meta Superintelligence Labs, distilled from Muse Spark and optimized for autonomous agents on consumer hardware. It is suited for long-horizon…

Kontekst
131K
Wejście
0,35 USD
Wyjście
1,5 USD
TekstObraz
Zobacz model
Aktywny

Muse Spark 1.1 is a multimodal reasoning model from Meta, built for agentic tasks. It accepts text, images, video, audio, and PDF documents and returns text, with a 1M-token context…

Kontekst
1,05M
Wejście
1,25 USD
Wyjście
4,25 USD
TekstObrazWideoPlik
Zobacz model
Aktywny

Muse Spark 1.2 is a reasoning model from Meta, designed for complex agentic tasks. It accepts text, images, video, audio, and PDF documents, returns text, and offers a 1M-token context…

Kontekst
1,05M
Wejście
1,25 USD
Wyjście
4,25 USD
TekstObrazWideoPlik
Zobacz model

Gemini 2.5 Flash Image, a.k.a. "Nano Banana," is now generally available. It is a state of the art image generation model with contextual understanding. It is capable of image generation,…

Kontekst
33K
Wejście
0,30 USD
Wyjście
2,5 USD
ObrazTekst
Zobacz model

Gemini 3.1 Flash Image Preview, a.k.a. "Nano Banana 2," is Google’s latest state of the art image generation and editing model, delivering Pro-level visual quality at Flash speed. It combines…

Kontekst
66K
Wejście
0,50 USD
Wyjście
3 USD
ObrazTekst
Zobacz model

Nano Banana 2 Lite (Gemini 3.1 Flash Lite Image) is Google's fastest, most cost-efficient Gemini image model, built for high-velocity developer pipelines and rapid-fire visual exploration. It delivers text-to-image generation…

Kontekst
66K
Wejście
0,25 USD
Wyjście
1,5 USD
ObrazTekst
Zobacz model

Nano Banana Pro is Google’s most advanced image-generation and editing model, built on Gemini 3 Pro. It extends the original Nano Banana with significantly improved multimodal reasoning, real-world grounding, and…

Kontekst
66K
Wejście
2 USD
Wyjście
12 USD
ObrazTekst
Zobacz model

Nano Banana Pro is Google’s most advanced image-generation and editing model, built on Gemini 3 Pro. It extends the original Nano Banana with significantly improved multimodal reasoning, real-world grounding, and…

Kontekst
66K
Wejście
2 USD
Wyjście
12 USD
ObrazTekst
Zobacz model

NVIDIA Nemotron 3 Embed 1B is an open text embedding model from NVIDIA, optimized for high-throughput, low-latency retrieval. It is suited for enterprise search, RAG, code retrieval, and agentic retrieval…

Kontekst
33K
Wejście
0,00 USD
Wyjście
0,00 USD
Tekst
Zobacz model

Nemotron-3-Nano-30B-A3B-BF16 is a large language model (LLM) trained from scratch by NVIDIA, and designed as a unified model for both reasoning and non-reasoning tasks. It responds to user queries and tasks by first generating a reasoning trace and then concluding with a final response. The model's reasoning capabilities can be configured through a flag in the chat template. If the user prefers the model to provide its final answer without intermediate reasoning traces, it can be configured to do so, albeit with a slight decrease in accuracy for harder prompts that require reasoning. Conversely, allowing the model to generate reasoning traces first generally results in higher-quality final solutions to queries and tasks.

Kontekst
262K
Wejście
0,05 USD
Wyjście
0,20 USD
Tekst
Zobacz model

NVIDIA Nemotron 3 Nano Omni is a multimodal large language model that unifies video, audio, image, and text understanding to support enterprise-grade Q&A, summarization, transcription, and document intelligence workflows. It extends the Nemotron Nano family with integrated video+speech comprehension, Graphical User Interface (GUI), Optical Character Recognition (OCR), and speech transcription capabilities, enabling end-to-end processing of rich enterprise content such as meeting recordings, M&E assets, training videos, and complex business documents. NVIDIA Nemotron 3 Nano Omni was developed by NVIDIA as part of the Nemotron model family.

Kontekst
256K
Wejście
0,00 USD
Wyjście
0,00 USD
TekstDźwiękObrazWideo
Zobacz model
Aktywny

> Use temperature=1.0 and top_p=0.95 across **all tasks and serving backends** — reasoning, tool calling, and general chat alike.

Kontekst
262K
Wejście
0,085 USD
Wyjście
0,40 USD
Tekst
Zobacz model
Aktywny

For more details on how to deploy and use the model - see the Quick Start Guide below!

Kontekst
512K
Wejście
0,60 USD
Wyjście
3,6 USD
Tekst
Zobacz model

Nemotron 3.5 ASR Streaming Multilingual 0.6B is a speech recognition model from NVIDIA. Its prompt-conditioned, cache-aware FastConformer-RNNT design targets low-latency transcription across more than 40 languages for real-time captioning, voice…

Kontekst
0
Wejście
3,33 USD
Wyjście
0,00 USD
Dźwięk
Zobacz model
Aktywny

NVIDIA Nemotron™ is a family of open models with open weights, training data, and recipes, delivering leading efficiency and accuracy for building specialized AI agents.

Kontekst
262K
Wejście
0,08 USD
Wyjście
0,20 USD
Tekst
Zobacz model

NVIDIA: Nemotron Nano 12B 2 VL (free) is included in the AIToolly model catalog.

Kontekst
128K
Wejście
0,00 USD
Wyjście
0,00 USD
ObrazTekstWideo
Zobacz model

NVIDIA-Nemotron-Nano-9B-v2 is a large language model (LLM) trained from scratch by NVIDIA, and designed as a unified model for both reasoning and non-reasoning tasks. It responds to user queries and tasks by first generating a reasoning trace and then concluding with a final response. The model's reasoning capabilities can be controlled via a system prompt. If the user prefers the model to provide its final answer without intermediate reasoning traces, it can be configured to do so, albeit with a slight decrease in accuracy for harder prompts that require reasoning. Conversely, allowing the model to generate reasoning traces first generally results in higher-quality final solutions to queries and tasks.

Kontekst
128K
Wejście
0,00 USD
Wyjście
0,00 USD
Tekst
Zobacz model
Aktywny

🤗 Model &nbsp&nbsp | &nbsp&nbsp 🔀 the provider catalog (Enjoy two weeks free starting June 9!) &nbsp&nbsp | &nbsp&nbsp 💻 Github &nbsp&nbsp | &nbsp&nbsp 🧭 ModelScope &nbsp&nbsp | &nbsp&nbsp 🚀 Nex-AGI

Kontekst
262K
Wejście
0,025 USD
Wyjście
0,10 USD
TekstObraz
Zobacz model

Nex AGI

Nex-N2-Pro

Aktywny

🤗 Model &nbsp&nbsp | &nbsp&nbsp 🔀 the provider catalog (Enjoy two weeks free starting June 9!) &nbsp&nbsp | &nbsp&nbsp 💻 Github &nbsp&nbsp | &nbsp&nbsp 🧭 ModelScope &nbsp&nbsp | &nbsp&nbsp 🚀 Nex-AGI

Kontekst
262K
Wejście
0,25 USD
Wyjście
1 USD
TekstObraz
Zobacz model
Aktywny

North Mini Code is Cohere's first agentic coding model and the debut of its North family. A sparse mixture-of-experts model with 30B total parameters and 3B active, it is optimized…

Kontekst
256K
Wejście
0,00 USD
Wyjście
0,00 USD
Tekst
Zobacz model
Aktywny

Nova 2 Lite is a fast, cost-effective reasoning model for everyday workloads that can process text, images, and videos to generate text. Nova 2 Lite demonstrates standout capabilities in processing…

Kontekst
1M
Wejście
0,30 USD
Wyjście
2,5 USD
TekstObrazWideoPlik
Zobacz model
Aktywny

Amazon Nova Premier is the most capable of Amazon’s multimodal models for complex reasoning tasks and for use as the best teacher for distilling custom models.

Kontekst
1M
Wejście
2,5 USD
Wyjście
12,5 USD
TekstObraz
Zobacz model

Deepgram

Nova-3

Aktywny

Deepgram Nova-3 general-purpose speech-to-text model with monolingual and multilingual transcription support.

Kontekst
0
Wejście
4300 USD
Wyjście
0,00 USD
Dźwięk
Zobacz model
Aktywny

We introduce Olmo 3, a new family of 7B and 32B models both Instruct and Think variants. Long chain-of-thought thinking improves reasoning tasks like math and coding.

Kontekst
66K
Wejście
0,15 USD
Wyjście
0,50 USD
Tekst
Zobacz model
Aktywny

This model always redirects to the latest model in the OpenAI GPT family.

Kontekst
1,05M
Wejście
2,5 USD
Wyjście
15 USD
PlikObrazTekst
Zobacz model

This model always redirects to the latest model in the OpenAI GPT Mini family.

Kontekst
400K
Wejście
0,75 USD
Wyjście
4,5 USD
PlikObrazTekst
Zobacz model

Canopy Labs

Orpheus 3B

Aktywny

Orpheus 3B is an English text-to-speech model from Canopy Labs, fine-tuned for natural prosody and expressive delivery. It offers 7 preset voices and is suited for narration, voice assistants, and…

Kontekst
4K
Wejście
7 USD
Wyjście
0,00 USD
Tekst
Zobacz model
Aktywny

Palmyra X5 is Writer's most advanced model, purpose-built for building and scaling AI agents across the enterprise. It delivers industry-leading speed and efficiency on context windows up to 1 million…

Kontekst
1,04M
Wejście
0,60 USD
Wyjście
6 USD
Tekst
Zobacz model
Aktywny

Parakeet TDT 0.6B v3 is NVIDIA's 600M-parameter multilingual speech-to-text model built on the FastConformer-TDT architecture. Trained on the Granary dataset (670,000+ hours of audio), it supports automatic language detection across…

Kontekst
0
Wejście
1500 USD
Wyjście
0,00 USD
Dźwięk
Zobacz model

Sentence Transformers

paraphrase-MiniLM-L6-v2

Aktywny

Sentence Transformers: paraphrase-MiniLM-L6-v2 is included in the AIToolly model catalog.

Kontekst
512
Wejście
0,005 USD
Wyjście
0,00 USD
Tekst
Zobacz model
Aktywny

The Pareto Router maintains a tiered shortlist of strong coding models, ranked by Artificial Analysis coding percentiles. Set min_coding_score between 0 and 1 on the pareto-router plugin to control how…

Kontekst
2M
Wejście
Nieznane
Wyjście
Nieznane
Tekst
Zobacz model

Perceptron

Perceptron Mk1

Aktywny

Perceptron Mk1 (Mark One) is Perceptron's highest-quality vision-language model for video and embodied reasoning.** It accepts image and video inputs paired with natural language queries, and produces detailed visual understanding…

Kontekst
33K
Wejście
0,15 USD
Wyjście
1,5 USD
TekstObrazWideo
Zobacz model
Aktywny

Qwen Image 3 is a unified image generation and editing model from Qwen. It supports precise rendering of text and details as small as 10px, along with a richer world…

Kontekst
66K
Wejście
0,00 USD
Wyjście
0,00 USD
TekstObraz
Zobacz model
Aktywny

Qwen Image 3 Pro is an image generation and editing model from Qwen. It supports precise rendering of text and details as small as 10px, along with richer world knowledge…

Kontekst
66K
Wejście
0,00 USD
Wyjście
0,00 USD
TekstObraz
Zobacz model
Aktywny

Qwen Plus 0728, based on the Qwen3 foundation model, is a 1 million context hybrid reasoning model with a balanced performance, speed, and cost combination.

Kontekst
1M
Wejście
0,26 USD
Wyjście
0,78 USD
Tekst
Zobacz model

Qwen-Audio-3.0-TTS Flash is Alibaba's fast, cost-efficient text-to-speech model, generating spoken audio from text via the DashScope Speech Synthesizer API.

Kontekst
0
Wejście
15 USD
Wyjście
0,00 USD
Tekst
Zobacz model

Qwen-Audio-3.0-TTS Plus is Alibaba's higher-quality text-to-speech model, generating spoken audio from text via the DashScope Speech Synthesizer API.

Kontekst
0
Wejście
20 USD
Wyjście
0,00 USD
Tekst
Zobacz model

Over the past three months, we have continued to scale the **thinking capability** of Qwen3-30B-A3B, improving both the **quality and depth** of reasoning. We are pleased to introduce **Qwen3-30B-A3B-Thinking-2507**, featuring the following key enhancements:

Kontekst
82K
Wejście
0,20 USD
Wyjście
2,4 USD
Tekst
Zobacz model
Aktywny

The Qwen3-ASR family includes Qwen3-ASR-1.7B and Qwen3-ASR-0.6B, which support language identification and ASR for 52 languages and dialects. Both leverage large-scale speech training data and the strong audio understanding capability of their foundation model, Qwen3-Omni. Experiments show that the 1.7B version achieves state-of-the-art performance among open-source ASR models and is competitive with the strongest proprietary commercial APIs. Here are the main features:

Kontekst
0
Wejście
3,33 USD
Wyjście
0,00 USD
Dźwięk
Zobacz model
Aktywny

The Qwen3-ASR family includes Qwen3-ASR-1.7B and Qwen3-ASR-0.6B, which support language identification and ASR for 52 languages and dialects. Both leverage large-scale speech training data and the strong audio understanding capability of their foundation model, Qwen3-Omni. Experiments show that the 1.7B version achieves state-of-the-art performance among open-source ASR models and is competitive with the strongest proprietary commercial APIs. Here are the main features:

Kontekst
0
Wejście
7,5 USD
Wyjście
0,00 USD
Dźwięk
Zobacz model
Aktywny

Qwen3-ASR-Flash is Alibaba's automatic speech recognition service, built on the Qwen3-Omni foundation and trained on tens of millions of hours of multimodal speech data. The model handles 11 languages —…

Kontekst
0
Wejście
35 USD
Wyjście
0,00 USD
Dźwięk
Zobacz model
Aktywny

Qwen3 Coder Flash is Alibaba's fast and cost efficient version of their proprietary Qwen3 Coder Plus. It is a powerful coding agent model specializing in autonomous programming via tool calling…

Kontekst
1M
Wejście
0,195 USD
Wyjście
0,975 USD
Tekst
Zobacz model
Aktywny

Qwen3 Coder Plus is Alibaba's proprietary version of the Open Source Qwen3 Coder 480B A35B. It is a powerful coding agent model specializing in autonomous programming via tool calling and…

Kontekst
1M
Wejście
0,65 USD
Wyjście
3,25 USD
Tekst
Zobacz model
Aktywny

The Qwen3 Embedding model series is the latest proprietary model of the Qwen family, specifically designed for text embedding and ranking tasks. Building upon the dense foundational models of the Qwen3 series, it provides a comprehensive range of text embeddings and reranking models in various sizes (0.6B, 4B, and 8B). This series inherits the exceptional multilingual capabilities, long-text understanding, and reasoning skills of its foundational model. The Qwen3 Embedding series represents significant advancements in multiple text embedding and ranking tasks, including text retrieval, code retrieval, text classification, text clustering, and bitext mining.

Kontekst
33K
Wejście
0,02 USD
Wyjście
0,00 USD
Tekst
Zobacz model
Aktywny

The Qwen3 Embedding model series is the latest proprietary model of the Qwen family, specifically designed for text embedding and ranking tasks. Building upon the dense foundational models of the Qwen3 series, it provides a comprehensive range of text embeddings and reranking models in various sizes (0.6B, 4B, and 8B). This series inherits the exceptional multilingual capabilities, long-text understanding, and reasoning skills of its foundational model. The Qwen3 Embedding series represents significant advancements in multiple text embedding and ranking tasks, including text retrieval, code retrieval, text classification, text clustering, and bitext mining.

Kontekst
33K
Wejście
0,01 USD
Wyjście
0,00 USD
Tekst
Zobacz model
Aktywny

Qwen3-Max is an updated release built on the Qwen3 series, offering major improvements in reasoning, instruction following, multilingual support, and long-tail knowledge coverage compared to the January 2025 version. It…

Kontekst
262K
Wejście
0,78 USD
Wyjście
3,9 USD
Tekst
Zobacz model
Aktywny

Qwen3-Max-Thinking is the flagship reasoning model in the Qwen3 series, designed for high-stakes cognitive tasks that require deep, multi-step reasoning. By significantly scaling model capacity and reinforcement learning compute, it…

Kontekst
262K
Wejście
0,78 USD
Wyjście
3,9 USD
Tekst
Zobacz model

Over the past few months, we have observed increasingly clear trends toward scaling both total parameters and context lengths in the pursuit of more powerful and agentic artificial intelligence (AI). We are excited to share our latest advancements in addressing these demands, centered on improving scaling efficiency through innovative model architecture. We call this next-generation foundation models **Qwen3-Next**.

Kontekst
262K
Wejście
0,10 USD
Wyjście
1,1 USD
Tekst
Zobacz model

Over the past few months, we have observed increasingly clear trends toward scaling both total parameters and context lengths in the pursuit of more powerful and agentic artificial intelligence (AI). We are excited to share our latest advancements in addressing these demands, centered on improving scaling efficiency through innovative model architecture. We call this next-generation foundation models **Qwen3-Next**.

Kontekst
131K
Wejście
0,15 USD
Wyjście
1,2 USD
Tekst
Zobacz model
Aktywny

Qwen3 Reranker 8B is a text reranking model from Alibaba Cloud built on the Qwen3 architecture. It evaluates query-document pairs to produce relevance scores for use in retrieval and RAG…

Kontekst
41K
Wejście
0,00 USD
Wyjście
0,00 USD
Tekst
Zobacz model

Meet Qwen3-VL — the most powerful vision-language model in the Qwen series to date.

Kontekst
131K
Wejście
0,21 USD
Wyjście
1,9 USD
TekstObraz
Zobacz model

Meet Qwen3-VL — the most powerful vision-language model in the Qwen series to date.

Kontekst
131K
Wejście
0,40 USD
Wyjście
4 USD
TekstObraz
Zobacz model

Meet Qwen3-VL — the most powerful vision-language model in the Qwen series to date.

Kontekst
131K
Wejście
0,13 USD
Wyjście
0,52 USD
TekstObraz
Zobacz model

Meet Qwen3-VL — the most powerful vision-language model in the Qwen series to date.

Kontekst
131K
Wejście
0,20 USD
Wyjście
2,4 USD
TekstObraz
Zobacz model

Meet Qwen3-VL — the most powerful vision-language model in the Qwen series to date.

Kontekst
131K
Wejście
0,104 USD
Wyjście
0,416 USD
TekstObraz
Zobacz model

Meet Qwen3-VL — the most powerful vision-language model in the Qwen series to date.

Kontekst
131K
Wejście
0,117 USD
Wyjście
0,455 USD
ObrazTekst
Zobacz model

Meet Qwen3-VL — the most powerful vision-language model in the Qwen series to date.

Kontekst
131K
Wejście
0,18 USD
Wyjście
2,1 USD
ObrazTekst
Zobacz model
Aktywny

> [!Note] > This repository contains model weights and configuration files for the post-trained model in the Hugging Face Transformers format. > > These artifacts are compatible with Hugging Face Transformers, vLLM, SGLang, KTransformers, etc.

Kontekst
262K
Wejście
0,39 USD
Wyjście
2,34 USD
TekstObrazWideo
Zobacz model

The Qwen3.5 native vision-language series Plus models are built on a hybrid architecture that integrates linear attention mechanisms with sparse mixture-of-experts models, achieving higher inference efficiency. In a variety of…

Kontekst
1M
Wejście
0,26 USD
Wyjście
1,56 USD
TekstObrazWideo
Zobacz model

Qwen3.5 Plus (April 2026) is a large-scale multimodal language model from Alibaba. It accepts text, image, and video input and produces text output, with a 1M token context window. This…

Kontekst
1M
Wejście
0,30 USD
Wyjście
1,8 USD
TekstObrazWideo
Zobacz model
Aktywny

> [!Note] > This repository contains model weights and configuration files for the post-trained model in the Hugging Face Transformers format. > > These artifacts are compatible with Hugging Face Transformers, vLLM, SGLang, KTransformers, etc.

Kontekst
262K
Wejście
0,26 USD
Wyjście
2,08 USD
TekstObrazWideo
Zobacz model
Aktywny

> [!Note] > This repository contains model weights and configuration files for the post-trained model in the Hugging Face Transformers format. > > These artifacts are compatible with Hugging Face Transformers, vLLM, SGLang, KTransformers, etc.

Kontekst
262K
Wejście
0,195 USD
Wyjście
1,56 USD
TekstObrazWideo
Zobacz model
Aktywny

> [!Note] > This repository contains model weights and configuration files for the post-trained model in the Hugging Face Transformers format. > > These artifacts are compatible with Hugging Face Transformers, vLLM, SGLang, KTransformers, etc.

Kontekst
262K
Wejście
0,225 USD
Wyjście
1,8 USD
TekstObrazWideo
Zobacz model
Aktywny

> [!Note] > This repository contains model weights and configuration files for the post-trained model in the Hugging Face Transformers format. > > These artifacts are compatible with Hugging Face Transformers, vLLM, SGLang, KTransformers, etc.

Kontekst
262K
Wejście
0,10 USD
Wyjście
0,15 USD
TekstObrazWideo
Zobacz model
Aktywny

The Qwen3.5 native vision-language Flash models are built on a hybrid architecture that integrates a linear attention mechanism with a sparse mixture-of-experts model, achieving higher inference efficiency. Compared to the…

Kontekst
1M
Wejście
0,065 USD
Wyjście
0,26 USD
TekstObrazWideo
Zobacz model
Aktywny

> [!Note] > This repository contains model weights and configuration files for the post-trained model in the Hugging Face Transformers format. > > These artifacts are compatible with Hugging Face Transformers, vLLM, SGLang, KTransformers, etc.

Kontekst
262K
Wejście
0,60 USD
Wyjście
3,6 USD
TekstObrazWideo
Zobacz model
Aktywny

> [!Note] > This repository contains model weights and configuration files for the post-trained model in the Hugging Face Transformers format. > > These artifacts are compatible with Hugging Face Transformers, vLLM, SGLang, KTransformers, etc.

Kontekst
262K
Wejście
0,14 USD
Wyjście
1 USD
TekstObrazWideo
Zobacz model
Aktywny

Qwen3.6 Flash is a fast, efficient language model from Alibaba's Qwen 3.6 series. It supports text, image, and video input with a 1M token context window. Tiered pricing kicks in…

Kontekst
1M
Wejście
0,1875 USD
Wyjście
1,13 USD
TekstObrazWideo
Zobacz model
Aktywny

Qwen3.6-Max-Preview is a proprietary frontier model from Alibaba Cloud built on a sparse mixture-of-experts architecture with approximately 1 trillion total parameters. It is optimized for agentic coding, tool use, and…

Kontekst
262K
Wejście
1,03 USD
Wyjście
6,16 USD
Tekst
Zobacz model
Aktywny

Qwen 3.6 Plus builds on a hybrid architecture that combines efficient linear attention with sparse mixture-of-experts routing, enabling strong scalability and high-performance inference. Compared to the 3.5 series, it delivers…

Kontekst
1M
Wejście
0,325 USD
Wyjście
1,95 USD
TekstObrazWideo
Zobacz model
Aktywny

Qwen3.7 Flash is a vision-language reasoning model from Alibaba. It is suited for multimodal agents, visual coding, search, and computer interaction, with strengths in object recognition, spatial understanding, and real-world…

Kontekst
1M
Wejście
0,03 USD
Wyjście
0,13 USD
TekstObrazWideo
Zobacz model
Aktywny

Qwen3.7-Max is the flagship model in Alibaba's Qwen3.7 series. It supports text input and output and is designed for agent-centric workloads, with particular strengths in coding, office and productivity tasks,…

Kontekst
1M
Wejście
1,48 USD
Wyjście
4,43 USD
Tekst
Zobacz model
Aktywny

Qwen3.7-Plus is a cost-effective model in Alibaba's Qwen3.7 series. It supports text and image input with text output, building on the series' text capabilities with a comprehensive upgrade to its…

Kontekst
1M
Wejście
0,32 USD
Wyjście
1,28 USD
TekstObraz
Zobacz model
Aktywny

> [!Note] > This repository contains model weights and configuration files for the post-trained model in the Hugging Face Transformers format. > > These artifacts are compatible with vLLM, SGLang, TokenSpeed, etc.

Kontekst
1M
Wejście
2 USD
Wyjście
6 USD
Tekst
Zobacz model
Aktywny

Qwen3.8 Max is the flagship model in Alibaba's Qwen3.8 series, the general-availability successor to the Qwen3.8 Max Preview. It is a multimodal reasoning model intended for complex reasoning, visual understanding,…

Kontekst
1M
Wejście
2 USD
Wyjście
6 USD
TekstObrazWideo
Zobacz model

Recraft

Recraft V3

Aktywny

Recraft V3 is an image generation model from Recraft. It supports text and image inputs with image output at ~1K resolution across multiple aspect ratios. Supports the following image_config parameters:…

Kontekst
66K
Wejście
0,00 USD
Wyjście
0,00 USD
TekstObraz
Zobacz model

Recraft

Recraft V4

Aktywny

Recraft V4 is an image generation model from Recraft. It supports text and image inputs with image output at ~1K resolution across multiple aspect ratios. It delivers stronger compositional judgment,…

Kontekst
66K
Wejście
0,00 USD
Wyjście
0,00 USD
TekstObraz
Zobacz model
Aktywny

Recraft V4 Pro is an image generation model from Recraft. It supports text and image inputs with image output at ~2K resolution across multiple aspect ratios, double the resolution of…

Kontekst
66K
Wejście
0,00 USD
Wyjście
0,00 USD
TekstObraz
Zobacz model
Aktywny

Recraft V4 Pro Vector is the vector (SVG) variant of Recraft V4 Pro. It supports text and image inputs and produces vector image output across multiple aspect ratios at the…

Kontekst
66K
Wejście
0,00 USD
Wyjście
0,00 USD
TekstObraz
Zobacz model
Aktywny

Recraft V4 Vector is the vector (SVG) variant of Recraft V4. It supports text and image inputs and produces vector image output across multiple aspect ratios. Compared to the raster…

Kontekst
66K
Wejście
0,00 USD
Wyjście
0,00 USD
TekstObraz
Zobacz model
Aktywny

Recraft V4.1 is an image generation model from Recraft tuned for high aesthetics. It supports text and image inputs with image output at ~1K resolution across multiple aspect ratios, with…

Kontekst
66K
Wejście
0,00 USD
Wyjście
0,00 USD
TekstObraz
Zobacz model
Aktywny

Recraft V4.1 Pro is an image generation model from Recraft tuned for high aesthetics. It supports text and image inputs with image output at ~2K resolution across multiple aspect ratios…

Kontekst
66K
Wejście
0,00 USD
Wyjście
0,00 USD
TekstObraz
Zobacz model
Aktywny

Recraft V4.1 Pro Vector is the vector (SVG) variant of Recraft V4.1 Pro, tuned for high aesthetics. It supports text and image inputs and produces higher-resolution SVG image output across…

Kontekst
66K
Wejście
0,00 USD
Wyjście
0,00 USD
TekstObraz
Zobacz model
Aktywny

Recraft V4.1 Utility is a general-purpose image generation model from Recraft. It supports text and image inputs with image output at ~1K resolution across multiple aspect ratios, with typical generation…

Kontekst
66K
Wejście
0,00 USD
Wyjście
0,00 USD
TekstObraz
Zobacz model
Aktywny

Recraft V4.1 Utility Pro is a general-purpose image generation model from Recraft. It supports text and image inputs with image output at ~2K resolution across multiple aspect ratios — double…

Kontekst
66K
Wejście
0,00 USD
Wyjście
0,00 USD
TekstObraz
Zobacz model
Aktywny

Recraft V4.1 Vector is the vector (SVG) variant of Recraft V4.1, tuned for high aesthetics. It supports text and image inputs and produces SVG image output across multiple aspect ratios,…

Kontekst
66K
Wejście
0,00 USD
Wyjście
0,00 USD
TekstObraz
Zobacz model
Aktywny

**Reka Edge** is an extremely efficient 7B multimodal vision-language model that accepts image/video+text inputs and generates text outputs. This model is optimized specifically to deliver industry-leading performance in image understanding, video analysis, object detection, and agentic tool-use.

Kontekst
16K
Wejście
0,10 USD
Wyjście
0,10 USD
ObrazTekstWideo
Zobacz model
Aktywny

Relace Apply 3 is a specialized code-patching LLM that merges AI-suggested edits straight into your source files. It can apply updates from GPT-4o, Claude, and others into your files at…

Kontekst
256K
Wejście
0,85 USD
Wyjście
1,25 USD
Tekst
Zobacz model
Aktywny

The relace-search model uses 4-12 view_file and grep tools in parallel to explore a codebase and return relevant files to the user request. In contrast to RAG, relace-search performs agentic…

Kontekst
256K
Wejście
1 USD
Wyjście
3 USD
Tekst
Zobacz model
Aktywny

Cohere's AI search foundation model for enhancing the relevance of information surfaced within search and RAG systems. Features a 32K context window, multilingual support across 100+ languages, no data pre-processing…

Kontekst
33K
Wejście
0,00 USD
Wyjście
0,00 USD
Tekst
Zobacz model
Aktywny

Cohere's AI search foundation model for enhancing the relevance of information surfaced within search and RAG systems. Features a 32K context window, multilingual support across 100+ languages, no data pre-processing…

Kontekst
33K
Wejście
0,00 USD
Wyjście
0,00 USD
Tekst
Zobacz model
Aktywny

Rerank v3.5 is designed to reorder search results for improved relevance. It supports multi-aspect and semi-structured data reranking over 100+ languages. Ideal for refining results from semantic or keyword search…

Kontekst
4K
Wejście
0,00 USD
Wyjście
0,00 USD
Tekst
Zobacz model

VoyageAI by MongoDB

rerank-2.5

Aktywny

rerank-2.5 is a cutting-edge reranker optimized for quality, delivering a 7.94% improvement in retrieval accuracy over Cohere Rerank v3.5 across 93 datasets. It also outperformed Cohere Rerank v3.5 by 12.70%…

Kontekst
32K
Wejście
0,00 USD
Wyjście
0,00 USD
Tekst
Zobacz model

VoyageAI by MongoDB

rerank-2.5-lite

Aktywny

rerank-2.5-lite is a reranker optimized for both latency and quality, delivering a 7.16% improvement in retrieval accuracy over Cohere Rerank v3.5 across 93 datasets. It also outperformed Cohere Rerank v3.5…

Kontekst
32K
Wejście
0,00 USD
Wyjście
0,00 USD
Tekst
Zobacz model

inclusionAI

Ring-2.6-1T

Aktywny

Ring-2.6-1T is a 1T-parameter-scale thinking model with 63B active parameters, built for real-world agent workflows that require both strong capability and operational efficiency. It is optimized for coding agents, tool…

Kontekst
262K
Wejście
0,075 USD
Wyjście
0,625 USD
Tekst
Zobacz model
Aktywny

Riverflow V2 Fast is the fastest variant of Sourceful's Riverflow 2.0 lineup, best for production deployments and latency-critical workflows. The Riverflow 2.0 series represents SOTA performance on image generation and…

Kontekst
8K
Wejście
0,00 USD
Wyjście
0,00 USD
TekstObraz
Zobacz model
Aktywny

Riverflow V2 Pro is the most powerful variant of Sourceful's Riverflow 2.0 lineup, best for top-tier control and perfect text rendering. The Riverflow 2.0 series represents SOTA performance on image…

Kontekst
8K
Wejście
0,00 USD
Wyjście
0,00 USD
TekstObraz
Zobacz model
Aktywny

Riverflow V2.5 Fast is the speed-optimized variant of Sourceful's Riverflow 2.5 lineup, best for production deployments and latency-critical workflows. The Riverflow 2.5 series is a unified text-to-image and image-to-image family…

Kontekst
33K
Wejście
0,00 USD
Wyjście
0,00 USD
TekstObraz
Zobacz model
Aktywny

Riverflow V2.5 Pro is the most powerful variant of Sourceful's Riverflow 2.5 lineup, best for top-tier control and quality-sensitive outputs. The Riverflow 2.5 series is a unified text-to-image and image-to-image…

Kontekst
33K
Wejście
0,00 USD
Wyjście
0,00 USD
TekstObraz
Zobacz model

Fish Audio

S1

Aktywny

S1 is a multilingual text-to-speech model from Fish Audio. It is suited for voice applications that need broad emotional expression, using parenthetical controls to guide speaking style across its supported…

Kontekst
0
Wejście
15 USD
Wyjście
0,00 USD
Tekst
Zobacz model

Fish Audio

S2 Pro

Aktywny

S2 Pro is a multilingual text-to-speech model from Fish Audio. It is suited for expressive narration and multi-speaker dialogue, with natural-language controls for speaking style and emotion.

Kontekst
0
Wejście
15 USD
Wyjście
0,00 USD
Tekst
Zobacz model

Fish Audio

S2.1 Pro

Aktywny

S2.1 Pro is a production-oriented text-to-speech model from Fish Audio. It is suited for multilingual voice applications, expressive narration, and dialogue synthesis, with open-ended natural-language controls for speaking style and…

Kontekst
0
Wejście
15 USD
Wyjście
0,00 USD
Tekst
Zobacz model
Aktywny

S2.1 Pro Free is the no-cost variant of Fish Audio S2.1 Pro, intended for testing, prototyping, and low-volume applications. It provides the same synthesis capabilities without production latency or availability…

Kontekst
0
Wejście
0,00 USD
Wyjście
0,00 USD
Tekst
Zobacz model
Aktywny

Sakana Namazu is a Japanese-specialized reasoning model from Sakana AI, based on Kimi K2.6 with additional training for Japanese language and business contexts. It is suited for Japanese instruction following,…

Kontekst
262K
Wejście
0,95 USD
Wyjście
4 USD
TekstObrazPlik
Zobacz model

ByteDance Seed

Seed 1.6

Aktywny

Seed 1.6 is a general-purpose model released by the ByteDance Seed team. It incorporates multimodal capabilities and adaptive deep thinking with a 256K context window.

Kontekst
262K
Wejście
0,25 USD
Wyjście
2 USD
ObrazTekstWideo
Zobacz model

ByteDance Seed

Seed 1.6 Flash

Aktywny

Seed 1.6 Flash is an ultra-fast multimodal deep thinking model by ByteDance Seed, supporting both text and visual understanding. It features a 256k context window and can generate outputs of…

Kontekst
262K
Wejście
0,075 USD
Wyjście
0,30 USD
ObrazTekstWideo
Zobacz model

ByteDance Seed

Seed 2.1 Turbo

Aktywny

Seed 2.1 Turbo is a multimodal model from ByteDance Seed for coding and long-horizon agent workflows. It is suited for end-to-end software delivery, multi-step task execution, and understanding visual and…

Kontekst
262K
Wejście
0,50 USD
Wyjście
2,5 USD
TekstObrazWideo
Zobacz model

ByteDance Seed

Seed-2.0-Code

Aktywny

Seed 2.0 Code is a model from ByteDance Seed optimized for agentic coding. It is suited for frontend development, multilingual programming tasks, and coding-agent workflows in tools such as Claude…

Kontekst
262K
Wejście
0,50 USD
Wyjście
3 USD
TekstObrazWideo
Zobacz model

ByteDance Seed

Seed-2.0-Lite

Aktywny

Seed-2.0-Lite is a versatile, cost‑efficient enterprise workhorse that delivers strong multimodal and agent capabilities while offering noticeably lower latency, making it a practical default choice for most production workloads across…

Kontekst
262K
Wejście
0,25 USD
Wyjście
2 USD
TekstObrazWideo
Zobacz model

ByteDance Seed

Seed-2.0-Mini

Aktywny

Seed-2.0-mini targets latency-sensitive, high-concurrency, and cost-sensitive scenarios, emphasizing fast response and flexible inference deployment. It delivers performance comparable to ByteDance-Seed-1.6, supports 256k context, four reasoning effort modes (minimal/low/medium/high), multimodal understanding,…

Kontekst
262K
Wejście
0,10 USD
Wyjście
0,40 USD
TekstObrazWideo
Zobacz model
Aktywny

ByteDance's next-generation audio-visual generation model with a 4.5B parameter Dual-Branch Diffusion Transformer architecture. Seedance 1.5 Pro generates video and audio simultaneously in a single unified pass — eliminating the timing…

Kontekst
0
Wejście
0,00 USD
Wyjście
0,00 USD
TekstObraz
Zobacz model

ByteDance

Seedance 2.0

Aktywny

Seedance 2.0 is a video generation model from ByteDance. It supports text-to-video, image-to-video with first and last frame control, and multimodal reference-to-video. It is particularly strong at preserving character consistency,…

Kontekst
0
Wejście
0,00 USD
Wyjście
0,00 USD
TekstObrazWideoDźwięk
Zobacz model
Aktywny

Seedance 2.0 Fast is a video generation model from ByteDance. It supports text-to-video, image-to-video with first and last frame control, and multimodal reference-to-video. It prioritizes generation speed and lower cost…

Kontekst
0
Wejście
0,00 USD
Wyjście
0,00 USD
TekstObrazWideoDźwięk
Zobacz model
Aktywny

Seedance 2.0 Mini is a video generation model from ByteDance. It supports text-to-video, image-to-video with first and last frame control, and multimodal reference-to-video with image, video, and audio inputs. It…

Kontekst
0
Wejście
0,00 USD
Wyjście
0,00 USD
TekstObrazWideoDźwięk
Zobacz model

ByteDance

Seedance 2.5

Aktywny

Seedance 2.5 is a video generation model from ByteDance. It is suited for long-form storytelling, multimodal reference-based generation, video editing, and video extension. It supports first-frame and first-and-last-frame control, up…

Kontekst
0
Wejście
0,00 USD
Wyjście
0,00 USD
TekstObrazWideoDźwięk
Zobacz model

ByteDance Seed

Seedream 4.5

Aktywny

Seedream 4.5 is the latest in-house image generation model developed by ByteDance. Compared with Seedream 4.0, it delivers comprehensive improvements, especially in editing consistency, including better preservation of subject details,…

Kontekst
4K
Wejście
0,00 USD
Wyjście
0,00 USD
ObrazTekst
Zobacz model

ByteDance Seed

Seedream 5.0 Lite

Aktywny

Seedream 5.0 Lite is an image generation model from ByteDance Seed. It is suited for professional visual creation that benefits from web-connected retrieval, complex-prompt comprehension, visual references, and broad knowledge…

Kontekst
0
Wejście
0,00 USD
Wyjście
0,00 USD
TekstObraz
Zobacz model

ByteDance Seed

Seedream 5.0 Pro

Aktywny

Seedream 5.0 Pro is an image generation and editing model from ByteDance Seed. It is suited for commercial visual-production workflows that require precise editing control, lifelike scenes, and natural rendering.

Kontekst
0
Wejście
0,00 USD
Wyjście
0,00 USD
TekstObraz
Zobacz model
Aktywny

Solar Pro 3 is Upstage's powerful Mixture-of-Experts (MoE) language model. With 102B total parameters and 12B active parameters per forward pass, it delivers exceptional performance while maintaining computational efficiency. Optimized…

Kontekst
131K
Wejście
0,15 USD
Wyjście
0,60 USD
Tekst
Zobacz model
Aktywny

Solar Pro 4 is Upstage's cost-efficient large language model, featuring a 524K context window. It is built for long-horizon tasks and agentic workflows, with strong capabilities in office productivity, document-intensive…

Kontekst
524K
Wejście
0,03 USD
Wyjście
0,12 USD
Tekst
Zobacz model
Aktywny

Exclusively available on the the provider catalog API, Sonar Pro's new Pro Search mode is Perplexity's most advanced agentic search system. It is designed for deeper reasoning and analysis. Pricing is based…

Kontekst
200K
Wejście
3 USD
Wyjście
15 USD
TekstObraz
Zobacz model
Aktywny

OpenAI's flagship video generation model, delivering production-quality video with physics-accurate motion, synchronized audio, and world-state persistence across shots. Sora 2 Pro follows intricate multi-shot instructions while maintaining consistent spatial relationships…

Kontekst
0
Wejście
0,00 USD
Wyjście
0,00 USD
TekstObraz
Zobacz model
Aktywny

MiniMax Speech 2.8 HD is a text-to-speech model from MiniMax. It is suited for applications that generate spoken audio from text and accepts arbitrary MiniMax voice IDs.

Kontekst
0
Wejście
100 USD
Wyjście
0,00 USD
Tekst
Zobacz model
Aktywny

MiniMax Speech 2.8 Turbo is a text-to-speech model from MiniMax. It is suited for applications that generate spoken audio from text and accepts arbitrary MiniMax voice IDs.

Kontekst
0
Wejście
60 USD
Wyjście
0,00 USD
Tekst
Zobacz model
Aktywny

Step 3.5 Flash is StepFun's most capable open-source foundation model. Built on a sparse Mixture of Experts (MoE) architecture, it selectively activates only 11B of its 196B parameters per token.…

Kontekst
262K
Wejście
0,10 USD
Wyjście
0,30 USD
Tekst
Zobacz model
Aktywny

Step 3.7 Flash is StepFun's latest high-efficiency multimodal Mixture-of-Experts model. It pairs a 196B-parameter language backbone with a vision encoder for native image and video understanding, activating roughly 11B parameters…

Kontekst
256K
Wejście
0,20 USD
Wyjście
1,15 USD
TekstObrazWideo
Zobacz model
Aktywny

text-embedding-3-large is OpenAI's most capable embedding model for both english and non-english tasks. Embeddings are a numerical representation of text that can be used to measure the relatedness between two…

Kontekst
8K
Wejście
0,13 USD
Wyjście
0,00 USD
Tekst
Zobacz model
Aktywny

text-embedding-3-small is OpenAI's improved, more performant version of the ada embedding model. Embeddings are a numerical representation of text that can be used to measure the relatedness between two pieces…

Kontekst
8K
Wejście
0,02 USD
Wyjście
0,00 USD
Tekst
Zobacz model
Aktywny

text-embedding-ada-002 is OpenAI's legacy text embedding model.

Kontekst
8K
Wejście
0,10 USD
Wyjście
0,00 USD
Tekst
Zobacz model

Fish Audio

Transcribe 1

Aktywny

Transcribe 1 is a speech-to-text model from Fish Audio. It is suited for audio transcription with automatic language detection and can return timestamped word-level segments when alignment details are requested.

Kontekst
0
Wejście
100 USD
Wyjście
0,00 USD
Dźwięk
Zobacz model
Aktywny

Trinity-Large-Thinking is a reasoning-optimized variant of Arcee AI's Trinity-Large family — a 398B-parameter sparse Mixture-of-Experts (MoE) model with approximately 13B active parameters per token. Built on Trinity-Large-Base and post-trained with extended chain-of-thought reasoning and agentic RL, Trinity-Large-Thinking delivers state-of-the-art performance on agentic benchmarks while maintaining strong general capabilities.

Kontekst
262K
Wejście
0,22 USD
Wyjście
0,85 USD
Tekst
Zobacz model

Google

Veo 3.1

Aktywny

Google's state-of-the-art video generation model, built for maximum visual fidelity in final production cuts. Veo 3.1 generates high-quality 1080p video from text or image prompts with native synchronized audio —…

Kontekst
0
Wejście
0,00 USD
Wyjście
0,00 USD
TekstObraz
Zobacz model
Aktywny

Google's mid-tier video generation model balancing speed and quality. Veo 3.1 Fast generates high-quality video from text or image prompts with native synchronized audio, offering faster turnaround than Veo 3.1…

Kontekst
0
Wejście
0,00 USD
Wyjście
0,00 USD
TekstObraz
Zobacz model
Aktywny

Google's most cost-effective video generation model, designed for high-volume applications and rapid iteration. Veo 3.1 Lite generates 720p and 1080p video from text or image prompts with native synchronized audio…

Kontekst
0
Wejście
0,00 USD
Wyjście
0,00 USD
TekstObraz
Zobacz model
Aktywny

Kling Video O1 is a video generation model from Kuaishou. It supports text and image inputs with video output, enabling text-to-video and image-to-video workflows. It is suited for cinematic content…

Kontekst
0
Wejście
0,00 USD
Wyjście
0,00 USD
TekstObraz
Zobacz model
Aktywny

Kling v3.0 Pro is Kuaishou's premium video generation model, offering higher visual quality than the Standard tier. It supports text-to-video and image-to-video workflows, with first-frame and last-frame control for precise…

Kontekst
0
Wejście
0,00 USD
Wyjście
0,00 USD
TekstObraz
Zobacz model
Aktywny

Kling v3.0 Standard is a video generation model from Kuaishou. It supports text-to-video and image-to-video workflows, with first-frame and last-frame control for guided scene composition. Clips range from 3 to…

Kontekst
0
Wejście
0,00 USD
Wyjście
0,00 USD
TekstObraz
Zobacz model
Aktywny

Voxtral Mini is an enhancement of Ministral 3B, incorporating state-of-the-art audio input capabilities while retaining best-in-class text performance. It excels at speech transcription, translation and audio understanding.

Kontekst
0
Wejście
16,67 USD
Wyjście
0,00 USD
Dźwięk
Zobacz model
Aktywny

Voxtral Mini Transcribe is Mistral's speech-to-text model, derived from the Voxtral Mini family. It accepts audio input and returns transcribed text via the standard transcription API. Suited for transcribing meetings,…

Kontekst
0
Wejście
3000 USD
Wyjście
0,00 USD
Dźwięk
Zobacz model
Aktywny

Voxtral Mini TTS is Mistral's text-to-speech model featuring zero-shot voice cloning and multilingual support. It converts text input into natural-sounding audio output.

Kontekst
4K
Wejście
16 USD
Wyjście
0,00 USD
Tekst
Zobacz model
Aktywny

Voxtral Small is an enhancement of Mistral Small 3, incorporating state-of-the-art audio input capabilities while retaining best-in-class text performance. It excels at speech transcription, translation and audio understanding.

Kontekst
32K
Wejście
0,10 USD
Wyjście
0,30 USD
TekstDźwiękPlik
Zobacz model

Voxtral Small is an enhancement of Mistral Small 3, incorporating state-of-the-art audio input capabilities while retaining best-in-class text performance. It excels at speech transcription, translation and audio understanding.

Kontekst
0
Wejście
50 USD
Wyjście
0,00 USD
Dźwięk
Zobacz model

VoyageAI by MongoDB

voyage-4

Aktywny

voyage-4 is a general-purpose (including multilingual) embedding model optimized for retrieval/search and AI applications. voyage-4 supports embeddings in 2048, 1024, 512, and 256 dimensions, with multiple quantization options. Learn more…

Kontekst
32K
Wejście
0,06 USD
Wyjście
0,00 USD
Tekst
Zobacz model

VoyageAI by MongoDB

voyage-4-large

Aktywny

voyage-4-large is a state-of-the-art general-purpose and multilingual embedding optimized for retrieval quality. Enabled by Matryoshka learning and quantization-aware training, voyage-4-large supports embeddings in 2048, 1024, 512, and 256 dimensions, with…

Kontekst
32K
Wejście
0,12 USD
Wyjście
0,00 USD
Tekst
Zobacz model

VoyageAI by MongoDB

voyage-4-lite

Aktywny

voyage-4-lite is a lightweight, general-purpose embedding model optimized for low latency and cost. Enabled by Matryoshka learning and quantization-aware training, voyage-4-lite supports embeddings in 2048, 1024, 512, and 256 dimensions,…

Kontekst
32K
Wejście
0,02 USD
Wyjście
0,00 USD
Tekst
Zobacz model

VoyageAI by MongoDB

voyage-code-4

Aktywny

voyage-code-4 is a code embedding model from Voyage AI, a MongoDB company. It is designed for coding agents and code retrieval, with Matryoshka embeddings at 2048, 1024, 512, and 256…

Kontekst
32K
Wejście
0,12 USD
Wyjście
0,00 USD
Tekst
Zobacz model

VoyageAI by MongoDB

voyage-multimodal-3.5

Aktywny

voyage-multimodal-3.5 is a state-of-the-art multimodal embedding model capable of vectorizing not only text, images, and video individually, but also content that interleaves all three modalities. It delivers excellent performance for…

Kontekst
32K
Wejście
0,12 USD
Wyjście
0,00 USD
TekstObraz
Zobacz model

Alibaba

Wan 2.6

Aktywny

Alibaba's most advanced video generation model, supporting over 10 visual creation capabilities in a unified system. Wan 2.6 generates 1080p video at 24fps from text, images, reference videos, or audio,…

Kontekst
0
Wejście
0,00 USD
Wyjście
0,00 USD
TekstObraz
Zobacz model

Alibaba

Wan 2.7

Aktywny

Wan 2.7 is a video generation model from Alibaba. It supports text-to-video, image-to-video with first and last frame control, and reference-to-video, where multiple reference images guide the style and content…

Kontekst
0
Wejście
0,00 USD
Wyjście
0,00 USD
TekstObraz
Zobacz model

OpenAI

Whisper 1

Aktywny

Whisper is OpenAI's open-source automatic speech recognition model, available via API as whisper-1. It supports transcription and translation across 50+ languages from audio files up to 25 MB. Accepts formats…

Kontekst
0
Wejście
6000 USD
Wyjście
0,00 USD
Dźwięk
Zobacz model
Aktywny

Whisper is a state-of-the-art model for automatic speech recognition (ASR) and speech translation, proposed in the paper Robust Speech Recognition via Large-Scale Weak Supervision by Alec Radford et al. from OpenAI. Trained on >5M hours of labeled data, Whisper demonstrates a strong ability to generalise to many datasets and domains in a zero-shot setting.

Kontekst
0
Wejście
7,5 USD
Wyjście
0,00 USD
Dźwięk
Zobacz model
Aktywny

Whisper is a state-of-the-art model for automatic speech recognition (ASR) and speech translation, proposed in the paper Robust Speech Recognition via Large-Scale Weak Supervision by Alec Radford et al. from OpenAI. Trained on >5M hours of labeled data, Whisper demonstrates a strong ability to generalise to many datasets and domains in a zero-shot setting.

Kontekst
0
Wejście
3,33 USD
Wyjście
0,00 USD
Dźwięk
Zobacz model
OUR METHOD

Jak AIToolly przetwarza dane modeli

Fakty dotyczące modelu i endpointu dostawcy są rozdzielone. Katalogi zewnętrzne służą do odkrywania i migawek, a oficjalna dokumentacja i zweryfikowane karty modeli mają pierwszeństwo dla danych modelu.

01Źródła i weryfikacja
02Zweryfikowano
03Możliwość