NVIDIA AI 모델

NVIDIA의 선별 AI 모델 13개를 컨텍스트, API 가격, 모달리티, 기능으로 비교합니다.

13개의 수집 모델

모델 디렉토리

현재 13개 모델

비교

Llama Nemotron Rerank VL 1B V2 is a 1.7B multimodal reranking model from NVIDIA. It evaluates the relevance of document images and text against user queries, designed for vision RAG…

컨텍스트
10K
입력
US$0.00
출력
US$0.00
텍스트이미지
모델 보기
사용 가능

NVIDIA Nemotron 3 Embed 1B is an open text embedding model from NVIDIA, optimized for high-throughput, low-latency retrieval. It is suited for enterprise search, RAG, code retrieval, and agentic retrieval…

컨텍스트
33K
입력
US$0.00
출력
US$0.00
텍스트
모델 보기
사용 가능

Nemotron-3-Nano-30B-A3B-BF16 is a large language model (LLM) trained from scratch by NVIDIA, and designed as a unified model for both reasoning and non-reasoning tasks. It responds to user queries and tasks by first generating a reasoning trace and then concluding with a final response. The model's reasoning capabilities can be configured through a flag in the chat template. If the user prefers the model to provide its final answer without intermediate reasoning traces, it can be configured to do so, albeit with a slight decrease in accuracy for harder prompts that require reasoning. Conversely, allowing the model to generate reasoning traces first generally results in higher-quality final solutions to queries and tasks.

컨텍스트
262K
입력
US$0.05
출력
US$0.20
텍스트
모델 보기
사용 가능

NVIDIA Nemotron 3 Nano Omni is a multimodal large language model that unifies video, audio, image, and text understanding to support enterprise-grade Q&A, summarization, transcription, and document intelligence workflows. It extends the Nemotron Nano family with integrated video+speech comprehension, Graphical User Interface (GUI), Optical Character Recognition (OCR), and speech transcription capabilities, enabling end-to-end processing of rich enterprise content such as meeting recordings, M&E assets, training videos, and complex business documents. NVIDIA Nemotron 3 Nano Omni was developed by NVIDIA as part of the Nemotron model family.

컨텍스트
256K
입력
US$0.00
출력
US$0.00
텍스트오디오이미지비디오
모델 보기
사용 가능

> Use temperature=1.0 and top_p=0.95 across **all tasks and serving backends** — reasoning, tool calling, and general chat alike.

컨텍스트
262K
입력
US$0.085
출력
US$0.40
텍스트
모델 보기
사용 가능

For more details on how to deploy and use the model - see the Quick Start Guide below!

컨텍스트
512K
입력
US$0.60
출력
US$3.6
텍스트
모델 보기

Nemotron 3.5 ASR Streaming Multilingual 0.6B is a speech recognition model from NVIDIA. Its prompt-conditioned, cache-aware FastConformer-RNNT design targets low-latency transcription across more than 40 languages for real-time captioning, voice…

컨텍스트
0
입력
US$3.33
출력
US$0.00
오디오
모델 보기
사용 가능

NVIDIA Nemotron™ is a family of open models with open weights, training data, and recipes, delivering leading efficiency and accuracy for building specialized AI agents.

컨텍스트
262K
입력
US$0.08
출력
US$0.20
텍스트
모델 보기
사용 가능

NVIDIA: Nemotron Nano 12B 2 VL (free) is included in the AIToolly model catalog.

컨텍스트
128K
입력
US$0.00
출력
US$0.00
이미지텍스트비디오
모델 보기
사용 가능

NVIDIA-Nemotron-Nano-9B-v2 is a large language model (LLM) trained from scratch by NVIDIA, and designed as a unified model for both reasoning and non-reasoning tasks. It responds to user queries and tasks by first generating a reasoning trace and then concluding with a final response. The model's reasoning capabilities can be controlled via a system prompt. If the user prefers the model to provide its final answer without intermediate reasoning traces, it can be configured to do so, albeit with a slight decrease in accuracy for harder prompts that require reasoning. Conversely, allowing the model to generate reasoning traces first generally results in higher-quality final solutions to queries and tasks.

컨텍스트
128K
입력
US$0.00
출력
US$0.00
텍스트
모델 보기
사용 가능

Parakeet TDT 0.6B v3 is NVIDIA's 600M-parameter multilingual speech-to-text model built on the FastConformer-TDT architecture. Trained on the Granary dataset (670,000+ hours of audio), it supports automatic language detection across…

컨텍스트
0
입력
US$1,500
출력
US$0.00
오디오
모델 보기
OUR METHOD

AIToolly의 모델 데이터 처리 방식

모델 고유 정보와 제공사 엔드포인트 정보를 분리합니다. 서드파티 카탈로그는 탐색과 스냅샷에 사용하며 모델 정보는 공식 문서와 검증된 모델 카드를 우선합니다.

01출처 및 검증
02검증일
03기능