NVIDIA Build / NVIDIA-hosted NIM
No-cost research and development inference across 102 models; not unrestricted production use.
https://integrate.api.nvidia.com/v1Models mentioned
01-ai/yi-largeadept/fuyu-8bai21labs/jamba-1.5-large-instructaisingapore/sea-lion-7b-instructbaai/bge-m3bigcode/starcoder2-15bdatabricks/dbrx-instructdeepseek-ai/deepseek-coder-6.7b-instructdeepseek-ai/deepseek-v4-flash-0731google/codegemma-1.1-7bgoogle/codegemma-7bgoogle/deplotgoogle/diffusiongemma-26b-a4b-itgoogle/gemma-2bgoogle/gemma-3-12b-itgoogle/gemma-3-4b-itgoogle/gemma-4-31b-itgoogle/recurrentgemma-2bibm/granite-3.0-3b-a800m-instructibm/granite-3.0-8b-instructibm/granite-34b-code-instructibm/granite-8b-code-instructmeta/codellama-70bmeta/llama-3.1-70b-instructmeta/llama-3.1-8b-instructmeta/llama-3.2-11b-vision-instructmeta/llama-3.2-1b-instructmeta/llama-3.2-3b-instructmeta/llama-3.2-90b-vision-instructmeta/llama-3.3-70b-instructmeta/llama-guard-4-12bmeta/llama2-70bmeta/muse-glimmer-30bmicrosoft/kosmos-2microsoft/phi-3-vision-128k-instructmicrosoft/phi-3.5-moe-instructminimaxai/minimax-m3mistralai/codestral-22b-instruct-v0.1mistralai/mistral-7b-instruct-v0.3mistralai/mistral-largemistralai/mistral-large-2-instructmistralai/mistral-nemotronmistralai/mixtral-8x22b-v0.1moonshotai/kimi-k2.6moonshotai/kimi-k3nv-mistralai/mistral-nemo-12b-instructnvidia/ai-synthetic-video-detectornvidia/cosmos-reason2-8bnvidia/embed-qa-4nvidia/ising-calibration-1.5-31bnvidia/llama-3.1-nemoguard-8b-content-safetynvidia/llama-3.1-nemoguard-8b-topic-controlnvidia/llama-3.1-nemotron-51b-instructnvidia/llama-3.1-nemotron-70b-instructnvidia/llama-3.1-nemotron-nano-8b-v1nvidia/llama-3.1-nemotron-nano-vl-8b-v1nvidia/llama-3.1-nemotron-safety-guard-8b-v3nvidia/llama-3.1-nemotron-ultra-253b-v1nvidia/llama-3.2-nemoretriever-1b-vlm-embed-v1nvidia/llama-3.2-nv-embedqa-1b-v1nvidia/llama-3.3-nemotron-super-49b-v1nvidia/llama-3.3-nemotron-super-49b-v1.5nvidia/llama-nemotron-embed-1b-v2nvidia/llama-nemotron-embed-vl-1b-v2nvidia/llama3-chatqa-1.5-70bnvidia/mistral-nemo-minitron-8b-8k-instructnvidia/nemoretriever-parsenvidia/nemotron-3-embed-1bnvidia/nemotron-3-nano-30b-a3bnvidia/nemotron-3-nano-omni-30b-a3b-reasoningnvidia/nemotron-3-super-120b-a12bnvidia/nemotron-3-ultra-550b-a55bnvidia/nemotron-3.5-content-safetynvidia/nemotron-3.5-lightning-30b-a3bnvidia/nemotron-4-340b-instructnvidia/nemotron-4-340b-rewardnvidia/nemotron-mini-4b-instructnvidia/nemotron-nano-12b-v2-vlnvidia/nemotron-nano-3-30b-a3bnvidia/nemotron-parseShowing 80 of 102. Use the provider discovery URL for the complete live catalog.
Equivalent paid value
How this was valued: Best-case rate-limit envelope for a single flagship model, meta/llama-3.3-70b-instruct, using NVIDIA's own published per-model limits on the build.nvidia.com model page ("Up to 40 rpm" and "10,000 requests per day") together with NVIDIA's documented per-request ceilings (131,072-token context length on the model page; max_tokens parameter maximum of 4,096 in the NVIDIA API reference). NVIDIA sells no per-token rate for this hosted route, so the paid comparison is the current OpenRouter price for the exact model ($0.10 input / $0.32 output per 1M tokens, snapshot 2026-08-22). Assumes uninterrupted saturation with maximal-context requests. This is one model's envelope only; the other 100+ catalog models have their own displayed limits and are excluded rather than summed, because the account-versus-model scope of the limits is not documented.
Partial because NVIDIA's own limit text says rates "may vary by model and traffic from other users may cause throttling", the phrasing is "up to", the allowance is development/prototyping-only, and the paid reference is a router price rather than an NVIDIA rate for the same service.
Limits and terms
- Public Numeric Limits
- No
- Official Rule
- Varies by model and concurrency; inspect the signed-in account.
- Current Marketing Phrase
- Unlimited prototyping
- Allowed Use
- prototypingresearchdevelopmenttestinglearning
What happens to your prompts?
The free developer terms permit storage and broad product, service, and underlying-model improvement uses.
- Plan Scope
- NVIDIA API Catalog and NVIDIA-hosted NIM trial access under the Technology Access Terms; a model or product-specific agreement can supersede those terms.
- Prompt Retention
- No fixed prompt-retention period is promised. The terms permit NVIDIA, affiliates, and service providers to host and store User Content to provide and support the service, for security, and to improve products or underlying technology.
- Response Retention
- No inference-output-specific retention commitment was found in the governing trial terms.
- Ordinary Logging
- NVIDIA may monitor, scan, or review communications and User Content transmitted through its servers for safety, security, moderation, or legal requests.
- Model Training
- The terms do not make a narrow no-training commitment. Their User Content license permits modification and improvement of NVIDIA products, services, and underlying technology, so users should treat training or equivalent improvement use as permitted unless a product-specific agreement says otherwise.
- Product Improvement
- Expressly permitted under the User Content license.
- Human Or Operator Access
- Monitoring, scanning, and review are permitted for security, moderation, and legal purposes; service providers may process content under the license.
- Subprocessors And Routing
- NVIDIA affiliates and service providers may process User Content, and separately identified product agreements can apply to individual models or services.
- Deletion Controls
- The terms do not guarantee permanent access or recovery if data is deleted or lost; public User Content may remain after reposting, and NVIDIA accepts no general duty to remove it.
- Caveat
- Unless a separate product agreement expressly permits it, User Content must not contain confidential information, personal data, protected health information, payment-card data, or sensitive human-subject research. This makes the free trial unsuitable for sensitive prompts.
Governing documents
- Technology Access Terms
- https://assets.ngc.nvidia.com/products/api-catalog/legal/NVIDIA_Technology_Access_TOU.pdf
Eligibility
- Account Required
- Yes
- Developer Program Membership Required
- Yes
- Membership Cost USD
- 0
- Payment Method Required
- not documented