Research / developmentNot privateHigh confidence

NVIDIA Build / NVIDIA-hosted NIM

No-cost research and development inference across 102 models; not unrestricted production use.

Visit provider
Free accessResearch / development
Payment cardNo
AccountRequired
Sources10 first-party links
API endpointhttps://integrate.api.nvidia.com/v1

Models mentioned

102
01-ai/yi-largeadept/fuyu-8bai21labs/jamba-1.5-large-instructaisingapore/sea-lion-7b-instructbaai/bge-m3bigcode/starcoder2-15bdatabricks/dbrx-instructdeepseek-ai/deepseek-coder-6.7b-instructdeepseek-ai/deepseek-v4-flash-0731google/codegemma-1.1-7bgoogle/codegemma-7bgoogle/deplotgoogle/diffusiongemma-26b-a4b-itgoogle/gemma-2bgoogle/gemma-3-12b-itgoogle/gemma-3-4b-itgoogle/gemma-4-31b-itgoogle/recurrentgemma-2bibm/granite-3.0-3b-a800m-instructibm/granite-3.0-8b-instructibm/granite-34b-code-instructibm/granite-8b-code-instructmeta/codellama-70bmeta/llama-3.1-70b-instructmeta/llama-3.1-8b-instructmeta/llama-3.2-11b-vision-instructmeta/llama-3.2-1b-instructmeta/llama-3.2-3b-instructmeta/llama-3.2-90b-vision-instructmeta/llama-3.3-70b-instructmeta/llama-guard-4-12bmeta/llama2-70bmeta/muse-glimmer-30bmicrosoft/kosmos-2microsoft/phi-3-vision-128k-instructmicrosoft/phi-3.5-moe-instructminimaxai/minimax-m3mistralai/codestral-22b-instruct-v0.1mistralai/mistral-7b-instruct-v0.3mistralai/mistral-largemistralai/mistral-large-2-instructmistralai/mistral-nemotronmistralai/mixtral-8x22b-v0.1moonshotai/kimi-k2.6moonshotai/kimi-k3nv-mistralai/mistral-nemo-12b-instructnvidia/ai-synthetic-video-detectornvidia/cosmos-reason2-8bnvidia/embed-qa-4nvidia/ising-calibration-1.5-31bnvidia/llama-3.1-nemoguard-8b-content-safetynvidia/llama-3.1-nemoguard-8b-topic-controlnvidia/llama-3.1-nemotron-51b-instructnvidia/llama-3.1-nemotron-70b-instructnvidia/llama-3.1-nemotron-nano-8b-v1nvidia/llama-3.1-nemotron-nano-vl-8b-v1nvidia/llama-3.1-nemotron-safety-guard-8b-v3nvidia/llama-3.1-nemotron-ultra-253b-v1nvidia/llama-3.2-nemoretriever-1b-vlm-embed-v1nvidia/llama-3.2-nv-embedqa-1b-v1nvidia/llama-3.3-nemotron-super-49b-v1nvidia/llama-3.3-nemotron-super-49b-v1.5nvidia/llama-nemotron-embed-1b-v2nvidia/llama-nemotron-embed-vl-1b-v2nvidia/llama3-chatqa-1.5-70bnvidia/mistral-nemo-minitron-8b-8k-instructnvidia/nemoretriever-parsenvidia/nemotron-3-embed-1bnvidia/nemotron-3-nano-30b-a3bnvidia/nemotron-3-nano-omni-30b-a3b-reasoningnvidia/nemotron-3-super-120b-a12bnvidia/nemotron-3-ultra-550b-a55bnvidia/nemotron-3.5-content-safetynvidia/nemotron-3.5-lightning-30b-a3bnvidia/nemotron-4-340b-instructnvidia/nemotron-4-340b-rewardnvidia/nemotron-mini-4b-instructnvidia/nemotron-nano-12b-v2-vlnvidia/nemotron-nano-3-30b-a3bnvidia/nemotron-parse

Showing 80 of 102. Use the provider discovery URL for the complete live catalog.

Equivalent paid value

Up to $4,263.78/mo
Daily$140.08
Weekly$980.58
Monthly$4,263.78
One-time

How this was valued: Best-case rate-limit envelope for a single flagship model, meta/llama-3.3-70b-instruct, using NVIDIA's own published per-model limits on the build.nvidia.com model page ("Up to 40 rpm" and "10,000 requests per day") together with NVIDIA's documented per-request ceilings (131,072-token context length on the model page; max_tokens parameter maximum of 4,096 in the NVIDIA API reference). NVIDIA sells no per-token rate for this hosted route, so the paid comparison is the current OpenRouter price for the exact model ($0.10 input / $0.32 output per 1M tokens, snapshot 2026-08-22). Assumes uninterrupted saturation with maximal-context requests. This is one model's envelope only; the other 100+ catalog models have their own displayed limits and are excluded rather than summed, because the account-versus-model scope of the limits is not documented.

Partial because NVIDIA's own limit text says rates "may vary by model and traffic from other users may cause throttling", the phrasing is "up to", the allowance is development/prototyping-only, and the paid reference is a router price rather than an NVIDIA rate for the same service.

Limits and terms

Public Numeric Limits
No
Official Rule
Varies by model and concurrency; inspect the signed-in account.
Current Marketing Phrase
Unlimited prototyping
Allowed Use
prototypingresearchdevelopmenttestinglearning

What happens to your prompts?

Not privateReviewed

The free developer terms permit storage and broad product, service, and underlying-model improvement uses.

Plan Scope
NVIDIA API Catalog and NVIDIA-hosted NIM trial access under the Technology Access Terms; a model or product-specific agreement can supersede those terms.
Prompt Retention
No fixed prompt-retention period is promised. The terms permit NVIDIA, affiliates, and service providers to host and store User Content to provide and support the service, for security, and to improve products or underlying technology.
Response Retention
No inference-output-specific retention commitment was found in the governing trial terms.
Ordinary Logging
NVIDIA may monitor, scan, or review communications and User Content transmitted through its servers for safety, security, moderation, or legal requests.
Model Training
The terms do not make a narrow no-training commitment. Their User Content license permits modification and improvement of NVIDIA products, services, and underlying technology, so users should treat training or equivalent improvement use as permitted unless a product-specific agreement says otherwise.
Product Improvement
Expressly permitted under the User Content license.
Human Or Operator Access
Monitoring, scanning, and review are permitted for security, moderation, and legal purposes; service providers may process content under the license.
Subprocessors And Routing
NVIDIA affiliates and service providers may process User Content, and separately identified product agreements can apply to individual models or services.
Deletion Controls
The terms do not guarantee permanent access or recovery if data is deleted or lost; public User Content may remain after reposting, and NVIDIA accepts no general duty to remove it.
Caveat
Unless a separate product agreement expressly permits it, User Content must not contain confidential information, personal data, protected health information, payment-card data, or sensitive human-subject research. This makes the free trial unsuitable for sensitive prompts.

Eligibility

Account Required
Yes
Developer Program Membership Required
Yes
Membership Cost USD
0
Payment Method Required
not documented

Primary sources

10