

Data is fundamentally changing the way companies do business, driving demand for data scientists and increasing the complexity in their workflows. Get the performance you need to transform massive amounts of data into insights and create amazing customer experiences with NVIDIA-powered data science workstations. Built by leading workstation providers to combine the power of Quadro RTX GPUs with accelerated CUDA-X AI data science software to deliver a new breed of fully-integrated desktop and mobile workstations for data science.


Run inference on trained machine learning or deep learning models from any framework on any processor—GPU, CPU, or other—with NVIDIA Triton Inference Server™. Part of the NVIDIA AI platform and available with NVIDIA AI Enterprise, Triton Inference Server is open-source software that standardizes AI model deployment and execution across every workload.

Enter a new frontier in professional graphics with unprecedented performance and scalability with 48 GB of high-speed GDDR6 memory and NVIDIA NVLink™. Designers and artists across industries can now expand the boundary of what’s possible, working with the largest and most complex ray tracing, deep learning, and visual computing workloads.

NVIDIA Nemotron-Nano-9B-v2 is a compact, open-source language model designed to deliver high-performance reasoning and agentic capabilities. Utilizing a hybrid Mamba-Transformer architecture, it efficiently processes long-context sequences up to 128,000 tokens, making it suitable for complex tasks requiring extensive context understanding. The model supports multiple languages, including English, German, French, Italian, Spanish, and Japanese, and excels in instruction following and code generation tasks. Key Features and Functionality: - Hybrid Architecture: Combines Mamba-2 state-space layers with Transformer attention layers, enhancing throughput and accuracy in reasoning tasks. - Efficient Long-Context Processing: Capable of handling sequences up to 128,000 tokens on a single NVIDIA A10G GPU, facilitating scalable long-context reasoning. - Multilingual Support: Trained on data spanning 15 languages and 43 programming languages, enabling broad multilingual and coding fluency. - Toggleable Reasoning Feature: Allows users to control the model's reasoning process using simple commands like "/think" or "/no_think," balancing accuracy and response speed. - Reasoning Budget Control: Introduces a "thinking budget" mechanism, enabling developers to set the number of tokens used during the reasoning process, optimizing for latency or cost. Primary Value and User Solutions: NVIDIA Nemotron-Nano-9B-v2 addresses the need for efficient, high-performance language models capable of handling extensive context and complex reasoning tasks. Its hybrid architecture and advanced features provide developers and researchers with a versatile tool for building AI applications that require deep understanding and rapid processing of large-scale textual data. The model's open-source nature and permissive licensing facilitate widespread adoption and customization, empowering users to deploy sophisticated AI solutions across various domains.

NVIDIA Nemotron is a family of open-source, multimodal AI models designed to empower developers and enterprises in building advanced agentic AI systems. These models excel in tasks such as complex reasoning, coding, visual understanding, and information retrieval, making them versatile tools for a wide range of applications. Key Features and Functionality: - Open Models: NVIDIA provides transparent and adaptable models, allowing developers to customize and deploy AI solutions with confidence. - High Compute Efficiency: The Nemotron family is optimized for computational efficiency, utilizing NVIDIA TensorRT-LLM to deliver higher throughput and on-demand reasoning capabilities. - High Accuracy: Post-trained with high-quality datasets, Nemotron models achieve top accuracy on leading benchmarks, ensuring reliable performance across various tasks. - Secure and Simple Deployment: Available as optimized NVIDIA NIM microservices, these models offer peak inference performance with flexible deployment options, ensuring superior security, privacy, and portability. Primary Value and Solutions: NVIDIA Nemotron addresses the growing need for transparent, efficient, and high-performing AI models in the development of agentic AI systems. By offering open models with high accuracy and compute efficiency, Nemotron enables developers and enterprises to create trustworthy AI agents capable of complex reasoning and decision-making. This empowers organizations to innovate and deploy AI solutions across various industries, enhancing productivity and driving business transformation.

This container provides a demonstration of how to deploy pre-trained models from NGC into an intelligent video analytics (IVA) pipeline in DeepStream. These models can be fine-tuned on additional data using Transfer Learning Toolkit.
Since its founding in 1993, NVIDIA (NASDAQ: NVDA) has been a pioneer in accelerated computing. The company’s invention of the GPU in 1999 sparked the growth of the PC gaming market, redefined computer graphics, ignited the era of modern AI and is fueling the creation of the metaverse. NVIDIA is now a full-stack computing company with data-center-scale offerings that are reshaping industry.