AI Lab

NVIDIA News

Official NVIDIA announcements — model releases, product launches and research, each one summarized with every source covering it, by AIBriefs. RSS

AnalysisCybersecurity1 source

Four Ways to Deploy More Secure AI Agents

NVIDIA blog post outlines four key security practices for deploying AI agents in enterprise workflows, including AI red teaming and continuous monitoring.

AnalysisAI Agents1 source

NVIDIA details six agent harness capabilities

The blog post explains how agent harness architecture—context rendering, execution planning, tool integration, and more—affects model performance. It covers six key capabilities to build better AI agents.

EventBusiness12 sources

Nvidia invests $1B in Naver, expands SK Group partnership for Korea AI buildout

Nvidia will invest $1 billion in Naver to finance an AI data center in South Korea, and expanded its partnership with SK Group to build over 2 gigawatts of AI data centers, with the two companies expecting to do more than $500 billion in business. South Korea plans to inject 20 trillion won ($13.9 billion) into its sovereign wealth fund for AI investments.

How-ToDevelopers3 sources

Customize NVIDIA Nemotron 3 Nano with Prime Intellect Lab

NVIDIA and Prime Intellect Lab release a guide for customizing Nemotron 3 Nano using reinforcement learning with verifiable rewards (RLVR) and LoRA adapters. The tutorial covers setup in a math-python environment and training steps to tailor the model for specific use cases.

How-ToDevelopers1 source

NVIDIA TensorRT adds observable and cancelable engine builds

TensorRT engine builds can now be made observable and cancelable in Python or C++, addressing long-running builds that can take minutes. The feature supports progress callbacks and cancellation for large strongly typed models.

AnalysisRobotics1 source

NVIDIA overviews state of simulation for physical AI

The blog post reviews current simulation platforms and techniques for training and testing physical AI systems, including robotics and autonomous vehicles. It covers key simulators, challenges in sim-to-real transfer, and the role of digital twins. The overview is published on Hugging Face as part of a collaboration between NVIDIA and the AI community.

EventHealth1 source

Bristol Myers Squibb to build AI factory on NVIDIA Vera Rubin

Bristol Myers Squibb (BMS) is deploying its second NVIDIA-powered AI cluster, called the 'SuperDuperPOD', on the new NVIDIA Vera Rubin platform. BMS already runs one of the largest AI clusters in life sciences, achieving significant results in drug discovery.

How-ToDevelopers1 source

Integrating Context-Aware Video AI Agents Into Enterprise Workflows

NVIDIA NemoClaw, a collection of open blueprints for autonomous agents, enables context-aware video AI agents to integrate with enterprise systems like content management, messaging, and databases. The approach moves from analysis to action by orchestrating blueprints for structured reports and multistep workflows.

EventBusiness7 sources

Japan, NVIDIA launch first national AI infrastructure

NVIDIA and Noetra Corp. will build an AI factory with 13,750 Vera CPUs and 27,500 Rubin GPUs, delivering 140 MW capacity. Supported by Japan's METI, it will create open multimodal foundation models for physical AI in manufacturing, logistics, and healthcare.

AnalysisAI Models2 sources

NVIDIA's Nemotron Challenge: 5,000 Kagglers Improve AI Reasoning

Over 5,000 participants across 4,000 teams competed in the NVIDIA Nemotron Model Reasoning Challenge on Kaggle. Winning approaches treated reasoning as a full engineering workflow, using LoRA adapters (rank ≤32) and synthetic chain-of-thought data to improve accuracy on the Nemotron-3-Nano-30B model.

How-ToAI Models4 sources

Post-Train NVIDIA Cosmos 3 in One Day Using Agent Skills

NVIDIA claims using TAO 7 agent skills, vision models can exceed 90% accuracy with minimal manual effort. Cosmos 3 is a new open world reasoning VLM, and TAO 7 provides a suite of tools for fine-tuning with coding agents and natural language prompts.

LaunchAI Models1 source

NVIDIA releases Nemotron-3-Embed-8B embedding model

NVIDIA released Nemotron-3-Embed-8B-BF16, an 8B-parameter embedding model in BF16 precision, available on HuggingFace. It is part of the Nemotron-3 series designed for text embedding tasks.

LaunchAI Models1 source

NVIDIA releases Nemotron-3-Embed-1B model

NVIDIA released Nemotron-3-Embed-1B-BF16 on HuggingFace, a 1B parameter embedding model. The model has garnered 50 likes and over 31,000 downloads since release.

AnalysisAI Models1 source

Guided generative models estimate extreme event likelihoods

NVIDIA presents guided generative models to efficiently estimate probabilities of rare, high-impact events across science, engineering, and finance. The approach improves sampling efficiency for extreme events critical for risk assessment.

How-ToRobotics3 sources

NVIDIA shares guide for evaluating general-purpose robot policies

NVIDIA published a blog post detailing RoboLab, its simulation benchmarking platform for evaluating robot foundation models. The guide covers real-world deployment challenges and best practices for testing general-purpose robot policies.

AnalysisAI Models1 source

NVIDIA explores hardware-friendly LLM co-design

Blog post discusses balancing accuracy, throughput, and latency in LLM design. Key dimensions: accuracy, throughput, and deployment latency must be optimized together.

AnalysisAI Models1 source

Synthetic Data Generation for Financial AI with NVIDIA NeMo

NVIDIA NeMo provides a pipeline for generating synthetic financial news data to fine-tune LLMs, addressing data scarcity and imbalance. The method uses domain-specific prompts to produce diverse, balanced datasets for sentiment analysis.

LaunchAI Models4 sources

NVIDIA releases Audex, unified audio-text LLM

Audex is a 30B-A3B MoE model built on Nemotron-Cascade-2, handling both audio understanding and generation. It retains the text intelligence of its backbone; a smaller 2B variant is also available under a noncommercial license.

EventAI Models3 sources

ICML 2026 highlights trends in open models and AI infrastructure

The 2026 International Conference on Machine Learning (ICML) in Seoul showcased a growing research focus on open frontier models and open AI infrastructure. Industry participants, including NVIDIA and Together AI, presented research on topics ranging from latent planning to inference optimization.

AnalysisBusiness1 source

Nations Deploy AI for Strategic Priorities

Nations are investing in domestic AI infrastructure to advance economies, protect data, and seize opportunities in transportation, healthcare, and other sectors. The blog from NVIDIA's Calista Redmond highlights AI as a key technology for national priorities.

AnalysisCybersecurity1 source

NVIDIA details hardware-rooted AI security for Blackwell

NVIDIA's blog post describes using Blackwell hardware features to secure AI inference without performance degradation. The solution integrates with TensorRT-LLM and Dynamo for runtime verification and attestation.

LaunchBusiness1 source

NVIDIA unlocks AI compute at scale with new capital partner model

NVIDIA introduces a revenue-sharing model enabling AI clouds to procure GPUs with credit support. Sharon AI is among the first partners, deploying up to 40,000 GB300 GPUs. NVIDIA earns standard product revenue plus a share of cloud revenue on supported capacity.