Skip to main content
Entity

NVIDIA

44 stories · since June 11, 2025

Illustration for: Build Applications on NVIDIA BlueField Faster with NVIDIA DO
Products & Tools

Build Applications on NVIDIA BlueField Faster with NVIDIA DOCA Agent Skills | NVIDIA Technical Blog Build Applications on NVIDIA BlueField Faster with NVIDIA DOCA Agent Skills By Claudia Martinez , Matan Raz and Tamir Tevet NVIDIA DOCA AI agent skills provide verified API signatures, hardware capability requirements, and build constraints so agents can reason like experienced DOCA developers.

NVIDIA Developer Blog9 min
Illustration for: Build Local AI Apps with C++ and NVIDIA TensorRT RTX Samples
Open Source

Build Local AI Apps with C++ and NVIDIA TensorRT RTX Samples | NVIDIA Technical Blog Build Local AI Apps with C++ and NVIDIA TensorRT RTX Samples Do Inference Now (DIN) Deploy provides open-source C++ samples that combine ONNX Runtime with the NVIDIA TensorRT RTX execution provider to accelerate local AI inference on Windows and Linux.

NVIDIA Developer Blog4 min
Illustration for: Fine-Tuning NVIDIA Nemotron for Saudi Arabic Dialects, with
Models & Research

Fine-Tuning NVIDIA Nemotron for Saudi Arabic Dialects, with a Path to Other Languages | NVIDIA Technical Blog Fine-Tuning NVIDIA Nemotron for Saudi Arabic Dialects, with a Path to Other Languages By Imane Khaouja , Amine El Khair , Meshari Alaeena , Zahra Al-Kaf and Abdulrahman Alkhamees NVIDIA Nemotron 3.5 ASR supports multilingual streaming transcription across 40 language-locales, but deployment-specific dialects...

NVIDIA Developer Blog14 min
Illustration for: Deploying an HSTU Generative Recommender with NVIDIA Dynamo-
News

Deploying an HSTU Generative Recommender with NVIDIA Dynamo-Triton | NVIDIA Technical Blog Deploying an HSTU Generative Recommender with NVIDIA Dynamo-Triton By Junyi Qiu , Shijie Liu , J Wyman and Mudit Aggarwal Generative recommender systems reformulate personalization as sequence modeling over user behavior, and NVIDIA recsys-examples now provides an end-to-end HSTU inference workflow with Dynamo-Triton.

NVIDIA Developer Blog10 min
Illustration for: Tracing Agent Harness Behavior with NVIDIA NeMo Relay
News

Tracing Agent Harness Behavior with NVIDIA NeMo Relay | NVIDIA Technical Blog Tracing Agent Harness Behavior with NVIDIA NeMo Relay Learn how to use execution traces to understand agent behavior and determine whether harness changes improve task outcomes.

NVIDIA Developer Blog13 min
Illustration for: AI Native by Design: Lessons Learned from Building NVIDIA Te
Open Source

AI Native by Design: Lessons Learned from Building NVIDIA TensorRT Model Connect | NVIDIA Technical Blog AI Native by Design: Lessons Learned from Building NVIDIA TensorRT Model Connect Parallel work, model family isolation, reversible changes, and GPU-backed validation shaped an open source project designed around coding agents NVIDIA TensorRT Model Connect is an open source collection of AI model reference...

NVIDIA Developer Blog10 min
Illustration for: Lower the Cost of Building and Running Visual AI Agents with
News

Lower the Cost of Building and Running Visual AI Agents with NVIDIA VSS Blueprint 3.3 | NVIDIA Technical Blog Lower the Cost of Building and Running Visual AI Agents with NVIDIA VSS Blueprint 3.3 By Hassan Moustafa , Debraj Sinha and Ashwani Agarwal The NVIDIA Metropolis Blueprint for Video Search and Summarization (VSS) 3.3 connects vision-language models such as NVIDIA...

NVIDIA Developer Blog11 min
Illustration for: NVIDIA Kumo Tabular Sets a New Accuracy-Efficiency Frontier
Open Source

NVIDIA Kumo Tabular Sets a New Accuracy-Efficiency Frontier for Tabular Prediction NVIDIA Kumo Tabular Sets a New Accuracy-Efficiency Frontier for Tabular Prediction Aleksandar S. Sokolovski aleksssokolovski Follow NVIDIA Kumo Tabular, part of the NVIDIA Kumo Structured model collection, is an open foundation model for tabular data now available on Hugging Face .

Hugging Face Blog8 min
Illustration for: NVIDIA Open Agent Safety Platform: A Reference for Continuou
Products & Tools

NVIDIA Open Agent Safety Platform: A Reference for Continuous In-Silicon Agent Monitoring | NVIDIA Technical Blog NVIDIA Open Agent Safety Platform: A Reference for Continuous In-Silicon Agent Monitoring Secure AI agents with a safety enforcement layer spanning software and hardware By John Myers , Alex Watson , Ali Golshan and Ofir Arkin NVIDIA OpenShell provides an open-source secure runtime that...

NVIDIA Developer Blog9 min
Illustration for: Add Runtime Controls to AI Agents with NVIDIA OpenShell
Models & Research

Add Runtime Controls to AI Agents with NVIDIA OpenShell | NVIDIA Technical Blog Add Runtime Controls to AI Agents with NVIDIA OpenShell NVIDIA OpenShell 0.1.0 provides an open-source runtime that enforces which systems and data an AI agent can access without rewriting the agent.

NVIDIA Developer Blog9 min
Illustration for: NVIDIA Nemotron 3.5 Lightning
Open Source

NVIDIA Nemotron 3.5 Lightning is now available on Ollama, and it runs completely on your own device. It's a 30 billion parameter (3B active) open model from NVIDIA built for agents that stay running: gathering context, calling tools, and working through multi-step tasks.

Ollama Blog2 min
Illustration for: NVIDIA Nemotron 3 Ultra
Open Source

NVIDIA Nemotron 3 Ultra is now available on Ollama’s cloud. It’s a 550 billion parameter (55B active) open model from NVIDIA built for long-running, agentic workflows with fast and affordable performance across hundreds of tool calls.

Ollama Blog1 min
Illustration for: How to Build, Run, and Scale High-Quality Creator Workflows
Guides

How to Build, Run, and Scale High-Quality Creator Workflows in ComfyUI | NVIDIA Technical Blog How to Build, Run, and Scale High-Quality Creator Workflows in ComfyUI By Joel Pennington , Margaret Zhang and Maitri Taneja NVIDIA GenAI Creator Toolkit provides three production-ready ComfyUI workflows that run locally on NVIDIA RTX GPUs without cloud dependencies.

ComfyUI Blog10 min