
Telecom Operators Shift to Open Models for AI
NVIDIA's report reveals a significant 89% shift towards open models among telecom operators, enabling trust, control, and customization of AI capabilities.
44 stories · since June 11, 2025

NVIDIA's report reveals a significant 89% shift towards open models among telecom operators, enabling trust, control, and customization of AI capabilities.

A new AI-powered breast cancer screening platform is commercially available in several US states, aiming to address the diagnosis gap.

Takeaways − Red Hat AI released an FP8 quantized build of NVIDIA Nemotron 3.5 Lightning 30B A3B. Cuts GPU memory and disk by roughly 50% versus the BF16 reference weights.

The era of generative AI has upended technology supply chains, and there is one ironclad rule in 2026: If it has memory or storage, it’s getting more expensive. Even devices with years-old tech inside are still apparently subject to that unwritten rule.

The US has arrested another suspect accused of smuggling high-end computer servers containing export-controlled Nvidia chips into China.

The era of generative AI has upended technology supply chains, and there is one ironclad rule in 2026: If it has memory or storage, it's getting more…

Local AI is becoming more useful by the token. As AI agents move from experiments into everyday development, increasingly capable open models are shrinking to…

GPT-6 Astra Ultrafast, running on NVIDIA Blackwell GPUs, is available now in the OpenAI API and to eligible ChatGPT Work and Codex users.

Build Applications on NVIDIA BlueField Faster with NVIDIA DOCA Agent Skills | NVIDIA Technical Blog Build Applications on NVIDIA BlueField Faster with NVIDIA DOCA Agent Skills By Claudia Martinez , Matan Raz and Tamir Tevet NVIDIA DOCA AI agent skills provide verified API signatures, hardware capability requirements, and build constraints so agents can reason like experienced DOCA developers.

Build Local AI Apps with C++ and NVIDIA TensorRT RTX Samples | NVIDIA Technical Blog Build Local AI Apps with C++ and NVIDIA TensorRT RTX Samples Do Inference Now (DIN) Deploy provides open-source C++ samples that combine ONNX Runtime with the NVIDIA TensorRT RTX execution provider to accelerate local AI inference on Windows and Linux.

In my previous post: Building persistent memory for multi-agent AI systems with Amazon S3 Vectors, we explored why memory engineering is the foundational…

Spooky season is streaming in. Alongside falling leaves, pumpkin spice and everything nice, 25 new games are joining GeForce NOW throughout October, including…

AI factories are built by the megawatt, even by the gigawatt. Each megawatt factory costs roughly $60 million, and AI factory operators will only commit…

Fine-Tuning NVIDIA Nemotron for Saudi Arabic Dialects, with a Path to Other Languages | NVIDIA Technical Blog Fine-Tuning NVIDIA Nemotron for Saudi Arabic Dialects, with a Path to Other Languages By Imane Khaouja , Amine El Khair , Meshari Alaeena , Zahra Al-Kaf and Abdulrahman Alkhamees NVIDIA Nemotron 3.5 ASR supports multilingual streaming transcription across 40 language-locales, but deployment-specific dialects...

Deploying an HSTU Generative Recommender with NVIDIA Dynamo-Triton | NVIDIA Technical Blog Deploying an HSTU Generative Recommender with NVIDIA Dynamo-Triton By Junyi Qiu , Shijie Liu , J Wyman and Mudit Aggarwal Generative recommender systems reformulate personalization as sequence modeling over user behavior, and NVIDIA recsys-examples now provides an end-to-end HSTU inference workflow with Dynamo-Triton.

On Tuesday, executives for Google , Anthropic , Meta , OpenAI , xAI , and Nvidia all signed “The White House Accord on Super Intelligence,” which was announced following a luncheon held by President Donald Trump.

Expanding AI Storage Access with NVIDIA cuObject and the NVIDIA SCADA Server SDK | NVIDIA Technical Blog Expanding AI Storage Access with NVIDIA cuObject and the NVIDIA SCADA Server SDK By Harish Arora , Vikram Sharma Mailthody , Kiran K.

Bringing together the world’s brightest minds and the latest accelerated computing technology leads to powerful breakthroughs that help tackle some of the…

Tracing Agent Harness Behavior with NVIDIA NeMo Relay | NVIDIA Technical Blog Tracing Agent Harness Behavior with NVIDIA NeMo Relay Learn how to use execution traces to understand agent behavior and determine whether harness changes improve task outcomes.

Building on nearly a decade of co-engineering, CoreWeave has built NVIDIA compute, networking and software into a cloud purpose-built for AI that’s still…

AI Native by Design: Lessons Learned from Building NVIDIA TensorRT Model Connect | NVIDIA Technical Blog AI Native by Design: Lessons Learned from Building NVIDIA TensorRT Model Connect Parallel work, model family isolation, reversible changes, and GPU-backed validation shaped an open source project designed around coding agents NVIDIA TensorRT Model Connect is an open source collection of AI model reference...

Lower the Cost of Building and Running Visual AI Agents with NVIDIA VSS Blueprint 3.3 | NVIDIA Technical Blog Lower the Cost of Building and Running Visual AI Agents with NVIDIA VSS Blueprint 3.3 By Hassan Moustafa , Debraj Sinha and Ashwani Agarwal The NVIDIA Metropolis Blueprint for Video Search and Summarization (VSS) 3.3 connects vision-language models such as NVIDIA...

NVIDIA Kumo Tabular Sets a New Accuracy-Efficiency Frontier for Tabular Prediction NVIDIA Kumo Tabular Sets a New Accuracy-Efficiency Frontier for Tabular Prediction Aleksandar S. Sokolovski aleksssokolovski Follow NVIDIA Kumo Tabular, part of the NVIDIA Kumo Structured model collection, is an open foundation model for tabular data now available on Hugging Face .

NVIDIA Open Agent Safety Platform: A Reference for Continuous In-Silicon Agent Monitoring | NVIDIA Technical Blog NVIDIA Open Agent Safety Platform: A Reference for Continuous In-Silicon Agent Monitoring Secure AI agents with a safety enforcement layer spanning software and hardware By John Myers , Alex Watson , Ali Golshan and Ofir Arkin NVIDIA OpenShell provides an open-source secure runtime that...

Add Runtime Controls to AI Agents with NVIDIA OpenShell | NVIDIA Technical Blog Add Runtime Controls to AI Agents with NVIDIA OpenShell NVIDIA OpenShell 0.1.0 provides an open-source runtime that enforces which systems and data an AI agent can access without rewriting the agent.

When COVID-19 emerged, scientists had a crucial advantage: Decades of prior research on coronaviruses meant they understood the virus’ key proteins well…

A warped Manhattan is waiting in the cloud this week. Remedy Entertainment’s CONTROL Resonant brings Dylan Faden’s extraordinary abilities and a paranatural…

When Sakeena Fiza describes her work as a validation engineer at NVIDIA, she does so in terms more befitting a detective story than a world-class engineering…

NVIDIA AI Day Singapore, which takes place Sept. 22-23 at the Raffles City Convention Centre, is offering attendees opportunities to explore the hands-on…

The kernels team at Together recently received access to the NVIDIA Vera Rubin NVL72 platform. We spent the past few days digging through the new ISA and poking the chip with micros.

The Information reports that Nvidia is discussing an equity investment in Perplexity at a valuation above $30 billion.

NVIDIA Nemotron 3.5 Lightning is now available on Ollama, and it runs completely on your own device. It's a 30 billion parameter (3B active) open model from NVIDIA built for agents that stay running: gathering context, calling tools, and working through multi-step tasks.

Figure 1: CUDA-to-MLX optimization translation map. CUDA optimization knowledge can be translated into architecture-native MLX strategies rather than copied…

NVIDIA Nemotron 3 Ultra is offering leading performance at lower cost than top closed models with the largest and most widely adopted AI agent orchestration…

NVIDIA Nemotron 3 Ultra is now available on Ollama’s cloud. It’s a 550 billion parameter (55B active) open model from NVIDIA built for long-running, agentic workflows with fast and affordable performance across hundreds of tool calls.

How to Build, Run, and Scale High-Quality Creator Workflows in ComfyUI | NVIDIA Technical Blog How to Build, Run, and Scale High-Quality Creator Workflows in ComfyUI By Joel Pennington , Margaret Zhang and Maitri Taneja NVIDIA GenAI Creator Toolkit provides three production-ready ComfyUI workflows that run locally on NVIDIA RTX GPUs without cloud dependencies.

Runway , an AI research and technology startup, said Tuesday that it has raised $315 million in a Series E round of funding. General Atlantic led the financing, which included participation from Nvidia , Adobe Ventures , AMD Ventures , Fidelity Management & Research Co.

Editor’s note: The name of NVIDIA DRIVE Hyperion was changed to NVIDIA Hyperion in September 2026. All references to the name have been updated in this blog.

Unveiling what it describes as the most capable model series yet for professional knowledge work, OpenAI launched GPT-5.2 in December.

Celtic languages — including Cornish, Irish, Scottish Gaelic and Welsh — are the U.K.’s oldest living languages.

For more than a century, meteorologists have chased storms with chalkboards, equations, and now, supercomputers.

Bringing together the world’s brightest minds and the latest accelerated computing technology leads to powerful breakthroughs that help tackle some of the…

Ceramics — the humble mix of earth, fire and artistry — have been part of a global conversation for millennia.

At GTC Paris — held alongside VivaTech, Europe’s largest tech event — NVIDIA founder and CEO Jensen Huang delivered a clear message: Europe isn’t just…