AI news about NVIDIA: product launches, model updates, and industry moves. Latest: How to Run Isolated Tenant Kubernetes Clusters on Shared GPU Infrastructure
NVIDIA’s article discusses a tenancy strategy for shared GPU platforms, arguing that per-team Kubernetes clusters can be overkill and that isolated tenant clusters can provide useful separation while keeping infrastructure shared.
NVIDIA’s developer post describes Vera Storage benchmarks that claim faster encryption, compression, integrity checks, and recovery features for storage workflows used in AI-native environments, including agentic AI operations.
NVIDIA’s developer blog discusses co-designing AI model attention mechanisms to improve speed for long-context, interactive workloads, noting that as such workloads grow, attention can become a larger portion of inference time.
NVIDIA released Video Codec SDK 13.1, highlighting features such as zero-copy transcode, AV1 B-frames, and frame-accurate seek to support stronger video performance amid rising demand across industries.
The NVIDIA Developer Blog item describes nvmath-python as a library intended to connect Python scientific users with NVIDIA CUDA-X math libraries for higher-performance core math at scale.
NVIDIA’s developer blog item says knowledge workers are increasingly using AI agents as digital coworkers and highlights four approaches aimed at deploying those agents more securely in workflows.
The NVIDIA blog discusses how AI infrastructure built with similar H100, GB200 NVL72, or GB300 NVL72 systems can still show notable differences in training throughput, and shares methods for extracting full performance.
NVIDIA says GeForce NOW can let users switch between study and gaming on a standard laptop, using cloud gaming to access GeForce RTX-class performance.
Unlike autonomous driving or industrial robotics, healthcare robotics can’t rely on internet-scale data collection or unlimited real-world experimentation....
Anyone can make a robot move; NVIDIA Jetson makes it think. As a discerning AI investor who values style and substance, Sarah Guo knows this season’s standout accessory isn’t the latest designer purse — but what’s inside it. In a recent video, Guo, founder of...
NVIDIA Ising Calibration is an open source vision language model (VLM) designed to interpret diagnostic outputs from quantum processors and determine how they...
Open source software is a critical pillar of the global economy. It underpins cloud computing, financial services, manufacturing, telecommunications, government and internet services by making technology accessible and observable to communities of experts. Cyb...
Building a great AI agent isn’t just about choosing the right models. The harness is the architecture surrounding the model. How it renders context, executes...
The complexity of modern chip design continues to grow as engineering teams work to develop increasingly sophisticated CPUs, GPUs and AI systems. To help meet that challenge, NVIDIA is collaborating with industry leaders Cadence and Synopsys to optimize critic...
Modern chip design is increasingly limited by engineering time. Register transfer level (RTL) development and verification require specialized hardware...
As AI workloads increase, explosive compute demand is pushing the semiconductor industry to meet unprecedented performance targets. Even small delays can have...
Every byte moved has a cost. As model checkpoints grow to hundreds of gigabytes or even a terabyte, that cost adds up quickly. To make things even worse, moving...
At this week’s AI Summit in San Francisco, South Korean President Jae Myung Lee and some of the country’s top business leaders and researchers are meeting with NVIDIA and ecosystem partners to chart Korea’s AI progress. Building on NVIDIA founder and CEO Jense...
NVIDIA OptiX ray tracing engine is an application framework for achieving optimal ray tracing performance on the GPU. Applications using OptiX can fail in ways...
Customization is what enables developers to take a general model and tailor it to use cases, domains, languages, and more. However, customization comes with a...
Lock in and load up the cloud. GFN Thursday brings fresh updates and new adventures, all ready to play without waiting for downloads. Set sail in Path of Exile: Curse of the Allflame and charge in Battlefield 6 Season 4 both launching major content for members...
NVIDIA founder and CEO Jensen Huang today visited the Naval Postgraduate School in Monterey, California, to commission an NVIDIA DGX GB300 system — bringing one of the world’s most powerful AI platforms fully online for the students, researchers and faculty at...
A TensorRT engine build can take seconds to many minutes. Large strongly typed models, deep tactic search, and a cold timing cache on a brand-new GPU SKU can...
Before a healthcare robot can be useful in the real world, it has to learn how the physical world pushes back. Anatomy varies. Instruments bend, press, slip and interact with tissue. Imaging can be noisy or incomplete. And the rare, edge scenarios developers m...
The AI era runs on AI infrastructure. Many of these advanced systems are built and tested in Texas. Wistron opened its first U.S. manufacturing facility today in Fort Worth — a 324,000-square-foot greenfield plant producing superchips at the heart of some of t...
NVIDIA Vera Rubin is here, and it’s going gigascale. Vera Rubin NVL72 production is ramping up with racks running at partners CoreWeave, Google Cloud, Microsoft Azure and Oracle Cloud Infrastructure. Spanning 350-plus factory sites in 30 countries, Vera Rubin...
AI has entered the gigascale era. The world’s most advanced AI factories are bringing together hundreds of thousands of GPUs and CPUs to train frontier models, power agentic AI and generate intelligence at unprecedented scale. At this level, networking becomes...
Agentic AI shifts more of the critical execution path onto the CPU. Agents operate in sandboxes to execute code, invoke tools, retrieve context, interact with...
What began as discrete AI model training and human-facing chat interfaces has evolved into always-on AI factories dedicated to producing intelligence at scale....
Source: NVIDIA Developer Blog
Live RSS items will appear after the first Cloudflare ingest run.