How to Run Isolated Tenant Kubernetes Clusters on Shared GPU Infrastructure
NVIDIA 的文章讨论了共享 GPU 平台中的租户架构,认为为每个团队单独建 Kubernetes 集群往往过于严格,提出通过隔离租户集群在共享资源下实现隔离与资源利用之间的平衡。
来源: NVIDIA Developer BlogNVIDIA 的 AI 新闻汇总:产品发布、模型更新与行业动态。最新:How to Run Isolated Tenant Kubernetes Clusters on Shared GPU Infrastructure
NVIDIA 的文章讨论了共享 GPU 平台中的租户架构,认为为每个团队单独建 Kubernetes 集群往往过于严格,提出通过隔离租户集群在共享资源下实现隔离与资源利用之间的平衡。
来源: NVIDIA Developer BlogNVIDIA开发者博客介绍了Vera Storage的基准测试,称其在用于AI原生场景、包括agentic AI流程的存储任务中,在加密、压缩、完整性校验和恢复速度方面更快。
来源: NVIDIA Developer BlogNVIDIA 开发者博客讨论了通过协同设计模型注意力机制来提升长上下文交互推理速度,文中指出在智能体和长上下文任务增多时,注意力计算在推理时间中的占比会上升。
来源: NVIDIA Developer BlogNVIDIA发布了视频编解码 SDK 13.1,强调零拷贝转码、AV1 B 帧和帧级精确寻址等特性,以应对各行业对高质量视频不断上升的需求。
来源: NVIDIA Developer BlogNVIDIA 开发者博客称 nvmath-python 是一个库,旨在连接 Python 科研用户与 NVIDIA CUDA-X 数学库,并用于扩展高性能核心数学计算。
来源: NVIDIA Developer BlogNVIDIA 的文章指出,知识工作者越来越多地将 AI 代理整合进工作流程,并将其作为“数字同事”使用,文中提出了提升这类代理安全部署的四种方式。
来源: NVIDIA Developer BlogNVIDIA 的博客指出,即使使用看似相同的 H100、GB200/GB300 NVL72 系统,训练吞吐量仍可能出现明显差异,并给出提升性能的做法。
来源: NVIDIA Developer BlogNVIDIA 表示,借助 GeForce NOW,普通笔记本也可在学习与游戏之间切换,实现 RTX 级的云端游戏体验。
来源: NVIDIA BlogNVIDIA 开发者博客介绍了使用 NeMo Guardrails 自建并验证 AI 编码助手的方法,面向受监管、主权或源代码敏感场景。
来源: NVIDIA Developer BlogUnlike autonomous driving or industrial robotics, healthcare robotics can’t rely on internet-scale data collection or unlimited real-world experimentation....
来源: NVIDIA Developer BlogAnyone can make a robot move; NVIDIA Jetson makes it think. As a discerning AI investor who values style and substance, Sarah Guo knows this season’s standout accessory isn’t the latest designer purse — but what’s inside it. In a recent video, Guo, founder of...
来源: NVIDIA BlogNVIDIA Ising Calibration is an open source vision language model (VLM) designed to interpret diagnostic outputs from quantum processors and determine how they...
来源: NVIDIA Developer BlogOpen source software is a critical pillar of the global economy. It underpins cloud computing, financial services, manufacturing, telecommunications, government and internet services by making technology accessible and observable to communities of experts. Cyb...
来源: NVIDIA BlogBuilding a great AI agent isn’t just about choosing the right models. The harness is the architecture surrounding the model. How it renders context, executes...
来源: NVIDIA Developer BlogThe complexity of modern chip design continues to grow as engineering teams work to develop increasingly sophisticated CPUs, GPUs and AI systems. To help meet that challenge, NVIDIA is collaborating with industry leaders Cadence and Synopsys to optimize critic...
来源: NVIDIA BlogModern chip design is increasingly limited by engineering time. Register transfer level (RTL) development and verification require specialized hardware...
来源: NVIDIA Developer BlogAs AI workloads increase, explosive compute demand is pushing the semiconductor industry to meet unprecedented performance targets. Even small delays can have...
来源: NVIDIA Developer BlogEvery byte moved has a cost. As model checkpoints grow to hundreds of gigabytes or even a terabyte, that cost adds up quickly. To make things even worse, moving...
来源: NVIDIA Developer BlogAt this week’s AI Summit in San Francisco, South Korean President Jae Myung Lee and some of the country’s top business leaders and researchers are meeting with NVIDIA and ecosystem partners to chart Korea’s AI progress. Building on NVIDIA founder and CEO Jense...
来源: NVIDIA BlogNVIDIA OptiX ray tracing engine is an application framework for achieving optimal ray tracing performance on the GPU. Applications using OptiX can fail in ways...
来源: NVIDIA Developer BlogCustomization is what enables developers to take a general model and tailor it to use cases, domains, languages, and more. However, customization comes with a...
来源: NVIDIA Developer BlogLock in and load up the cloud. GFN Thursday brings fresh updates and new adventures, all ready to play without waiting for downloads. Set sail in Path of Exile: Curse of the Allflame and charge in Battlefield 6 Season 4 both launching major content for members...
来源: NVIDIA BlogNVIDIA founder and CEO Jensen Huang today visited the Naval Postgraduate School in Monterey, California, to commission an NVIDIA DGX GB300 system — bringing one of the world’s most powerful AI platforms fully online for the students, researchers and faculty at...
来源: NVIDIA BlogA TensorRT engine build can take seconds to many minutes. Large strongly typed models, deep tactic search, and a cold timing cache on a brand-new GPU SKU can...
来源: NVIDIA Developer BlogBefore a healthcare robot can be useful in the real world, it has to learn how the physical world pushes back. Anatomy varies. Instruments bend, press, slip and interact with tissue. Imaging can be noisy or incomplete. And the rare, edge scenarios developers m...
来源: NVIDIA BlogThe AI era runs on AI infrastructure. Many of these advanced systems are built and tested in Texas. Wistron opened its first U.S. manufacturing facility today in Fort Worth — a 324,000-square-foot greenfield plant producing superchips at the heart of some of t...
来源: NVIDIA BlogNVIDIA Vera Rubin is here, and it’s going gigascale. Vera Rubin NVL72 production is ramping up with racks running at partners CoreWeave, Google Cloud, Microsoft Azure and Oracle Cloud Infrastructure. Spanning 350-plus factory sites in 30 countries, Vera Rubin...
来源: NVIDIA BlogAI has entered the gigascale era. The world’s most advanced AI factories are bringing together hundreds of thousands of GPUs and CPUs to train frontier models, power agentic AI and generate intelligence at unprecedented scale. At this level, networking becomes...
来源: NVIDIA BlogAgentic AI shifts more of the critical execution path onto the CPU. Agents operate in sandboxes to execute code, invoke tools, retrieve context, interact with...
来源: NVIDIA Developer BlogWhat began as discrete AI model training and human-facing chat interfaces has evolved into always-on AI factories dedicated to producing intelligence at scale....
来源: NVIDIA Developer BlogCloudflare 第一次抓取 RSS 后,实时新闻会出现在这里。