NVIDIA Vera Rubin NVL72 Delivers Leading Performance in MLPerf Inference v6.1 Debut
https://blogs.nvidia.com/blog/vera-rubin-nvl72-mlperf-inference/
https://blogs.nvidia.com/blog/vera-rubin-nvl72-mlperf-inference/
NVIDIA Blog
NVIDIA Vera Rubin NVL72 Delivers Leading Performance in MLPerf Inference v6.1 Debut
First preview submission using NVIDIA Vera Rubin NVL72 systems delivers up to 3.7x better throughput than NVIDIA GB300 NVL72.
TensorRT Edge-LLM Completes the MLPerf Edge Agentic Benchmark 6.4x Faster on Jetson AGX Thor
https://developer.nvidia.com/blog/tensorrt-edge-llm-completes-the-mlperf-edge-agentic-benchmark-6-4x-faster-on-jetson-agx-thor/
https://developer.nvidia.com/blog/tensorrt-edge-llm-completes-the-mlperf-edge-agentic-benchmark-6-4x-faster-on-jetson-agx-thor/
NVIDIA Technical Blog
TensorRT Edge-LLM Completes the MLPerf Edge Agentic Benchmark 6.4x Faster on Jetson AGX Thor
AI agents are moving from cloud data centers to vehicles, robots, and other edge devices. Unlike a chatbot that answers a single prompt, an agent works through a sequence of steps. It selects tools…
How to Use AI Agents to Prepare 3D Scenes for Simulation
https://developer.nvidia.com/blog/how-to-use-ai-agents-to-prepare-3d-scenes-for-simulation/
https://developer.nvidia.com/blog/how-to-use-ai-agents-to-prepare-3d-scenes-for-simulation/
NVIDIA Technical Blog
How to Use AI Agents to Prepare 3D Scenes for Simulation
Agentic AI workflows can be used to prepare and validate digital twins for physical AI systems. Agents can inspect 3D scenes, author simulation-relevant data in OpenUSD, add physics properties…
Cute Critters Come to the Cloud: ‘Aniimo’ Launches on GeForce NOW
https://blogs.nvidia.com/blog/geforce-now-thursday-aniimo/
https://blogs.nvidia.com/blog/geforce-now-thursday-aniimo/
NVIDIA Blog
Cute Critters Come to the Cloud: ‘Aniimo’ Launches on GeForce NOW
Explore the creature-catching adventure at launch, alongside ‘Active Matter,’ ‘Outer Wilds’ and '007 First Light' with a path-tracing update.
👍13
Benchmarking LLM Inference at Scale with AIPerf
https://developer.nvidia.com/blog/benchmarking-llm-inference-at-scale-with-aiperf/
https://developer.nvidia.com/blog/benchmarking-llm-inference-at-scale-with-aiperf/
NVIDIA Technical Blog
Benchmarking LLM Inference at Scale with AIPerf
You’re deploying a model on a system. It starts up, prompts are getting responses. Now the hard question: Is this fast? Your instincts might lead you to send curl commands, hand-roll an asyncio script…
👍8
Turn Your Latest Observations Into Timely Weather Decisions With NVIDIA Earth-2
https://developer.nvidia.com/blog/turn-your-latest-observations-into-timely-weather-decisions-with-nvidia-earth-2/
https://developer.nvidia.com/blog/turn-your-latest-observations-into-timely-weather-decisions-with-nvidia-earth-2/
NVIDIA Technical Blog
Turn Your Latest Observations Into Timely Weather Decisions With NVIDIA Earth-2
Weather-sensitive industries increasingly have access to observations that offer an earlier, more local view of changing conditions. Energy companies collect measurements across wind and solar assets…
AI Security Is an Engineering Problem — How to Solve It at Every Layer of the Agent Stack
https://blogs.nvidia.com/blog/ai-security-agent-stack/
https://blogs.nvidia.com/blog/ai-security-agent-stack/
NVIDIA Blog
AI Security Is an Engineering Problem — How to Solve It at Every Layer of the Agent Stack
Open research, controls across the agent stack and continuous testing help defenders build and operate more secure AI systems.
From Enablement to Execution, Egypt’s AI Ecosystem Reaches Production Scale
https://blogs.nvidia.com/blog/egypt-africa-ai-ecosystem/
https://blogs.nvidia.com/blog/egypt-africa-ai-ecosystem/
NVIDIA Blog
From Enablement to Execution, Egypt’s AI Ecosystem Reaches Production Scale
AI factories are also coming online across South Africa, Morocco and Nigeria, enabling Africa’s developers and enterprises to reach production-scale compute onshore.
Why Deploying Physical AI at Scale Demands Safety at Every Layer
https://blogs.nvidia.com/blog/physical-ai-halos-safety/
https://blogs.nvidia.com/blog/physical-ai-halos-safety/
NVIDIA Blog
Why Deploying Physical AI at Scale Demands Safety at Every Layer
Physical AI safety spans hardware, software and validation. See how NVIDIA Halos brings full-stack safety to autonomous vehicles and robots.
NVIDIA Launches DSX Ready to Qualify Power and Cooling Products for AI Factories
https://blogs.nvidia.com/blog/dsx-ready-ai-factories-power-cooling/
https://blogs.nvidia.com/blog/dsx-ready-ai-factories-power-cooling/
NVIDIA Blog
NVIDIA Launches DSX Ready to Qualify Power and Cooling Products for AI Factories
The new qualification program helps AI factory builders identify products that meet applicable NVIDIA DSX requirements, starting with battery energy storage systems and cooling distribution units.
How to Evaluate AI Agents From Tool Calls to Task Completion
https://developer.nvidia.com/blog/how-to-evaluate-ai-agents-from-tool-calls-to-task-completion/
https://developer.nvidia.com/blog/how-to-evaluate-ai-agents-from-tool-calls-to-task-completion/
NVIDIA Technical Blog
How to Evaluate AI Agents From Tool Calls to Task Completion
When you ship an AI agent, the key question is whether it can execute a chain of work across dozens of sequential tool calls against a live environment, and recover when a step fails.
Simplifying Model Serving Across Multiple GPUs with NVIDIA TensorRT Multi-Device Integration in NVIDIA Dynamo-Triton
https://developer.nvidia.com/blog/simplifying-model-serving-across-multiple-gpus-with-nvidia-tensorrt-multi-device-integration-in-nvidia-dynamo-triton/
https://developer.nvidia.com/blog/simplifying-model-serving-across-multiple-gpus-with-nvidia-tensorrt-multi-device-integration-in-nvidia-dynamo-triton/
NVIDIA Technical Blog
Simplifying Model Serving Across Multiple GPUs with NVIDIA TensorRT Multi-Device Integration in NVIDIA Dynamo-Triton
The compute and memory demands of generative AI increasingly exceed what a single GPU can provide. NVIDIA TensorRT multi-device inference is a new capability that enables a single TensorRT network to…
👍8
Accelerating a ROS 2 Node with an AI Agent and NVIDIA Isaac ROS
https://developer.nvidia.com/blog/accelerating-a-ros-2-node-with-an-ai-agent-and-nvidia-isaac-ros/
https://developer.nvidia.com/blog/accelerating-a-ros-2-node-with-an-ai-agent-and-nvidia-isaac-ros/
NVIDIA Technical Blog
Accelerating a ROS 2 Node with an AI Agent and NVIDIA Isaac ROS
GPU acceleration can speed up compute-intensive robotics workloads, but a fast CUDA kernel alone does not guarantee a fast ROS 2 graph. As messages move between nodes, they may continue to be…
What’s New for Game Developers: DLSS 5 with 3D-Guided Neural Rendering, NVIDIA ACE Updates, and New RTX Kit Capabilities
https://developer.nvidia.com/blog/whats-new-for-game-developers-dlss-5-with-3d-guided-neural-rendering-nvidia-ace-updates-and-new-rtx-kit-capabilities/
https://developer.nvidia.com/blog/whats-new-for-game-developers-dlss-5-with-3d-guided-neural-rendering-nvidia-ace-updates-and-new-rtx-kit-capabilities/
NVIDIA Technical Blog
What’s New for Game Developers: DLSS 5 with 3D-Guided Neural Rendering, NVIDIA ACE Updates, and New RTX Kit Capabilities
NVIDIA DLSS 5 introduces DLSS 3D-Guided Neural Rendering and granular controls that help game developers add lifelike lighting and material detail while preserving their artistic intent.
NVIDIA Isaac ROS 5.0 Advances Agentic, Open Source Robotics Development
https://blogs.nvidia.com/blog/isaac-ros-5-0-agentic-open-source-robotics/
https://blogs.nvidia.com/blog/isaac-ros-5-0-agentic-open-source-robotics/
NVIDIA Blog
NVIDIA Isaac ROS 5.0 Advances Agentic, Open Source Robotics Development
The latest release brings AI agent capabilities to the ROS developer ecosystem while expanding open source physical AI libraries and deployment across the NVIDIA Jetson platform.
Topology-Aware Workload Scheduling with NVIDIA Topograph
https://developer.nvidia.com/blog/topology-aware-workload-scheduling-with-nvidia-topograph/
https://developer.nvidia.com/blog/topology-aware-workload-scheduling-with-nvidia-topograph/
NVIDIA Technical Blog
Topology-Aware Workload Scheduling with NVIDIA Topograph
AI factories are power-limited systems that deliver maximum value when fully optimized. GPU workload placement is a key optimization. Poor workload placement fragments topology domains and forces…
Enabling Private High-Performance Production AI Inference with NVIDIA Confidential Computing
https://developer.nvidia.com/blog/enabling-private-high-performance-production-ai-inference-with-nvidia-confidential-computing/
https://developer.nvidia.com/blog/enabling-private-high-performance-production-ai-inference-with-nvidia-confidential-computing/
NVIDIA Technical Blog
Enabling Private High-Performance Production AI Inference with NVIDIA Confidential Computing
As large language model (LLM) inference increasingly processes sensitive information and proprietary model context across personal, enterprise, and regulated settings, data must be processed inside a…
At AI Day Singapore, NVIDIA and Partners Showcase AI Advancements Across Southeast Asia
https://blogs.nvidia.com/blog/ai-day-singapore/
https://blogs.nvidia.com/blog/ai-day-singapore/
NVIDIA Blog
At AI Day Singapore, NVIDIA and Partners Showcase AI Advancements Across Southeast Asia
NVIDIA is helping nations move AI from experimentation to production-scale deployment through open models, developer tools and a broad partner ecosystem.
Sakeena Fiza Helps NVIDIA Hardware Succeed at Scale
https://blogs.nvidia.com/blog/nvidia-life-sakeena-fiza/
https://blogs.nvidia.com/blog/nvidia-life-sakeena-fiza/
NVIDIA Blog
Sakeena Fiza Helps NVIDIA Hardware Succeed at Scale
By testing systems from lab bring-up to production, validation engineers turn cutting-edge hardware into dependable infrastructure.
👍4
How SWE-Serve Exposes the Gap Between Local Tests and Live Serving
https://developer.nvidia.com/blog/how-swe-serve-exposes-the-gap-between-local-tests-and-live-serving/
https://developer.nvidia.com/blog/how-swe-serve-exposes-the-gap-between-local-tests-and-live-serving/
NVIDIA Technical Blog
How SWE-Serve Exposes the Gap Between Local Tests and Live Serving
An AI coding agent’s patch can pass tests yet fail when the server loads a real model and handles requests. Evaluating changes to inference-serving software therefore requires checking the full…
👍4