From Megawatts to Tokens: How NVIDIA Maximizes AI Factory Production
https://blogs.nvidia.com/blog/from-megawatts-to-tokens-how-nvidia-maximizes-ai-factory-production/
https://blogs.nvidia.com/blog/from-megawatts-to-tokens-how-nvidia-maximizes-ai-factory-production/
NVIDIA Blog
From Megawatts to Tokens: How NVIDIA Maximizes AI Factory Production
On a sweltering August evening in Silicon Valley, as the sun dropped and air conditioning loads spiked, Silicon Valley Power sent a signal to an AI factory to adjust its power consumption. Varun Sivaram was watching on Zoom with about forty others — his team…
Scaling Federated Learning Across Docker, Kubernetes, and Slurm with NVIDIA FLARE
https://developer.nvidia.com/blog/scaling-federated-learning-across-docker-kubernetes-and-slurm-with-nvidia-flare/
https://developer.nvidia.com/blog/scaling-federated-learning-across-docker-kubernetes-and-slurm-with-nvidia-flare/
NVIDIA Technical Blog
Scaling Federated Learning Across Docker, Kubernetes, and Slurm with NVIDIA FLARE
Federated learning (FL) projects often begin with a straightforward setup: one server, a few clients, and one dataset at each site. As those projects grow, the challenge shifts from running an…
How NVIDIA Groq 3 LPX Deterministic Execution Drives Power-Efficient High-Interactivity Inference on NVIDIA Vera Rubin
https://developer.nvidia.com/blog/how-nvidia-groq-3-lpx-deterministic-execution-drives-power-efficient-high-interactivity-inference-on-nvidia-vera-rubin/
https://developer.nvidia.com/blog/how-nvidia-groq-3-lpx-deterministic-execution-drives-power-efficient-high-interactivity-inference-on-nvidia-vera-rubin/
NVIDIA Technical Blog
How NVIDIA Groq 3 LPX Deterministic Execution Drives Power-Efficient High-Interactivity Inference on NVIDIA Vera Rubin
Power is a defining constraint for AI factories. As AI workloads demand a full compute platform to serve them, each component of that platform must maximize output within the factory’s limited power…
How NVIDIA NVLink 6 Delivers Multi-Layer Resiliency for AI Factories
https://developer.nvidia.com/blog/how-nvidia-nvlink-6-delivers-multi-layer-resiliency-for-ai-factories/
https://developer.nvidia.com/blog/how-nvidia-nvlink-6-delivers-multi-layer-resiliency-for-ai-factories/
NVIDIA Technical Blog
How NVIDIA NVLink 6 Delivers Multi-Layer Resiliency for AI Factories
For operators of large-scale AI factories, maximizing continuous output is essential for productivity. In massive-scale AI training, every GPU in the cluster must synchronize gradients across…
Dense vs. MoE Models: Active Parameters, Throughput, and When to Choose Each
https://developer.nvidia.com/blog/dense-vs-moe-models-active-parameters-throughput-and-when-to-choose-each/
https://developer.nvidia.com/blog/dense-vs-moe-models-active-parameters-throughput-and-when-to-choose-each/
NVIDIA Technical Blog
Dense vs. MoE Models: Active Parameters, Throughput, and When to Choose Each
How can a 30B-parameter model activate only 3B parameters per token, and still use the capacity of the larger model? Nemotron 3.5 Lightning illustrates the answer: It uses a Mixture-of-Experts (MoE)…
‘Now We Can Know Everything and Do Anything,’ Jensen Huang Says at Dreamforce
https://blogs.nvidia.com/blog/jensen-huang-dreamforce/
https://blogs.nvidia.com/blog/jensen-huang-dreamforce/
NVIDIA Blog
‘Now We Can Know Everything and Do Anything,’ Jensen Huang Says at Dreamforce
At Dreamforce, Salesforce and NVIDIA announced Koa, a CRM reasoning model for Agentforce, post-trained from Nemotron on 27 years of Salesforce CRM intelligence — and running entirely within Salesforce’s own infrastructure.
University of Manchester Uses NVIDIA Earth-2 to Forecast Air Pollution Across the UK
https://blogs.nvidia.com/blog/uk-air-pollution-research-earth-2/
https://blogs.nvidia.com/blog/uk-air-pollution-research-earth-2/
NVIDIA Blog
University of Manchester Uses NVIDIA Earth-2 to Forecast Air Pollution Across the UK
Manchester physicists David Topping and Hao Zhang applied NVIDIA Earth-2 generative models to air quality forecasting — unlocking predictions fast enough to inform public health.
Emerald AI, Google and NVIDIA Launch Alliance to Advance Flexible AI Data Centers
https://blogs.nvidia.com/blog/ai-energy-management-alliance/
https://blogs.nvidia.com/blog/ai-energy-management-alliance/
NVIDIA Blog
Emerald AI, Google and NVIDIA Launch Alliance to Advance Flexible AI Data Centers
The AI Energy Management Alliance brings together the full AI and power value chain to accelerate the interconnection of flexible, grid-enhancing data centers, strengthen reliability and protect affordability.
Translating CUDA Tile Operations from Python to Rust Using Agentic AI
https://developer.nvidia.com/blog/translating-cuda-tile-operations-from-python-to-rust-using-agentic-ai/
https://developer.nvidia.com/blog/translating-cuda-tile-operations-from-python-to-rust-using-agentic-ai/
NVIDIA Technical Blog
Translating CUDA Tile Operations from Python to Rust Using Agentic AI
cuTile Rust () is a tile-based system for safe, idiomatic GPU kernel authoring in the Rust programming language. Extending the Rust ownership model to tile-based GPU kernels, it splits mutable outputs…
NVIDIA Vera Rubin NVL72 Delivers Leading Performance in MLPerf Inference v6.1 Debut
https://blogs.nvidia.com/blog/vera-rubin-nvl72-mlperf-inference/
https://blogs.nvidia.com/blog/vera-rubin-nvl72-mlperf-inference/
NVIDIA Blog
NVIDIA Vera Rubin NVL72 Delivers Leading Performance in MLPerf Inference v6.1 Debut
First preview submission using NVIDIA Vera Rubin NVL72 systems delivers up to 3.7x better throughput than NVIDIA GB300 NVL72.
TensorRT Edge-LLM Completes the MLPerf Edge Agentic Benchmark 6.4x Faster on Jetson AGX Thor
https://developer.nvidia.com/blog/tensorrt-edge-llm-completes-the-mlperf-edge-agentic-benchmark-6-4x-faster-on-jetson-agx-thor/
https://developer.nvidia.com/blog/tensorrt-edge-llm-completes-the-mlperf-edge-agentic-benchmark-6-4x-faster-on-jetson-agx-thor/
NVIDIA Technical Blog
TensorRT Edge-LLM Completes the MLPerf Edge Agentic Benchmark 6.4x Faster on Jetson AGX Thor
AI agents are moving from cloud data centers to vehicles, robots, and other edge devices. Unlike a chatbot that answers a single prompt, an agent works through a sequence of steps. It selects tools…
How to Use AI Agents to Prepare 3D Scenes for Simulation
https://developer.nvidia.com/blog/how-to-use-ai-agents-to-prepare-3d-scenes-for-simulation/
https://developer.nvidia.com/blog/how-to-use-ai-agents-to-prepare-3d-scenes-for-simulation/
NVIDIA Technical Blog
How to Use AI Agents to Prepare 3D Scenes for Simulation
Agentic AI workflows can be used to prepare and validate digital twins for physical AI systems. Agents can inspect 3D scenes, author simulation-relevant data in OpenUSD, add physics properties…
Cute Critters Come to the Cloud: ‘Aniimo’ Launches on GeForce NOW
https://blogs.nvidia.com/blog/geforce-now-thursday-aniimo/
https://blogs.nvidia.com/blog/geforce-now-thursday-aniimo/
NVIDIA Blog
Cute Critters Come to the Cloud: ‘Aniimo’ Launches on GeForce NOW
Explore the creature-catching adventure at launch, alongside ‘Active Matter,’ ‘Outer Wilds’ and '007 First Light' with a path-tracing update.
👍13
Benchmarking LLM Inference at Scale with AIPerf
https://developer.nvidia.com/blog/benchmarking-llm-inference-at-scale-with-aiperf/
https://developer.nvidia.com/blog/benchmarking-llm-inference-at-scale-with-aiperf/
NVIDIA Technical Blog
Benchmarking LLM Inference at Scale with AIPerf
You’re deploying a model on a system. It starts up, prompts are getting responses. Now the hard question: Is this fast? Your instincts might lead you to send curl commands, hand-roll an asyncio script…
👍8
Turn Your Latest Observations Into Timely Weather Decisions With NVIDIA Earth-2
https://developer.nvidia.com/blog/turn-your-latest-observations-into-timely-weather-decisions-with-nvidia-earth-2/
https://developer.nvidia.com/blog/turn-your-latest-observations-into-timely-weather-decisions-with-nvidia-earth-2/
NVIDIA Technical Blog
Turn Your Latest Observations Into Timely Weather Decisions With NVIDIA Earth-2
Weather-sensitive industries increasingly have access to observations that offer an earlier, more local view of changing conditions. Energy companies collect measurements across wind and solar assets…
AI Security Is an Engineering Problem — How to Solve It at Every Layer of the Agent Stack
https://blogs.nvidia.com/blog/ai-security-agent-stack/
https://blogs.nvidia.com/blog/ai-security-agent-stack/
NVIDIA Blog
AI Security Is an Engineering Problem — How to Solve It at Every Layer of the Agent Stack
Open research, controls across the agent stack and continuous testing help defenders build and operate more secure AI systems.
From Enablement to Execution, Egypt’s AI Ecosystem Reaches Production Scale
https://blogs.nvidia.com/blog/egypt-africa-ai-ecosystem/
https://blogs.nvidia.com/blog/egypt-africa-ai-ecosystem/
NVIDIA Blog
From Enablement to Execution, Egypt’s AI Ecosystem Reaches Production Scale
AI factories are also coming online across South Africa, Morocco and Nigeria, enabling Africa’s developers and enterprises to reach production-scale compute onshore.
Why Deploying Physical AI at Scale Demands Safety at Every Layer
https://blogs.nvidia.com/blog/physical-ai-halos-safety/
https://blogs.nvidia.com/blog/physical-ai-halos-safety/
NVIDIA Blog
Why Deploying Physical AI at Scale Demands Safety at Every Layer
Physical AI safety spans hardware, software and validation. See how NVIDIA Halos brings full-stack safety to autonomous vehicles and robots.
NVIDIA Launches DSX Ready to Qualify Power and Cooling Products for AI Factories
https://blogs.nvidia.com/blog/dsx-ready-ai-factories-power-cooling/
https://blogs.nvidia.com/blog/dsx-ready-ai-factories-power-cooling/
NVIDIA Blog
NVIDIA Launches DSX Ready to Qualify Power and Cooling Products for AI Factories
The new qualification program helps AI factory builders identify products that meet applicable NVIDIA DSX requirements, starting with battery energy storage systems and cooling distribution units.
How to Evaluate AI Agents From Tool Calls to Task Completion
https://developer.nvidia.com/blog/how-to-evaluate-ai-agents-from-tool-calls-to-task-completion/
https://developer.nvidia.com/blog/how-to-evaluate-ai-agents-from-tool-calls-to-task-completion/
NVIDIA Technical Blog
How to Evaluate AI Agents From Tool Calls to Task Completion
When you ship an AI agent, the key question is whether it can execute a chain of work across dozens of sequential tool calls against a live environment, and recover when a step fails.