Build an AI Scientist for Life Science Discovery with NVIDIA BioNeMo Agent Toolkit
https://developer.nvidia.com/blog/build-an-ai-scientist-for-life-science-discovery-with-nvidia-bionemo-agent-toolkit/
https://developer.nvidia.com/blog/build-an-ai-scientist-for-life-science-discovery-with-nvidia-bionemo-agent-toolkit/
NVIDIA Technical Blog
Build an AI Scientist for Life Science Discovery with NVIDIA BioNeMo Agent Toolkit
AI scientists are emerging as a new interface for scientific computing. These agents can read papers, write code, generate hypotheses, call APIs, inspect files, and iterate on results. But science isn’…
Boost Inference Performance up to 15x on NVIDIA Blackwell Using DFlash Speculative Decoding
https://developer.nvidia.com/blog/boost-inference-performance-up-to-15x-on-nvidia-blackwell-using-dflash-speculative-decoding/
https://developer.nvidia.com/blog/boost-inference-performance-up-to-15x-on-nvidia-blackwell-using-dflash-speculative-decoding/
NVIDIA Technical Blog
Boost Inference Performance up to 15x on NVIDIA Blackwell Using DFlash Speculative Decoding
As AI systems move from single-turn interactions to coordinated multiagent workflows, low-latency inference becomes increasingly important. Autoregressive LLMs generate tokens sequentially…
Maximize AI Factory Energy Efficiency Through Full-Stack Inference and Training Optimizations
https://developer.nvidia.com/blog/maximize-ai-factory-energy-efficiency-through-full-stack-inference-and-training-optimizations/
https://developer.nvidia.com/blog/maximize-ai-factory-energy-efficiency-through-full-stack-inference-and-training-optimizations/
NVIDIA Technical Blog
Maximize AI Factory Energy Efficiency Through Full-Stack Inference and Training Optimizations
Power can account for 40% of the operating expenses (OpEx) to run an AI factory. Each watt can be spent on overhead, data ingestion, training, or generating tokens for customers. And most sites are…
NVIDIA and AWS Collaborate to Bring AI to Production at Scale
https://blogs.nvidia.com/blog/nvidia-aws-ai-production-scale/
https://blogs.nvidia.com/blog/nvidia-aws-ai-production-scale/
NVIDIA Blog
NVIDIA and AWS Collaborate to Bring AI to Production at Scale
Across Amazon OpenSearch and Amazon EC2, NVIDIA AI infrastructure is giving enterprises more practical paths to deploy AI at production scale.
Accelerating BEV Pooling on NVIDIA GPUs for Physical AI Applications
https://developer.nvidia.com/blog/accelerating-bev-pooling-on-nvidia-gpus-for-physical-ai-applications/
https://developer.nvidia.com/blog/accelerating-bev-pooling-on-nvidia-gpus-for-physical-ai-applications/
NVIDIA Technical Blog
Accelerating BEV Pooling on NVIDIA GPUs for Physical AI Applications
An increasingly common design pattern for autonomous vehicles (AVs), robotics, and spatial AI systems is bird’s-eye-view (BEV) perception. BEV models project multicamera image features into a shared…
The Ultimate Summer Sale Pairing: Steam Sale Meets GeForce NOW Discounts
https://blogs.nvidia.com/blog/geforce-now-thursday-steam-summer-sale-2026/
https://blogs.nvidia.com/blog/geforce-now-thursday-steam-summer-sale-2026/
NVIDIA Blog
The Ultimate Summer Sale Pairing: Steam Sale Meets GeForce NOW Discounts
Big summer deals arrive along with ‘Dark Scrolls’ and ‘The Adventures of Elliot: The Millennium Tales’ joining six new games this week.
How KRAFTON Built PUBG Ally, a Co-Playable Character Powered by NVIDIA ACE
https://developer.nvidia.com/blog/how-krafton-built-pubg-ally-a-co-playable-character-powered-by-nvidia-ace/
https://developer.nvidia.com/blog/how-krafton-built-pubg-ally-a-co-playable-character-powered-by-nvidia-ace/
NVIDIA Technical Blog
Q&A: How KRAFTON Built PUBG Ally, a Co-Playable Character Powered by NVIDIA ACE
AI companions in games have long been constrained by fixed dialogue. PUBG Ally is a different kind of system. Built by KRAFTON for PUBG: BATTLEGROUNDS, this AI teammate is powered by NVIDIA ACE and…
Scaling AI Inference Across Multiple GPUs Using NVIDIA TensorRT with Multi-Device Inference Support
https://developer.nvidia.com/blog/scaling-ai-inference-across-multiple-gpus-using-nvidia-tensorrt-with-multi-device-inference-support/
https://developer.nvidia.com/blog/scaling-ai-inference-across-multiple-gpus-using-nvidia-tensorrt-with-multi-device-inference-support/
NVIDIA Technical Blog
Scaling AI Inference Across Multiple GPUs Using NVIDIA TensorRT with Multi-Device Inference Support
Generative AI workloads are rapidly outgrowing the memory and compute budget of single GPUs. For inference developers building media generation pipelines, the challenge is scaling across multiple…
Streamlining Resource Binding with End-to-End Support for Vulkan Descriptor Heaps
https://developer.nvidia.com/blog/streamlining-resource-binding-with-end-to-end-support-for-vulkan-descriptor-heaps/
https://developer.nvidia.com/blog/streamlining-resource-binding-with-end-to-end-support-for-vulkan-descriptor-heaps/
NVIDIA Technical Blog
Streamlining Resource Binding with End-to-End Support for Vulkan Descriptor Heaps
Shaders are GPU programs that process visual data—such as rays, pixels, geometry, and textures—to produce specific rendering effects. Shaders find necessary data through a process called resource…
Creating the NVIDIA Nemotron 3 Ultra NVFP4 Checkpoint with NVIDIA Model Optimizer
https://developer.nvidia.com/blog/creating-the-nvidia-nemotron-3-ultra-nvfp4-checkpoint-with-nvidia-model-optimizer/
https://developer.nvidia.com/blog/creating-the-nvidia-nemotron-3-ultra-nvfp4-checkpoint-with-nvidia-model-optimizer/
NVIDIA Technical Blog
Creating the NVIDIA Nemotron 3 Ultra NVFP4 Checkpoint with NVIDIA Model Optimizer
As context windows grow longer, moving large model weights efficiently becomes critical to performance. A common way to address this is quantization, an optimization technique that compresses model…
Deploy a Production-Ready NVIDIA AI-Q Blueprint on Oracle Cloud Infrastructure
https://developer.nvidia.com/blog/deploy-a-production-ready-nvidia-ai-q-blueprint-on-oracle-cloud-infrastructure/
https://developer.nvidia.com/blog/deploy-a-production-ready-nvidia-ai-q-blueprint-on-oracle-cloud-infrastructure/
NVIDIA Technical Blog
Deploy a Production-Ready NVIDIA AI-Q Blueprint on Oracle Cloud Infrastructure
AI agents have changed a lot in the last two years. The first could only answer one question at a time. Then came multi-turn chat, where the model could keep some context across a session. Today…
How to Govern Autonomous Agents in Enterprise AI Factories
https://developer.nvidia.com/blog/how-to-govern-autonomous-agents-in-enterprise-ai-factories/
https://developer.nvidia.com/blog/how-to-govern-autonomous-agents-in-enterprise-ai-factories/
NVIDIA Technical Blog
How to Govern Autonomous Agents in Enterprise AI Factories
AI agents are quickly moving beyond chat. They inspect code, run tests, read documents, search knowledge bases, query internal systems, and operate for hours on behalf of a user.
Open Models, Closed Environments: Palantir Brings Secure AI to US Agencies With NVIDIA Nemotron
https://blogs.nvidia.com/blog/palantir-secure-ai-us-agencies-nemotron-open-models/
https://blogs.nvidia.com/blog/palantir-secure-ai-us-agencies-nemotron-open-models/
NVIDIA Blog
Open Models, Closed Environments: Palantir Brings Secure AI to US Agencies With NVIDIA Nemotron
NVIDIA Nemotron open models help Palantir advance US technology leadership with mission-specific sovereign AI for government agencies and critical infrastructure operators.
Firefly Aerospace Operates NVIDIA Jetson in Lunar Orbit for the First Time
https://blogs.nvidia.com/blog/firefly-aerospace-nvidia-jetson-lunar-orbit/
https://blogs.nvidia.com/blog/firefly-aerospace-nvidia-jetson-lunar-orbit/
NVIDIA Blog
Firefly Aerospace Operates NVIDIA Jetson in Lunar Orbit for the First Time
The NVIDIA Inception member’s Ocula moon imaging service will harness the NVIDIA Jetson platform for edge AI, running inference directly in space to significantly accelerate insights compared with downlinking all data back down to Earth.
Claude Meets Blackwell Ultra: Anthropic’s Models Now Run on NVIDIA GB300 in Azure
https://blogs.nvidia.com/blog/anthropic-nvidia-gb300-blackwell-ultra-microsoft-azure/
https://blogs.nvidia.com/blog/anthropic-nvidia-gb300-blackwell-ultra-microsoft-azure/
NVIDIA Blog
Claude Meets Blackwell Ultra: Anthropic’s Models Now Run on NVIDIA GB300 in Azure
Now generally available in Microsoft Foundry, Claude on NVIDIA GB300 Blackwell Ultra gives Azure-native enterprises a new foundation for building autonomous and domain-specific AI agents.
Into the Omniverse: Three Workflows for Improving Vision AI Agent Accuracy With Synthetic Data and Fine-Tuning
https://blogs.nvidia.com/blog/vision-ai-agent-skills-omniverse-metropolis/
https://blogs.nvidia.com/blog/vision-ai-agent-skills-omniverse-metropolis/
NVIDIA Blog
Into the Omniverse: Three Workflows for Improving Vision AI Agent Accuracy With Synthetic Data and Fine-Tuning
Vision AI agents are becoming a practical way to automatically turn video data from the physical world into operational intelligence in factories, cities, warehouses and transportation systems.
👍5
How Jaiveer Singh Is Helping Robots — and Developers — Move Faster
https://blogs.nvidia.com/blog/nvidia-life-jaiveer-singh/
https://blogs.nvidia.com/blog/nvidia-life-jaiveer-singh/
NVIDIA Blog
How Jaiveer Singh Is Helping Robots — and Developers — Move Faster
When Jaiveer Singh talks about robots, he doesn’t begin with spectacle. He begins with infrastructure: the boards inside machines, the software that lets developers see through a robot’s cameras and the engineering required before a robot can leave a demo…
👍8
How NVIDIA’s Inference Software Stack Powers the Lowest Token Cost
https://blogs.nvidia.com/blog/inference-software-lowest-token-cost/
https://blogs.nvidia.com/blog/inference-software-lowest-token-cost/
NVIDIA Blog
How NVIDIA’s Inference Software Stack Powers the Lowest Token Cost
Baseten, Cognition, Deep Infra, Together AI and Cursor are seeing compounding value from NVIDIA’s software and open source ecosystem.
NVIDIA BioNeMo Agent Toolkit Brings Accelerated AI to Life Sciences Researchers in Claude Science
https://blogs.nvidia.com/blog/claude-science-bionemo-agent-toolkit/
https://blogs.nvidia.com/blog/claude-science-bionemo-agent-toolkit/
NVIDIA Blog
NVIDIA BioNeMo Agent Toolkit Brings Accelerated AI to Life Sciences Researchers in Claude Science
Anthropic’s Claude Science, an AI workbench for science research, lets scientists use agents to build workflows harnessing NVIDIA BioNeMo models, libraries and accelerated workflows.
Optimizing a Neural Reconstruction Pipeline Using NVIDIA Nsight Developer Tools
https://developer.nvidia.com/blog/optimizing-a-neural-reconstruction-pipeline-using-nvidia-nsight-developer-tools/
https://developer.nvidia.com/blog/optimizing-a-neural-reconstruction-pipeline-using-nvidia-nsight-developer-tools/
NVIDIA Technical Blog
Optimizing a Neural Reconstruction Pipeline Using NVIDIA Nsight Developer Tools
NVIDIA Omniverse NuRec is a neural reconstruction pipeline for building high-fidelity 3D representations of real-world environments from multisensor data such as cameras and lidar. It is used to…
Designing GPU-Accelerated Query Engines with NVIDIA GQE
https://developer.nvidia.com/blog/designing-gpu-accelerated-query-engines-with-nvidia-gqe/
https://developer.nvidia.com/blog/designing-gpu-accelerated-query-engines-with-nvidia-gqe/
NVIDIA Technical Blog
Designing GPU-Accelerated Query Engines with NVIDIA GQE
GPU-accelerated query engines are often constrained by memory and I/O bandwidth. NVIDIA hardware advances—including high bandwidth memory (HBM), NVIDIA NVLink-C2C, and dedicated decompression engines…