NVIDIA Nemotron 3.5 Lightning Delivers Fast, Accurate Specialized Task Execution for Long-Running Agents | NVIDIA Technical Blog
Nvidia Developer Blogby Tanya Lenz·11 Aug 2026
Long-running AI agents spend most of their time on high-volume execution: tool calls, result validation, and subagent delegation.
Route AI Agents Across Models with NVIDIA NeMo Switchyard | NVIDIA Technical Blog
Nvidia Developer Blogby Michelle Horton·11 Aug 2026
Building an AI agent does not end with choosing a single model. Each model has its own strengths, weaknesses, and cost profile, which can shift from one workload to another—or even within the same...
Run Local Agentic AI Workflows with Meta’s Muse Glimmer on NVIDIA | NVIDIA Technical Blog
Nvidia Developer Blogby Michelle Horton·10 Aug 2026
Meta returns to the open source ecosystem with the release of Muse Glimmer, a 30B open-weight dense model with a 120K+ context window built for local AI agentic work.
Beyond VLAs: How World Action Models Reshape Robot Manipulation | NVIDIA Technical Blog
Nvidia Developer Blogby Michelle Horton·4 Aug 2026
A central challenge in robotics is building policies that generalize beyond the demonstrations they’re trained on.
Generate Trajectories, Reasoning Traces, and Auto-Labels with NVIDIA Alpamayo 2 Super | NVIDIA Technical Blog
Nvidia Developer Blogby Elizabeth Goodman·4 Aug 2026
Autonomous vehicle (AV) development often relies on separate models for trajectory generation, high-level intent prediction, scene understanding, and data labeling.
NVIDIA Vera Storage Benchmarks: Faster Encryption, Compression, Integrity Checking, and Recovery for AI-Native Storage | NVIDIA Technical Blog
Nvidia Developer Blogby Elizabeth Goodman·3 Aug 2026
Storage is an active part of every agentic AI workflow. As agents retrieve enterprise knowledge, access persistent memory, reuse key-value (KV) cache data, execute tools, and generate new results...
How to Run Isolated Tenant Kubernetes Clusters on Shared GPU Infrastructure | NVIDIA Technical Blog
Nvidia Developer Blogby Tanya Lenz·3 Aug 2026
Running a dedicated Kubernetes cluster per team often results in more isolation than an organization requires.
Co-Designing AI Model Attention for Fast, Interactive Long-Context Inference | NVIDIA Technical Blog
Nvidia Developer Blogby Tanya Lenz·31 Jul 2026
As agentic and long-context workloads become common, the context lengths increase and attention consumes a larger share of inference time (Figure 1).
NVIDIA Video Codec SDK 13.1: Zero-Copy Transcode, AV1 B-Frames, and Frame-Accurate Seek | NVIDIA Technical Blog
Nvidia Developer Blogby Elizabeth Goodman·31 Jul 2026
The demand for high-quality video continues to accelerate across industries, powering everything from immersive streaming experiences to remote collaboration, generative AI media tools, and...
Run High-Performance Core Math at Scale with NVIDIA nvmath-python | NVIDIA Technical Blog
Nvidia Developer Blogby Michelle Horton·30 Jul 2026
NVIDIA nvmath-python is a library designed to bridge the gap between the Python scientific community and NVIDIA CUDA-X math libraries.
Four Ways to Deploy More Secure AI Agents | NVIDIA Technical Blog
Nvidia Developer Blogby Michelle Horton·30 Jul 2026
Knowledge workers are increasingly integrating AI agents into their workflows. Agents that function as “digital coworkers” offer clear benefits.
NVIDIA Exemplar Cloud: Lessons for Unlocking Full Performance on AI Infrastructure | NVIDIA Technical Blog
Nvidia Developer Blogby Elizabeth Goodman·30 Jul 2026
Two AI computing clusters built from identical NVIDIA H100, GB200 NVL72, or GB300 NVL72 systems can deliver materially different training throughput.
How to Self-Host a Validated AI Coding Assistant with NVIDIA NeMo Guardrails | NVIDIA Technical Blog
Nvidia Developer Blogby Tanya Lenz·29 Jul 2026
Deploying an AI coding assistant in a regulated, sovereign, or source-sensitive environment, often comes with challenges.
Developing Healthcare Robotics with GPU-Native Medical Physics Simulation | NVIDIA Technical Blog
Nvidia Developer Blogby Michelle Horton·28 Jul 2026
Unlike autonomous driving or industrial robotics, healthcare robotics can’t rely on internet-scale data collection or unlimited real-world experimentation.
NVIDIA Ising Enables Fully Automated Quantum Computer Calibration with Enhanced In-Context Learning | NVIDIA Technical Blog
Nvidia Developer Blogby Tanya Lenz·27 Jul 2026
NVIDIA Ising Calibration is an open source vision language model (VLM) designed to interpret diagnostic outputs from quantum processors and determine how they should be tuned to continue operating.
Six Agent Harness Capabilities for Higher Model Performance | NVIDIA Technical Blog
Nvidia Developer Blogby Michelle Horton·27 Jul 2026
Building a great AI agent isn’t just about choosing the right models. The harness is the architecture surrounding the model.
Advancing Semiconductor Innovation Across Materials Engineering and Manufacturing | NVIDIA Technical Blog
Nvidia Developer Blogby Tanya Lenz·27 Jul 2026
As AI workloads increase, explosive compute demand is pushing the semiconductor industry to meet unprecedented performance targets.
NVIDIA Nemotron 3 Ultra Leads Open Models on Accuracy and Efficiency in Agentic RTL Coding | NVIDIA Technical Blog
Nvidia Developer Blogby Nirmal Kumar Juluru·27 Jul 2026
Modern chip design is increasingly limited by engineering time. Register transfer level (RTL) development and verification require specialized hardware knowledge, precise reasoning, and repeated...
ModelExpress: Distributing Model Artifacts at the Speed of Light | NVIDIA Technical Blog
Nvidia Developer Blogby Elizabeth Goodman·24 Jul 2026
Every byte moved has a cost. As model checkpoints grow to hundreds of gigabytes or even a terabyte, that cost adds up quickly.
Debugging Ray Tracing Applications Using NVIDIA OptiX Toolkit | NVIDIA Technical Blog
Nvidia Developer Blogby Tanya Lenz·23 Jul 2026
NVIDIA OptiX ray tracing engine is an application framework for achieving optimal ray tracing performance on the GPU.
Your filters hide everything on this page. Adjust them in preferences.