How to Self-Host a Validated AI Coding Assistant with NVIDIA NeMo Guardrails
https://developer.nvidia.com/blog/how-to-self-host-a-validated-ai-coding-assistant-with-nvidia-nemo-guardrails/
https://developer.nvidia.com/blog/how-to-self-host-a-validated-ai-coding-assistant-with-nvidia-nemo-guardrails/
NVIDIA Technical Blog
How to Self-Host a Validated AI Coding Assistant with NVIDIA NeMo Guardrails
Deploying an AI coding assistant in a regulated, sovereign, or source-sensitive environment, often comes with challenges. Three common issues are: the source cannot leave the network…
Best in Class: Stream PC Games and Study on the Same Laptop With GeForce NOW
https://blogs.nvidia.com/blog/geforce-now-thursday-back-to-school-2026/
https://blogs.nvidia.com/blog/geforce-now-thursday-back-to-school-2026/
NVIDIA Blog
Best in Class: Stream PC Games and Study on the Same Laptop With GeForce NOW
Get ready for back to school with GeForce NOW on the same laptops used for studying. Start this week with eight new games.
NVIDIA Exemplar Cloud: Lessons for Unlocking Full Performance on AI Infrastructure
https://developer.nvidia.com/blog/nvidia-exemplar-cloud-lessons-for-unlocking-full-performance-on-ai-infrastructure/
https://developer.nvidia.com/blog/nvidia-exemplar-cloud-lessons-for-unlocking-full-performance-on-ai-infrastructure/
NVIDIA Technical Blog
NVIDIA Exemplar Cloud: Lessons for Unlocking Full Performance on AI Infrastructure
Two AI computing clusters built from identical NVIDIA H100, GB200 NVL72, or GB300 NVL72 systems can deliver materially different training throughput. We routinely see 8% to 12%
Four Ways to Deploy More Secure AI Agents
https://developer.nvidia.com/blog/four-ways-to-deploy-more-secure-ai-agents/
https://developer.nvidia.com/blog/four-ways-to-deploy-more-secure-ai-agents/
NVIDIA Technical Blog
Four Ways to Deploy More Secure AI Agents
Knowledge workers are increasingly integrating AI agents into their workflows. Agents that function as “digital coworkers” offer clear benefits. For example, they can review a bug report…
Run High-Performance Core Math at Scale with NVIDIA nvmath-python
https://developer.nvidia.com/blog/run-high-performance-core-math-at-scale-with-nvidia-nvmath-python/
https://developer.nvidia.com/blog/run-high-performance-core-math-at-scale-with-nvidia-nvmath-python/
NVIDIA Technical Blog
Run High-Performance Core Math at Scale with NVIDIA nvmath-python
NVIDIA nvmath-python is a library designed to bridge the gap between the Python scientific community and NVIDIA CUDA-X math libraries. It gives Python users access to CUDA-X performance for common…
NVIDIA Video Codec SDK 13.1: Zero-Copy Transcode, AV1 B-Frames, and Frame-Accurate Seek
https://developer.nvidia.com/blog/nvidia-video-codec-sdk-13-1-zero-copy-transcode-av1-b-frames-and-frame-accurate-seek/
https://developer.nvidia.com/blog/nvidia-video-codec-sdk-13-1-zero-copy-transcode-av1-b-frames-and-frame-accurate-seek/
NVIDIA Technical Blog
NVIDIA Video Codec SDK 13.1: Zero-Copy Transcode, AV1 B-Frames, and Frame-Accurate Seek
The demand for high-quality video continues to accelerate across industries, powering everything from immersive streaming experiences to remote collaboration, generative AI media tools…
Co-Designing AI Model Attention for Fast, Interactive Long-Context Inference
https://developer.nvidia.com/blog/co-designing-ai-model-attention-for-fast-interactive-long-context-inference/
https://developer.nvidia.com/blog/co-designing-ai-model-attention-for-fast-interactive-long-context-inference/
NVIDIA Technical Blog
Co-Designing AI Model Attention for Fast, Interactive Long-Context Inference
As agentic and long-context workloads become common, the context lengths increase and attention consumes a larger share of inference time (Figure 1). Because attention now dominates that cost…
How to Run Isolated Tenant Kubernetes Clusters on Shared GPU Infrastructure
https://developer.nvidia.com/blog/how-to-run-isolated-tenant-kubernetes-clusters-on-shared-gpu-infrastructure/
https://developer.nvidia.com/blog/how-to-run-isolated-tenant-kubernetes-clusters-on-shared-gpu-infrastructure/
NVIDIA Technical Blog
How to Run Isolated Tenant Kubernetes Clusters on Shared GPU Infrastructure
Running a dedicated Kubernetes cluster per team often results in more isolation than an organization requires. While one cluster can be successfully shared across many teams…
👍17
AI Leaders Propose SAFE Guidelines for Cybersecurity Transparency
https://blogs.nvidia.com/blog/open-secure-ai-alliance-contributions/
https://blogs.nvidia.com/blog/open-secure-ai-alliance-contributions/
NVIDIA Blog
AI Leaders Propose SAFE Guidelines for Cybersecurity Transparency
Members of the Open Secure AI Alliance — now more than 120 organizations strong — are developing new guidelines to strengthen agentic AI cybersecurity as the annual Black Hat conference begins in Las Vegas today. The Linux Foundation today shared a Request…
NVIDIA Alpamayo 2 Super, the Frontier Open Model for Robotaxis and Autonomous Vehicles, Now Available for Commercial Use
https://blogs.nvidia.com/blog/alpamayo-2-super-open-model-now-available/
https://blogs.nvidia.com/blog/alpamayo-2-super-open-model-now-available/
NVIDIA Blog
NVIDIA Alpamayo 2 Super, the Frontier Open Model for Robotaxis and Autonomous Vehicles, Now Available for Commercial Use
Open commercial licensing, benchmark‑leading reasoning and inspectable decisions bring autonomous vehicles, including robotaxis, closer to production and widescale deployment.
NVIDIA Joins NSF State and Regional AI Hubs Program to Expand AI Research and Education Across the US
https://blogs.nvidia.com/blog/nsf-state-regional-ai-hub-program/
https://blogs.nvidia.com/blog/nsf-state-regional-ai-hub-program/
NVIDIA Blog
NVIDIA Joins NSF State and Regional AI Hubs Program to Expand AI Research and Education Across the US
NVIDIA is participating in the NSF's State and Regional Artificial Intelligence Infrastructure Hubs program, an effort launching today to expand access to the advanced computing, data, software and expertise needed for AI-enabled research and education.
Generate Trajectories, Reasoning Traces, and Auto-Labels with NVIDIA Alpamayo 2 Super
https://developer.nvidia.com/blog/generate-trajectories-reasoning-traces-and-auto-labels-with-nvidia-alpamayo-2-super/
https://developer.nvidia.com/blog/generate-trajectories-reasoning-traces-and-auto-labels-with-nvidia-alpamayo-2-super/
NVIDIA Technical Blog
Generate Trajectories, Reasoning Traces, and Auto-Labels with NVIDIA Alpamayo 2 Super
Autonomous vehicle (AV) development often relies on separate models for trajectory generation, high-level intent prediction, scene understanding, and data labeling. This separation makes it hard to…
Beyond VLAs: How World Action Models Reshape Robot Manipulation
https://developer.nvidia.com/blog/beyond-vlas-how-world-action-models-reshape-robot-manipulation/
https://developer.nvidia.com/blog/beyond-vlas-how-world-action-models-reshape-robot-manipulation/
NVIDIA Technical Blog
Beyond VLAs: How World Action Models Reshape Robot Manipulation
A central challenge in robotics is building policies that generalize beyond the demonstrations they’re trained on. A policy that succeeds in a training scene often fails when object shapes, positions…
Into the Omniverse: How Open World Models Push the Frontier of Physical AI
https://blogs.nvidia.com/blog/open-world-models-physical-ai/
https://blogs.nvidia.com/blog/open-world-models-physical-ai/
NVIDIA Blog
Into the Omniverse: How Open World Models Push the Frontier of Physical AI
Open models, which anyone can download, inspect, modify and run on their own infrastructure, are what make that possible. Nowhere is that more crucial than in physical AI, where every deployment is a specialization problem.
👍3
Firebird Launches CIS Region’s Largest AI Factory in Armenia
https://blogs.nvidia.com/blog/firebird-ai-factory-armenia-blackwell-rubin-dsx/
https://blogs.nvidia.com/blog/firebird-ai-factory-armenia-blackwell-rubin-dsx/
NVIDIA Blog
Firebird Launches CIS Region’s Largest AI Factory in Armenia
The global buildout of AI infrastructure reached a new milestone today — Firebird, an emerging AI cloud, launched the CIS region’s largest AI factory in Armenia, establishing a new AI computing hub powered by NVIDIA accelerated computing and Dell Technologies…
👍11
Run Local Agentic AI Workflows with Meta’s Muse Glimmer on NVIDIA
https://developer.nvidia.com/blog/run-local-agentic-ai-workflows-with-metas-muse-glimmer-on-nvidia/
https://developer.nvidia.com/blog/run-local-agentic-ai-workflows-with-metas-muse-glimmer-on-nvidia/
NVIDIA Technical Blog
Run Local Agentic AI Workflows with Meta’s Muse Glimmer on NVIDIA
Meta returns to the open source ecosystem with the release of Muse Glimmer, a 30B open-weight dense model with a 120K+ context window built for local AI agentic work. Optimized to run across a range…
Route AI Agent Workloads Across Models with NVIDIA NeMo Switchyard
https://developer.nvidia.com/blog/route-ai-agent-workloads-across-models-with-nvidia-nemo-switchyard/
https://developer.nvidia.com/blog/route-ai-agent-workloads-across-models-with-nvidia-nemo-switchyard/
NVIDIA Technical Blog
Route AI Agent Workloads Across Models with NVIDIA NeMo Switchyard
Learn how NVIDIA NeMo Switchyard routes AI agent workloads across models using tuning-free and tunable routers that balance model capability, cost, and latency.
NVIDIA Nemotron 3.5 Lightning Delivers Fast, Accurate Specialized Task Execution for Long-Running Agents
https://developer.nvidia.com/blog/nvidia-nemotron-3-5-lightning-delivers-fast-accurate-specialized-task-execution-for-long-running-agents/
https://developer.nvidia.com/blog/nvidia-nemotron-3-5-lightning-delivers-fast-accurate-specialized-task-execution-for-long-running-agents/
NVIDIA Technical Blog
NVIDIA Nemotron 3.5 Lightning Delivers Fast, Accurate Specialized Task Execution for Long-Running Agents
Long-running AI agents spend most of their time on high-volume execution: tool calls, result validation, and subagent delegation. Using a frontier reasoning model for every execution step adds cost…
NVIDIA Nemotron 3.5 Lightning and NeMo Switchyard Deliver Faster, Smarter, More Efficient Agentic AI
https://blogs.nvidia.com/blog/nemotron-lightning-switchyard-rtx-dgx/
https://blogs.nvidia.com/blog/nemotron-lightning-switchyard-rtx-dgx/
NVIDIA Blog
NVIDIA Nemotron 3.5 Lightning and NeMo Switchyard Deliver Faster, Smarter, More Efficient Agentic AI
The new lightweight open model and routing library delivers greater control over AI, data and workflows across edge devices, PCs, workstations, data centers and the cloud.