NVIDIA XR AI
NVIDIA XR AI Deploy enterprise-grade AI agents that see, understand, and augment your hands-on workforce. Now available in public beta. What Is NVIDIA XR AI? NVIDIA XR AI brings enterprise…
With agents running 24 hours a day, seven days a week on increasingly complex tasks, efficient local compute matters even more. NVIDIA has collaborated with the open source community to enhance the top inference backends for agents, llama.cpp and vLLM. llama.cpp now delivers 2x performance on Qwen 3.5 and 3.6 27B dense models, and 1.6x performance on Qwen 3.5 and 3.6 35B mixture-of-expert (MoE) models. The following two techniques make this possible: Multi-Token Prediction (MTP): An advanced speculative decoding technique, where a smaller draft model proposes several tokens ahead that the targ
Build Personal AI Agents on Windows PCs with New Tools from Microsoft and NVIDIA | NVIDIA Technical BlogAA-AgentPerf is a hardware benchmark created by Artificial Analysis that measures the number of concurrent AI agents an inference system can support while meeting predefined, model-specific performance service level objective (SLO) tiers. An SLO is defined as a specific threshold of output token speed and time-to-first-token (TTFT). The benchmark results are normalized per accelerator and per megawatt to enable comparison across hardware configurations.
NVIDIA Achieves Leading Agentic Coding Performance on First Agentic AI Benchmark | NVIDIA Technical BlogEarlier this week at GTC Taipei, NVIDIA unveiled the NVIDIA RTX Spark product family, including small form factor desktops and laptops built for the age of personal assistants. These desktops and laptops deliver 1 petaflop of AI power, up to 128 GB of memory, and CUDA-accelerated AI frameworks for running large models alongside everyday work. Microsoft is creating an RTX Spark special developer edition—the Microsoft Surface NVIDIA RTX Spark Dev Box—preloaded with a modified Windows configured for developers and the top developer tools you need to get started. To learn more, see Building the n
Build Personal AI Agents on Windows PCs with New Tools from Microsoft and NVIDIA | NVIDIA Technical BlogOne popular way to run AI locally has been to use multiple GPUs to access more memory and compute. While cloud frameworks like vLLM are well optimized for multiple GPUs thanks to their use in data centers, PC frameworks like llama.cpp and the ComfyUI implementation in PyTorch are not optimized for it. To solve this challenge, NVIDIA has collaborated with both llama.cpp and ComfyUI to enhance performance for RTX PCs with two equivalent GPUs. This enables you to run larger models and use the compute of both GPUs for better performance. llama.cpp now supports tensor parallelism (TP), fully utiliz
Build Personal AI Agents on Windows PCs with New Tools from Microsoft and NVIDIA | NVIDIA Technical BlogNVIDIA XR AI Deploy enterprise-grade AI agents that see, understand, and augment your hands-on workforce. Now available in public beta. What Is NVIDIA XR AI? NVIDIA XR AI brings enterprise…
…For developers building advanced in-cabin AI assistants or robotic dialogue agents, deploying highly capable language models at the edge presents a significant memory and latency challenge. Nemotron 2 Nano addresses this…
AI agents have changed a lot in the last two years. The first could only answer one question at a time. Then came multi-turn chat, where the model could keep some…
…Choose the Nsight AI Tool for Your Workflow Connect an AI coding agent to current CUDA knowledge, deploy a self-hosted AI assistant, or bring AI-guided performance analysis into your Nsight…
Agentic AI / Generative AI NVIDIA Ising Introduces AI-Powered Workflows to Build Fault-Tolerant Quantum Systems Apr 14, 2026 By Tom Lubowe , Christopher Chamberland , Shuxiang Cao , Ivan Basov and Jan Olle Discuss…
Agentic AI / Generative AI Deploy Long-Context Reasoning and Agentic Workflows with MiniMax M3 on NVIDIA Accelerated Infrastructure Jun 12, 2026 By Anu Srivastava Discuss (0) Discuss (0) L T F R…
AR / VR Building AI Agents for AR Glasses and XR Devices with NVIDIA XR AI Jun 16, 2026 By Greg Barbone Discuss (0) Discuss (0) L T F R E AI-Generated…
…Discuss (0) Discuss (0) Tags Agentic AI / Generative AI | Computer Vision / Video Analytics | Simulation / Modeling / Design | General | Cosmos | Metropolis | NIM | TAO Toolkit | Intermediate Technical | Tutorial | AI Agent | Fine-Tuning | Physical AI | Training…
Agentic AI / Generative AI Build AI-Ready Knowledge Systems Using 5 Essential Multimodal RAG Capabilities Feb 17, 2026 By Shruthii Sathyanarayanan , Sumit Bhattacharya , Punit Kumar , Pranjal Doshi and Nikhil Kulkarni Discuss (1…
Agentic AI / Generative AI Build a Retrieval-Augmented Generation (RAG) Agent with NVIDIA Nemotron Sep 23, 2025 By Edward Li , Vanessa Bellotti , Ryan Kraus and Rebecca Kao Discuss (0) Discuss (0) L…