惯性聚合 高效追踪和阅读你感兴趣的博客、新闻、科技资讯
阅读原文 在惯性聚合中打开

推荐订阅源

量子位
Threat Intelligence Blog | Flashpoint
Threat Intelligence Blog | Flashpoint
雷峰网
雷峰网
酷 壳 – CoolShell
酷 壳 – CoolShell
博客园 - 司徒正美
N
News | PayPal Newsroom
WordPress大学
WordPress大学
freeCodeCamp Programming Tutorials: Python, JavaScript, Git & More
钛媒体:引领未来商业与生活新知
钛媒体:引领未来商业与生活新知
The Cloudflare Blog
S
Secure Thoughts
让小产品的独立变现更简单 - ezindie.com
让小产品的独立变现更简单 - ezindie.com
Security Archives - TechRepublic
Security Archives - TechRepublic
博客园 - 【当耐特】
博客园 - 聂微东
S
Securelist
宝玉的分享
宝玉的分享
爱范儿
爱范儿
IT之家
IT之家
T
The Exploit Database - CXSecurity.com
S
SegmentFault 最新的问题
Exploit-DB.com RSS Feed
Exploit-DB.com RSS Feed
AI
AI
Security Latest
Security Latest
博客园 - 叶小钗
OSCHINA 社区最新新闻
OSCHINA 社区最新新闻
The Last Watchdog
The Last Watchdog
月光博客
月光博客
D
Darknet – Hacking Tools, Hacker News & Cyber Security
S
Schneier on Security
人人都是产品经理
人人都是产品经理
Webroot Blog
Webroot Blog
Jina AI
Jina AI
阮一峰的网络日志
阮一峰的网络日志
J
Java Code Geeks
N
News and Events Feed by Topic
Recent Commits to openclaw:main
Recent Commits to openclaw:main
奇客Solidot–传递最新科技情报
奇客Solidot–传递最新科技情报
Last Week in AI
Last Week in AI
S
Security Affairs
美团技术团队
Hugging Face - Blog
Hugging Face - Blog
V
V2EX
罗磊的独立博客
Spread Privacy
Spread Privacy
Help Net Security
Help Net Security
T
Tailwind CSS Blog
C
Cybersecurity and Infrastructure Security Agency CISA
博客园_首页
Apple Machine Learning Research
Apple Machine Learning Research

NVIDIA Blog

GeForce NOW Turns Up the Heat With New GeForce RTX 5080-Powered Toronto Server NVIDIA Nemotron Achieves Benchmark-Leading Performance With LangChain Deep Agents Harness AI Innovators Adopt NVIDIA Vera — Why Max Single-Threaded CPU at Scale Matters NVIDIA and Hugging Face Bring New Models and Frameworks to LeRobot for the Open Robotics Community How Open Models Are Driving AI Research How Nations Are Deploying AI for Strategic Priorities Joyride Through July With 12 Games Coming to GeForce NOW NVIDIA Unlocks AI Compute at Scale, Inviting Partners to Power the AI Infrastructure Buildout NVIDIA and Partners Build in America, for America NVIDIA BioNeMo Agent Toolkit Brings Accelerated AI to Life Sciences Researchers in Claude Science How NVIDIA’s Inference Software Stack Powers the Lowest Token Cost How Jaiveer Singh Is Helping Robots — and Developers — Move Faster Into the Omniverse: Three Workflows for Improving Vision AI Agent Accuracy With Synthetic Data and Fine-Tuning Claude Meets Blackwell Ultra: Anthropic’s Models Now Run on NVIDIA GB300 in Azure Firefly Aerospace Operates NVIDIA Jetson in Lunar Orbit for the First Time Open Models, Closed Environments: Palantir Brings Secure AI to US Agencies With NVIDIA Nemotron The Ultimate Summer Sale Pairing: Steam Sale Meets GeForce NOW Discounts NVIDIA and AWS Collaborate to Bring AI to Production at Scale How Businesses Are Building Specialized AI They Can Trust NVIDIA Powers Over 400 of the World’s 500 Fastest Supercomputers NVIDIA Brings Trusted, 24/7 AI Agents to Telecom Operations At ISC, JUPITER Shows What Exascale Science Looks Like NAIRR Science Program Reshapes Scientific Research, Powered by NVIDIA AI Infrastructure From Materials Simulation to Experimental Astronomy, New NVIDIA AI Software Unlocks Scientific Discoveries NVIDIA Vera CPU Opens the Way for Agentic Scientific AI at Los Alamos National Laboratory Eco Wave Power Turns Waves Into Watts With NVIDIA AI Infrastructure and Digital Twins Hotter Than a Hot Tub: The 45°C Breakthrough to Cool AI’s Biggest Machines How FERC’s Large-Load Interconnection Actions Help Address Grid Stress, Improve Affordability At Cannes Lions, NVIDIA Partners Reshape Advertising and Marketing With AI Sync and Stream: GeForce NOW Connects to Members’ Game Libraries Across Devices France Advances Europe’s AI Future With NVIDIA Technologies Hands Free, AIs Forward: NVIDIA XR AI Brings Agents to AR Glasses Coherent Breaks Ground on Expanded Texas Facility, Scaling AI’s Optical Backbone HPE AI Factory With NVIDIA Expands for the Era of Agents Fastest, Largest, Strongest: NVIDIA Blackwell Sweeps MLPerf Training 6.0 NVIDIA Blackwell Leads on First Agentic AI Infrastructure Benchmark Save Big and Play Bigger: GeForce NOW Summer Sale Brings Major Membership Savings For Robotaxis, Safety Must Be Built In, Not Bolted On NVIDIA Accelerates Google DeepMind’s DiffusionGemma for Local AI NVIDIA Confidential Computing to Help Expand Apple’s Private Cloud Compute How the UK Is Turning Sovereign AI Ambition Into Action With NVIDIA Technologies NVIDIA and LG Group Build an AI Factory to Advance Physical AI, Mobility and AI Infrastructure NVIDIA and Doosan Group Collaborate to Advance Physical AI and AI Factory Infrastructure NVIDIA, KRAFTON, NC and Reigning ‘League of Legends’ Champions T1 Celebrate RTX Spark at Korea’s PC Bangs Seoul Purpose: How NVIDIA and South Korea Are Building the Future of AI Forecast: Fun Ahead — 18 Games Join in June to Stream on GeForce NOW NVIDIA Research Unlocks Advanced Grasping, Smarter Autonomous Driving and Agent Training at Scale NVIDIA Enables the Next Era Of Physical AI Research With Agent Skills For Autonomous Vehicles, Robotics And Vision AI Industrial Software Leaders Build Secure, Autonomous AI Engineers With NVIDIA NemoClaw Why Financial Institutions Are Converging on Transaction Foundation Models to Build Their Own Intelligence NVIDIA Jetson Brings Agentic AI to the Physical World NVIDIA AI Cloud Ecosystem Expands Worldwide to Meet Global AI Compute Demand NVIDIA Factory Operations Blueprint Gives Factories a New AI Brain Taiwan’s Industry Titans Turbocharge World’s AI Infrastructure Buildout With NVIDIA How Cosmos 3 Helps Physical AI Think Before It Acts NVIDIA Levels Up Local AI Agents Across RTX PCs and DGX Spark NVIDIA Research Advances Robotics From Simulation to the Real World The Name’s Gaming … Cloud Gaming: ‘007 First Light’ Launches on GeForce NOW AI Factories: The New Infrastructure of Intelligence NVIDIA Vera CPU Is ‘Packing a Heavy-Hitting Punch’ Against Competition NVIDIA GTC Taipei at COMPUTEX: Live Updates on What’s Next in AI License to Stream: ‘007 First Light’ Coming to GeForce NOW With an Ultimate Bundle NVIDIA and Google Cloud Empower the Next Wave of AI Builders NVIDIA CEO Jensen Huang at Dell Technologies World: ‘Demand Is Going Parabolic, Utterly Parabolic’ Vera Arrives: NVIDIA’s First CPU Built for Agents Lands at Top AI Labs Sea You in the Cloud: ‘Subnautica 2’ Early Access Dives Onto GeForce NOW NVIDIA, Ineffable Intelligence Team Up to Build the Future of Reinforcement Learning Infrastructure Hermes Unlocks Self-Improving AI Agents, Powered by NVIDIA RTX PCs and DGX Spark NVIDIA and SAP Bring Trust to Specialized Agents Linked and Loaded: Gaijin Single Sign-On Now Available on GeForce NOW NVIDIA and ServiceNow Partner on New Autonomous AI Agents for Enterprises It’s Gonna Be May: 16 Games Hit the Cloud This Month, With More NVIDIA GeForce RTX 5080 Power NVIDIA Launches Nemotron 3 Nano Omni Model, Unifying Vision, Audio and Language for up to 9x More Efficient AI Agents Into the Omniverse: Manufacturing’s Simulation-First Era Has Arrived Tag, You’re It: GeForce NOW Levels Up Game Discovery With Xbox Game Pass and Ubisoft+ Labels Making Sense of the Early Universe From Rainforests to Recycling Plants: 5 Ways NVIDIA AI Is Protecting the Planet NVIDIA and Google Cloud Collaborate to Advance Agentic and Physical AI Autonomous AI at Scale: Adobe Agents Unlock Breakthrough Creative Intelligence With NVIDIA and WPP No Need for Space Gear — Capcom’s ‘PRAGMATA’ Joins GeForce NOW on Launch Day Rethinking AI TCO: Why Cost per Token Is the Only Metric That Matters New Adobe Premiere Color Grading Mode Accelerated on NVIDIA GPUs Strength and Destiny Collide: ‘Samson: A Tyndalston Story’ Arrives in the Cloud National Robotics Week — Latest Physical AI Research, Breakthroughs and Resources From RTX to Spark: NVIDIA Accelerates Gemma 4 for Local Agentic AI Press Start on April: GeForce NOW Brings 10 Games to the Cloud Efficiency at Scale: NVIDIA, Energy Leaders Accelerating Power‑Flexible AI Factories to Fortify the Grid Into the Omniverse: NVIDIA GTC Showcases Virtual Worlds Powering the Physical AI Era Game On: Five New Titles Now Streaming on GeForce NOW The Future of AI Is Open and Proprietary Blowing Off Steam: How Power-Flexible AI Factories Can Stabilize the Global Energy Grid Advancing Open Source AI, NVIDIA Donates Dynamic Resource Allocation Driver for GPUs to Kubernetes Community How Autonomous AI Agents Become Secure by Design With NVIDIA OpenShell NVIDIA GTC 2026: Live Updates on What’s Next in AI Smooth Moves: 90 Frames-Per-Second Virtual Reality Arrives on GeForce NOW From Simulation to Production: How to Build Robots With AI More Than Meets the Eye: NVIDIA RTX-Accelerated Computers Now Connect Directly to Apple Vision Pro NVIDIA, Telecom Leaders Build AI Grids to Optimize Inference on Distributed Networks GTC Spotlights NVIDIA RTX PCs and DGX Sparks Running Latest Open Models and AI Agents Locally Snap Decisions: How Open Libraries for Accelerated Data Processing Boost A/B Testing for Snapchat
NVIDIA Partners With Microsoft on Unified Stack for Agentic AI Deployment, From Windows Devices to Cloud to Local
Dave Salvator · 2026-06-03 · via NVIDIA Blog

The agentic AI moment has arrived, but delivering on its promise requires more than good models. It also takes fast hardware, secure runtimes, a responsive data layer and models tuned for long-running reasoning. NVIDIA and Microsoft are bringing that full stack to developers across Windows devices, Azure cloud and local deployments.

At Microsoft Build, NVIDIA founder and CEO Jensen Huang joined Microsoft chairman and CEO Satya Nadella’s keynote via livestream from Taipei to discuss the expanded partnership: NVIDIA RTX Spark and DGX Station for Windows, NVIDIA GPU-accelerated Microsoft Fabric, NVIDIA open models on Microsoft Foundry, the NVIDIA OpenShell secure runtime in GitHub Copilot and the next generation of NVIDIA-powered AI factories.

Reinventing Windows for Agents: From RTX Spark to DGX Station for Windows

NVIDIA and Microsoft are reimagining Windows PCs for the age of AI agents. With RTX Spark laptops and small desktops, and DGX Station for Windows deskside AI supercomputers, developers can build, tune and run agents natively on Windows.

RTX Spark is a new beginning, powering the world’s first Windows PCs purpose-built for personal agents, with 1 petaflop of AI performance, up to 128GB of unified memory, all-day battery life, and full AI and graphics performance unplugged. Bringing over 30 years of NVIDIA innovation, including CUDA, RTX, DLSS and TensorRT, systems arrive this fall from Microsoft Surface, ASUS, Dell, HP, Lenovo and MSI.

DGX Station for Windows is the most powerful deskside AI supercomputer for building and running agents on Windows enterprise applications and workflows. Powered by the NVIDIA GB300 Grace Blackwell Ultra Desktop Superchip with up to 748GB of coherent memory and 20 petaflops of FP4 performance, it runs frontier models of up to 1 trillion parameters for always-on enterprise agents. Systems are expected from ASUS, Dell, GIGABYTE, HP, MSI and Supermicro in Q4. Both products run NVIDIA OpenShell, a secure-by-design runtime for autonomous agents.

Read more in this Microsoft blog: “Introducing a powerful new chapter for Windows PCs, accelerated by NVIDIA RTX Spark

Powering Agentic Workflows at Enterprise Scale With NVIDIA Open Models on Microsoft Foundry

Agentic AI runs on a system of models. With NVIDIA, Anthropic and OpenAI models  plus Hermes special agents — now on the hosted agents in Foundry Agent Service, enterprises can bring agentic systems to life on Azure with built-in identity and governance. Anthropic’s Claude models now run natively on NVIDIA GB300 Blackwell Ultra systems on Azure, with customer availability in the weeks ahead.

NVIDIA Nemotron 3 Ultra, a new open frontier reasoning model for long-running agents across coding, research and enterprise workflows, is available this month on Foundry managed compute, alongside Nemotron 3.5 ASR for speech recognition and Nemotron 3.5 Content Safety. Developers can compose Nemotron alongside frontier and local models, optimizing cost and quality for each workflow.

NVIDIA’s open model portfolio on Foundry now spans agentic, physical and scientific AI. NVIDIA Cosmos 3, the first fully open omnimodel for physical AI, brings vision reasoning, world simulation and action generation. NVIDIA Earth-2 AI weather models are available through Microsoft Planetary Computer Pro and Foundry for enterprise forecasting and risk analysis.

NVIDIA Agent Toolkit and NVIDIA NemoClaw blueprints give developers an open source platform to build production agents on Foundry. NVIDIA CUDA-X libraries including cuDF, cuOpt, AI-Q and NeMo are now accessible to agents as domain-specific skills.

Learn more in this Build breakout session: “Orchestrate Special Agents with NVIDIA Nemotron Models on Microsoft Foundry.”

Accelerating Enterprise Data Warehouses for the AI Era

Data fuels agentic AI, and fast access to it is critical. 

NVIDIA accelerated computing is now built into Microsoft Fabric Data Warehouse, with Microsoft’s internal benchmarking delivering SQL execution up to 6x faster than the CPU-powered baseline and up to 7x faster than three other leading cloud data warehouse providers for high-concurrency workloads. 

The enterprise data layer can now keep pace with AI agents that continuously query and reason over data, the result of years of deep engineering collaboration between NVIDIA and Microsoft, from research to production.

Read more in this Microsoft blog: “Microsoft Build 2026: Building agentic apps with Microsoft Fabric and Microsoft Databases

Advancing Physical AI and Autonomous Systems

Physical AI is the next frontier for agents. 

Microsoft is integrating NVIDIA’s open source physical AI skills and tools with Azure and its Physical AI Toolchain. Developers get a unified platform, powered by Cosmos 3’s mixture-of-transformers architecture, to simulate, train and deploy autonomous systems, including robots, autonomous vehicles and industrial systems that can perceive, reason, plan and act in the physical world. Cosmos 3 ranks first among open models on key benchmarks for vision reasoning, world generation and action generation.

Enhancing Azure Local and Foundry Local With NVIDIA RTX PRO 6000 Blackwell Server Edition and Nemotron Models

Agentic AI is moving beyond the cloud. 

Microsoft is bringing Foundry Local on Azure Local to the NVIDIA RTX PRO 6000 Blackwell Server Edition platform. Paired with the NVIDIA Nemotron open model family, enterprises can run high-performance AI workloads where their data resides, whether in on-premises, hybrid or sovereign environments, without sacrificing performance or governance. 

Foundry Local on Azure Local now supports multinode deployments and the vLLM runtime, scaling inference for manufacturing, energy, sovereign data centers and other latency-sensitive scenarios.

Learn more in these Microsoft blogs: “Build, deploy and govern sovereign AI with Foundry Local on Azure Local” and “Scale On-Prem AI with Foundry Local on Azure Local.” 

Bringing Secure Agent Development to GitHub Copilot With NVIDIA OpenShell

As agents move from coding assistance to autonomous execution, they need real capability without real credentials. 

NVIDIA OpenShell, now integrated into GitHub Copilot, solves this: Each agent runs isolated in its own sandboxed container, and every outbound call is evaluated against policy before it can reach files, networks or credentials. Policies are written as code, versioned in the repository and updatable on the fly. OpenShell is open source under Apache 2.0, model-agnostic and spans on-premises, hybrid and cloud environments.

Learn more in this Build lightning session: “Secure Agent Workflows with GitHub Copilot and NVIDIA OpenShell.

Fairwater Wisconsin Goes Live, Validated for NVIDIA Vera Rubin

Microsoft’s Fairwater Wisconsin AI factory is now live, ahead of schedule, running hundreds of thousands of NVIDIA Grace Blackwell systems as a single AI factory, and connected with a similar AI factory in Georgia to deliver a scalable and distributed AI system for the most demanding frontier models. Through joint engineering on power, cooling, NVIDIA Spectrum-X Ethernet and the new Multipath Reliable Connection (MRC) transport protocol, Microsoft’s Fairwater AI data center designs are optimizing token economics.  

In addition, Microsoft has already validated the NVIDIA Vera Rubin platform, now in full production, for deployment across Azure data centers. 

Vera Rubin slots in alongside Blackwell with no retrofits, delivering up to 10x inference throughput per megawatt and reducing cost per agentic token by an order of magnitude. Built-in NVIDIA Confidential Computing protects models and data as agents reason at scale. The NVIDIA Dynamo inference framework extends those gains into software, accelerating model cold starts on AKS and bringing Kubernetes-native distributed inference orchestration via NVIDIA Grove.

Read more in this Microsoft blog: “Scaling multi-node LLM inference with NVIDIA Dynamo-Grove on AKS (Part 4)

Explore the full lineup of NVIDIA sessions, demos and hands-on labs at Microsoft Build.