← Back
NVIDIA Blog

Claude Meets Blackwell Ultra: Anthropic’s Models Now Run on NVIDIA GB300 in Azure

4 min read
#agents#anthropic#nvidia
TL;DR

Anthropic’s Claude models are now generally available on Microsoft Azure via Microsoft Foundry, running on NVIDIA’s new GB300 Blackwell Ultra GPUs. This partnership gives Azure‑native enterprises a high‑performance, low‑latency platform for building autonomous and domain‑specific AI agents. The deployment leverages Azure’s managed services while exposing the cutting‑edge GPU architecture for scalable inference, though it currently supports only Claude models and requires an Azure subscription with GPU access.

⚡ Key Takeaways

  • Claude models now run on NVIDIA GB300 Blackwell Ultra GPUs in Azure.
  • Deployment is managed through Microsoft Foundry’s Azure AI Studio.
  • Requires Azure subscription and GPU availability, limiting immediate access to organizations without GB300 capacity.
  • Deploy by selecting the GB300 GPU type in the Azure AI Studio deployment settings.
  • Currently limited to Claude; other LLMs are not yet supported on this GPU stack.
  • WhyItMatters: Engineers shipping production AI can now leverage Azure’s managed services combined with NVIDIA’s latest GPU architecture to achieve faster inference and lower latency for Claude‑based agents, directly impacting throughput and cost efficiency.
  • TechnicalLevel
💡 Why It Matters

Engineers shipping production AI can now leverage Azure’s managed services combined with NVIDIA’s latest GPU architecture to achieve faster inference and lower latency for Claude‑based agents, directly impacting throughput and cost efficiency. TechnicalLevel

Want the full story? Read the original article.

Read on NVIDIA Blog

More like this

57% of enterprises have watched AI agents be confidently wrong. The fix is an agentic context layer, but who has one?

VentureBeat AI#agents

Choosing the Right AI Agent Memory Strategy: A Decision-Tree Approach

Machine Learning Mastery#agents

Fine-tune NVIDIA Nemotron 3 models with Amazon SageMaker AI serverless model customization

AWS ML Blog#llm

NVIDIA Nemotron Achieves Benchmark-Leading Performance With LangChain Deep Agents Harness

NVIDIA Blog#agents

EXPLORE AI NEWS

Daily hand-picked stories on LLMs, RAG, agents and production AI — curated for engineers who ship.

BROWSE NEWS

GET THE WEEKLY DIGEST

Join engineers getting the Monday signal-over-noise AI breakdown. No spam, unsubscribe anytime.

LEARN AI ENGINEERING

Curated courses, research papers, repos and tutorials built for engineers leveling up in AI.

START LEARNING