THE APEX TIMES
Anthropic’s Claude on Microsoft Foundry becomes generally available on NVIDIA GB300 in Azure
NVIDIA says Claude models accessed through Microsoft’s Foundry platform can now run on its next-generation Blackwell Ultra GB300 systems, with a reference architecture aimed at deploying governed, autonomous AI agents in enterprise environments.
Anthropic’s Claude models, delivered through Microsoft Foundry and hosted on Microsoft Azure, have moved into general availability on NVIDIA’s GB300 Blackwell Ultra GPU infrastructure, according to NVIDIA. The change is positioned as a way for Azure-native enterprises to build and run more capable “agentic” AI systems, including autonomous and domain-specific sub-agents that can tackle multi-step business tasks.
In NVIDIA’s announcement, the company links improved inference performance and efficiency to lower total cost of ownership for enterprises deploying large-scale AI. “Inference” refers to the compute used to run a trained model on new inputs, rather than the training process. In NVIDIA’s framing, faster or more efficient inference can matter for both responsiveness and ongoing operating expenses when AI assistants are used continuously.
NVIDIA said Claude in Microsoft Foundry running on NVIDIA GB300 is deployed on NVIDIA NVL72 systems, paired with NVIDIA Quantum-X800 InfiniBand networking. It also says this configuration is intended to support “more powerful agentic systems,” including specialized agents that can operate across business domains.
A key part of the push is how developers can connect Claude to enterprise capabilities. NVIDIA said it is working with Anthropic to extend developer tooling by integrating NVIDIA tools into the Anthropic stack. The company described this as enabling “domain-specific abilities” for Claude agents through what NVIDIA calls NVIDIA verified agent skills, enabled by access to NVIDIA accelerated computing.
NVIDIA further pointed to deployment and governance as the practical hurdle for enterprise agents. Enterprises can run Claude agents on Azure using NVIDIA’s Secure Agent Workspace Reference Design, which NVIDIA describes as a blueprint for running autonomous agents in a governed environment. The design, as outlined by NVIDIA, aims to keep identity, network access, credentials, and runtime policy controlled at the infrastructure level.
The announcement also situates the GB300 availability within a broader partnership. NVIDIA said the expanded access is built on a strategic collaboration announced in November involving Microsoft, NVIDIA, and Anthropic, aimed at offering enterprise access to Claude and making Anthropic models available on NVIDIA-accelerated computing.
From a market perspective, the move underscores how the largest AI platform providers are pushing beyond chat-style assistants toward systems that can execute tasks with autonomy, while still being controlled inside corporate IT environments. For NVIDIA, adding GB300-backed capacity behind a major cloud distribution channel such as Azure is a way to align its newest data center platform with where enterprises are actually deploying models.
For Microsoft and its customers, the Foundry framing matters because it targets developers and enterprises already operating in Azure ecosystems. By making Claude available generally on NVIDIA GB300-backed infrastructure, the companies are effectively turning cloud compute choice and AI model access into a tighter bundle for building and scaling agent-based workloads.
Still, NVIDIA’s disclosure is largely focused on availability, architecture, and integration rather than performance benchmarks, pricing, or service-level details. The announcement does not provide specific latency, throughput, or cost figures for Claude on GB300, nor does it spell out the scope of “NVIDIA verified agent skills” beyond the concept of domain-specific abilities and the need for accelerated compute.
Looking ahead, what to watch is how widely enterprises adopt Claude through Foundry on GB300, and whether NVIDIA and Anthropic provide additional technical detail on the Secure Agent Workspace Reference Design, including how enterprises can operationalize identity and runtime controls for real-world agent deployments. Developers will also look for documentation that connects NVIDIA tools and verified agent skills to concrete application workflows in the Anthropic model stack.
Why It Matters
- Cloud distribution of Claude on the newest NVIDIA GPU platform can affect how quickly enterprises can prototype and scale agentic AI workloads within Azure environments.
- Efficient inference and supported networking and infrastructure configurations may influence the total cost and operational viability of running agents in production.
- The emphasis on governance and reference architecture suggests enterprises are prioritizing control, security, and policy enforcement as AI systems become more autonomous.
- Integration of developer tooling and “verified agent skills” indicates a push to make it easier to connect model capabilities to domain workflows rather than building everything from scratch.
Key Facts
- Claude models in Microsoft Foundry on Azure are now generally available on NVIDIA GB300 Blackwell Ultra GPU systems.
- NVIDIA said this deployment uses NVIDIA NVL72 systems and NVIDIA Quantum-X800 InfiniBand networking.
- The company linked the move to enabling more powerful agentic AI workloads, including autonomous and domain-specific sub-agents.
- NVIDIA said it is integrating NVIDIA tools into the Anthropic stack and enabling “NVIDIA verified agent skills.”
- Enterprises can run Claude agents on Azure using NVIDIA’s Secure Agent Workspace Reference Design, which focuses on governance controls such as identity, network access, credentials, and runtime policy.
Technology Related
Elon Musk’s chip preference spotlights Nvidia’s edge over AMD, but investors still watch execution
A Yahoo Finance analysis highlighted Nvidia’s faster growth relative to AMD, drawing attention to how high-profile tech users, including Elon Musk, frame the semiconductor race.
Ming-Chi Kuo says Nvidia has revived Rubin CPX after it seemingly vanished from the AI roadmap
The analyst Ming-Chi Kuo says Nvidia’s Rubin CPX accelerator is back, with what he characterizes as a substantial redesign after the chip appeared to be shelved earlier this year.
Apple’s next CEO arrives with a different kind of power: money, and an AI test
A new leadership chapter at Apple, as reported by Yahoo Finance, raises a central question for investors and customers alike: will Apple use its unusual financial profile to change its AI direction, or simply defend its status quo?
ZonPrep buys inbound-inventory software and services, betting on Amazon logistics automation
The Amazon-focused supply chain and FBA prep company says it acquired Wizard-Industries and FNSKU Studio, tools aimed at helping sellers get inventory into Amazon faster and with fewer process steps.
Nvidia pauses part of its AI customer financing after a strong quarter, raising questions about timing
After delivering another heavy AI-related quarter, Nvidia indicated it is stepping back from a portion of its financing approach for customers. Market coverage framed the move as potentially awkward, given investor expectations tied to continued momentum in AI infrastructure spending.
Apple CEO transition hands AI test to John Ternus as AAPL slips
John Ternus takes over as Apple’s chief executive role as Phil Schiller steps back, with market attention focused on how leadership changes could affect ongoing work on artificial intelligence initiatives. Apple shares slid in early trading following the transition reports.
Anthropic reportedly signs $35 billion cloud deal involving Nvidia-backed Lambda and a Texas data-center lease
A Yahoo Finance report says Anthropic has agreed to a long-term cloud-computing arrangement worth $35 billion, with the infrastructure and data-center lease tied to Lambda, an Nvidia-backed provider.
FTC and 22 states sue Amazon, alleging it overcharged advertisers using its retail platform
The U.S. Federal Trade Commission and a coalition of state attorneys general accused Amazon of misleading businesses about pricing tied to advertising on its shopping marketplace, alleging the conduct resulted in billions in gains for the company.
Intel’s push toward on-prem, privacy-focused AI gets a partnership spotlight as Xeon 6 platform work expands
A new extension to Kasm Technologies’ deal work with Intel highlights a market trend toward running large language model workloads locally on enterprise hardware, aiming to reduce data exposure and reliance on GPUs.
Broadcom (AVGO) set to report earnings Wednesday after the bell, with investors focused on guidance and demand outlines
The fabless chip and software maker Broadcom will release its next quarterly results this Wednesday after market close, according to a preview posted by Yahoo Finance.