Business Wire
BusinessAMD says Instinct AI systems are now operating in Saudi Arabia, highlighting a potential ramp tied to additional data-center powerThe Apex TimesBusinessCostco and Old Navy promotions, Apple leadership change, and other retail and tech themes surfaced in a market roundupThe Apex TimesBusinessDeere named among stocks making notable moves in late-Thursday trading recapThe Apex TimesBusinessSalesforce says AI-driven revenue momentum is building as Agentforce adoption spreadsThe Apex TimesBusinessSalesforce backs HiBob to bolster workforce AI, and adds a new AgentExchange email toolThe Apex TimesBusinessEverPass Media expands NFL distribution via multi-year Netflix deal for 2026 slateThe Apex TimesBusinessUnitedHealth shares rise as it moves to drop prior-authorization checks for about 30% of servicesThe Apex TimesBusinessBerkshire Hathaway CEO Greg Abel to Appear on TV in Rare Interview, With Focus Likely on Insurance and BNSFThe Apex TimesBusinessCoinbase expands Webull crypto trading footprint into CanadaThe Apex TimesBusinessBroadcom leans harder into VMware AI with a push aimed at enterprise rivalsThe Apex TimesBusinessModerna shares jump after GSK advances a rival mRNA flu vaccine to Phase IIIThe Apex TimesBusinessYahoo Finance points to “buy zones” for Microsoft, Palantir, Shopify and ServiceNowThe Apex TimesBusinessAMD says Instinct AI systems are now operating in Saudi Arabia, highlighting a potential ramp tied to additional data-center powerThe Apex TimesBusinessCostco and Old Navy promotions, Apple leadership change, and other retail and tech themes surfaced in a market roundupThe Apex TimesBusinessDeere named among stocks making notable moves in late-Thursday trading recapThe Apex TimesBusinessSalesforce says AI-driven revenue momentum is building as Agentforce adoption spreadsThe Apex TimesBusinessSalesforce backs HiBob to bolster workforce AI, and adds a new AgentExchange email toolThe Apex TimesBusinessEverPass Media expands NFL distribution via multi-year Netflix deal for 2026 slateThe Apex TimesBusinessUnitedHealth shares rise as it moves to drop prior-authorization checks for about 30% of servicesThe Apex TimesBusinessBerkshire Hathaway CEO Greg Abel to Appear on TV in Rare Interview, With Focus Likely on Insurance and BNSFThe Apex TimesBusinessCoinbase expands Webull crypto trading footprint into CanadaThe Apex TimesBusinessBroadcom leans harder into VMware AI with a push aimed at enterprise rivalsThe Apex TimesBusinessModerna shares jump after GSK advances a rival mRNA flu vaccine to Phase IIIThe Apex TimesBusinessYahoo Finance points to “buy zones” for Microsoft, Palantir, Shopify and ServiceNowThe Apex TimesBusinessAMD says Instinct AI systems are now operating in Saudi Arabia, highlighting a potential ramp tied to additional data-center powerThe Apex TimesBusinessCostco and Old Navy promotions, Apple leadership change, and other retail and tech themes surfaced in a market roundupThe Apex TimesBusinessDeere named among stocks making notable moves in late-Thursday trading recapThe Apex TimesBusinessSalesforce says AI-driven revenue momentum is building as Agentforce adoption spreadsThe Apex TimesBusinessSalesforce backs HiBob to bolster workforce AI, and adds a new AgentExchange email toolThe Apex TimesBusinessEverPass Media expands NFL distribution via multi-year Netflix deal for 2026 slateThe Apex TimesBusinessUnitedHealth shares rise as it moves to drop prior-authorization checks for about 30% of servicesThe Apex TimesBusinessBerkshire Hathaway CEO Greg Abel to Appear on TV in Rare Interview, With Focus Likely on Insurance and BNSFThe Apex TimesBusinessCoinbase expands Webull crypto trading footprint into CanadaThe Apex TimesBusinessBroadcom leans harder into VMware AI with a push aimed at enterprise rivalsThe Apex TimesBusinessModerna shares jump after GSK advances a rival mRNA flu vaccine to Phase IIIThe Apex TimesBusinessYahoo Finance points to “buy zones” for Microsoft, Palantir, Shopify and ServiceNowThe Apex TimesBusinessAMD says Instinct AI systems are now operating in Saudi Arabia, highlighting a potential ramp tied to additional data-center powerThe Apex TimesBusinessCostco and Old Navy promotions, Apple leadership change, and other retail and tech themes surfaced in a market roundupThe Apex TimesBusinessDeere named among stocks making notable moves in late-Thursday trading recapThe Apex TimesBusinessSalesforce says AI-driven revenue momentum is building as Agentforce adoption spreadsThe Apex TimesBusinessSalesforce backs HiBob to bolster workforce AI, and adds a new AgentExchange email toolThe Apex TimesBusinessEverPass Media expands NFL distribution via multi-year Netflix deal for 2026 slateThe Apex TimesBusinessUnitedHealth shares rise as it moves to drop prior-authorization checks for about 30% of servicesThe Apex TimesBusinessBerkshire Hathaway CEO Greg Abel to Appear on TV in Rare Interview, With Focus Likely on Insurance and BNSFThe Apex TimesBusinessCoinbase expands Webull crypto trading footprint into CanadaThe Apex TimesBusinessBroadcom leans harder into VMware AI with a push aimed at enterprise rivalsThe Apex TimesBusinessModerna shares jump after GSK advances a rival mRNA flu vaccine to Phase IIIThe Apex TimesBusinessYahoo Finance points to “buy zones” for Microsoft, Palantir, Shopify and ServiceNowThe Apex Times
Back to front
Nvidia says Groq 3 LPX inference chip is now in production, with Nebius planned as a launch customer before year-end
The Apex Times

THE APEX TIMES

Business/The Apex Times/Aug 24, 7:16 PM EDT

Nvidia says Groq 3 LPX inference chip is now in production, with Nebius planned as a launch customer before year-end

The move, described as part of a roughly $20 billion Groq push dating back about eight months, outlines Nvidia is accelerating its in-house approach to powering AI inference workloads, not just training.

Nvidia NVDA said its Groq 3 LPX inference chip is now in production and that Nebius has been lined up as the launch customer before the end of 2026, according to a report published Tuesday by Yahoo Finance.

The announcement is the latest milestone in what Nvidia has described as a major, fast-moving bet on Groq technology. The report frames it as an eight-month sprint that began after Nvidia was reported to be pursuing an approximately $20 billion Groq deal, with the company now moving from planning to hardware availability.

Groq 3 LPX is positioned as an “inference” chip, meaning it is aimed at the stage of an AI workload where models generate answers from prompts and not the earlier “training” phase where models learn from large data sets. Inference is increasingly where enterprises are spending to run models at scale, especially once a model is selected and deployed.

By naming Nebius as the launch customer, Nvidia is also providing a concrete path for early deployment. Launch customers typically serve as first reference sites and can help shorten the timeline between chip availability and real-world system integration, even when volume ramp details are not disclosed publicly.

The report does not spell out performance specifications, manufacturing capacity, or pricing for Groq 3 LPX. It also does not provide any breakdown of what portions of Nvidia’s broader infrastructure stack will be used alongside the chip, beyond the production status and the timing expectation for Nebius.

Still, the timing matters for Nvidia’s broader strategy. Over the past year, investors have increasingly focused on whether semiconductor leaders can maintain momentum not only in accelerators for model training, but also in the systems needed to run inference efficiently across cloud providers and enterprise deployments.

Sector context is also important: AI spending is shifting toward “time-to-answer” and total cost of ownership, where inference efficiency can directly affect how many requests a system can handle per dollar. A chip moving into production suggests Nvidia is working to reduce friction for deployments rather than keeping Groq technology in the research or pilot phase.

What remains unclear is how broadly Groq 3 LPX will be rolled out after Nebius. Nvidia did not, in the reporting cited here, disclose forecasted shipment quantities, qualification timelines for additional customers, or whether future configurations will target particular model types, data center power constraints, or specific inference software stacks.

Why It Matters

  • If Groq 3 LPX is truly ready for production, Nvidia can potentially accelerate inference deployments, which are increasingly central to enterprise AI spending.
  • Naming a launch customer reduces uncertainty around early adoption timing, though broader rollout details were not provided.
  • Inference hardware availability can affect how quickly customers can scale model usage while controlling cost and latency.
  • The development reinforces Nvidia’s push to compete across more than training accelerators, expanding influence across the inference layer of AI systems.

Sources

Key Facts

  • Nvidia said Groq 3 LPX, an inference-focused chip, is now in production.
  • The company indicated Nebius is set to serve as the launch customer before the end of 2026.
  • The report characterizes Nvidia’s Groq effort as an eight-month push following a reported roughly $20 billion Groq bet.
  • The cited report does not include disclosed chip specifications, pricing, or production volume details.
  • The announcement points to a move from planning toward deployable hardware with a named early customer.

Technology Related