This website uses cookies
We use cookies to personalise content and ads, to provide social media features and to analyse our traffic. We also share information about your use of our site with our social media, advertising and analytics partners who may combine it with other information that you’ve provided to them or that they’ve collected from your use of their services.
Consent Selection
Details
  • Necessary cookies help make a website usable by enabling basic functions like page navigation and access to secure areas of the website. The website cannot function properly without these cookies.
  • Preference cookies enable a website to remember information that changes the way the website behaves or looks, like your preferred language or the region that you are in.
    • We do not use cookies of this type.

  • Statistic cookies help website owners to understand how visitors interact with websites by collecting and reporting information anonymously.
    • We do not use cookies of this type.

  • Marketing cookies are used to track visitors across websites. The intention is to display ads that are relevant and engaging for the individual user and thereby more valuable for publishers and third party advertisers.
    • We do not use cookies of this type.

  • Unclassified cookies are cookies that we are in the process of classifying, together with the providers of individual cookies.
    • __emg_sidPending
      Maximum Storage Duration: 1 dayType: HTTP Cookie
      __emg_vidPending
      Maximum Storage Duration: 1 yearType: HTTP Cookie
      nl-read-countPending
      Maximum Storage Duration: PersistentType: HTML Local Storage
Cookie declaration last updated on 8/12/26 by Cookiebot
[#IABV2_TITLE#]
[#IABV2_BODY_INTRO#]
[#IABV2_BODY_LEGITIMATE_INTEREST_INTRO#]
[#IABV2_BODY_PREFERENCE_INTRO#]
[#IABV2_BODY_PURPOSES_INTRO#]
[#IABV2_BODY_PURPOSES#]
[#IABV2_BODY_FEATURES_INTRO#]
[#IABV2_BODY_FEATURES#]
[#IABV2_BODY_PARTNERS_INTRO#]
[#IABV2_BODY_PARTNERS#]
About
Cookies are small text files that can be used by websites to make a user's experience more efficient.

The law states that we can store cookies on your device if they are strictly necessary for the operation of this site. For all other types of cookies we need your permission.

This site uses different types of cookies. Some cookies are placed by third party services that appear on our pages.

You can at any time change or withdraw your consent from the Cookie Declaration on our website.

Learn more about who we are, how you can contact us and how we process personal data in our Privacy Policy.

Please state your consent ID and date when you contact us regarding your consent.
NewsLayer.com
NewsLayer PulseLIVEBTC$79,714+2.96%ETH$2,499+1.96%SOL$97.16+1.87%XRP$1.51+0.47%DOGE$0.092-0.67%ADA$0.2238-0.45%Total Cap$2.81T+2.40%Layer Index69 Greed

Nvidia's dedicated inference accelerator Groq 3 LPX enters full production to supercharge AI agents

Chipmaker Nvidia Corp. says its dedicated artificial intelligence inference accelerator Groq 3 LPX has now entered full production as it strives to maintain its dominance in the world of AI compute.

SiliconANGLE

Publisher

Aug 24, 2026 at 3:00 PM UTC · 5 分钟阅读

Nvidia's dedicated inference accelerator Groq 3 LPX enters full production to supercharge AI agents
Image via SiliconANGLE
翻译中…

Chipmaker Nvidia Corp. says its dedicated artificial intelligence inference accelerator Groq 3 LPX has now entered full production as it strives to maintain its dominance in the world of AI compute.

The new chip, announced today at Hot Chips 2026, is described as a purpose-built extension to Nvidia’s flagship Vera Rubin data center platform. According to Nvidia, it’s designed to deliver ultra-fast token generation speeds, which are necessary to run highly responsive agentic AI workloads. The chipmaker said the neocloud provider Nebius Group N.V. has already signed on as the first customer to commit to using the new chip.

Inference is the AI industry’s lingo for the process of running fully trained AI models in production, and it’s increasingly focused on autonomous AI agents that can perform tasks on behalf of humans. These agents must be able to do everything from reason and plan, write and execute code, inspect system files and use third-party tools in continuous loops. They can quickly crunch through thousands of tokens across these complex chains, but this sometimes results in massive “decode latency” that can create frustrating delays.

To prevent this from happening, AI data centers need more specialized compute architectures that can disaggregate the enormous context processing from token generation, in order to increase the speed at which AI agents can reason and work.

Article Intelligence

Topics

Sponsored

Ad
House — Advertise on NewsLayer
NewsLayerLearn more

NewsLayer Premium

Unlock deeper intelligence.

Ad-free reading, exclusive research, and real-time onchain insights.

Go Premium