NVIDIA announced the Groq 3 LPX as the interactive AI inference accelerator for the Vera Rubin platform. It is integrated with the NVIDIA Vera Rubin NVL72 component.
The accelerator focuses on improving inference speed and interactivity for models that require processing long context windows. This can enhance performance in applications demanding real-time responses.
Details on hardware specifications, such as size or performance metrics, are not provided. The platform aims to support efficient AI inference at scale.
