OpenAI and Broadcom announce LLM inference chip
OpenAI and Broadcom have announced a chip designed for LLM inference at scale, targeting the workload that runs trained large language models in production. The collaboration points to growing pressure on AI infrastructure as demand for model-driven services continues to test available compute capacity.
The move places the companies in a widening silicon race, where specialized hardware is becoming central to supporting large-scale AI deployments. The announcement frames inference, rather than model training, as the focus, signaling attention on the operational side of serving LLMs to users at high volume.
Originally reported by devtalk.comRead the source →
Related coverage
Leaked EU proposal would broaden AI access to personal data
1 day ago
Republicans split over AI guardrails as tech leaders warn of risks
3 days ago
AI safety moves and infrastructure deals dominate venture news
3 days ago
New AI releases push agents from chat to work
5 days ago