New Accelerator from OpenAI and Broadcom Improves LLM Inference Performance Per Watt
Key Takeaways: OpenAI and Broadcom have taken another step into the AI infrastructure race with the debut of Jalapeño, a purpose-built inference accelerator engineered specifically around large language model worklo...