OpenAI and Broadcom announce chip designed for LLM inference at scale
… OpenAI claims that “early testing shows that Jalapeño will deliver performance per watt substantially better than current state-of-the-art,” but notes that it is not done measuring performance, and that a “detailed technical report will be presented in the coming months.” Until then, we don’t have … …