OpenAI and Broadcom have announced a new chip called Jalapeño, designed specifically for large language model inference in data centers. The companies plan to deploy the chip at large data centers, positioning it as the first generation in a long-term project that will see ongoing refinements. The chip aims to optimize LLM inference workloads at high volume, addressing the growing demand for efficient AI inference at scale.
OpenAI and Broadcom Unveil Jalapeño Chip for Large-Scale LLM Inference
vidgetc
Tech, gaming & AI news — always at hand
Google Play · Soon
App Store · Soon
Comments