OpenAI and Broadcom Unveil Jalapeño Chip for Large-Scale LLM Inference

OpenAI and Broadcom have announced a new chip called Jalapeño, designed specifically for large language model inference in data centers. The companies plan to deploy the chip at large data centers, positioning it as the first generation in a long-term project that will see ongoing refinements. The chip aims to optimize LLM inference workloads at high volume, addressing the growing demand for efficient AI inference at scale.

vidgetc Tech, gaming & AI news — always at hand Google Play · Soon App Store · Soon
💬 Discuss

Comments

Next articleWeb Scraper Declares 'Google and Reddit Do Not Own the Internet' After Court Victory
Start typing to search