→ Back to Home
Generative AI

OpenAI unveils Jalapeño chip, aiming to reduce Nvidia reliance for LLM inference

OpenAI has officially introduced its inaugural custom-designed artificial intelligence chip, dubbed Jalapeño, a strategic collaboration with Broadcom. This new silicon is specifically tailored for Large Language Model (LLM) inference, marking a significant step in OpenAI's ambition to bolster its end-to-end technical capabilities and reduce its dependency on third-party hardware providers, particularly Nvidia. The Jalapeño chip is not a general-purpose accelerator adapted from existing AI hardware; instead, it has been meticulously designed from scratch with LLM inference in mind. OpenAI emphasized that the chip's architecture is informed by the operational demands of its own flagship products, such as ChatGPT and Codex, while also being compatible with other LLMs across the industry. The primary goal is to achieve a blend of high throughput and power efficiency, making it ideal for scalable, interactive LLM applications. According to OpenAI, the development process for Jalapeño was remarkably swift, with the chip moving from concept to tape-out in just nine months. This accelerated timeline was reportedly aided by the use of OpenAI's own AI models in the design phase, showcasing a synergistic application of AI in hardware development. Broadcom's role in the partnership involves silicon implementation, networking, and connectivity technologies, while Celestica provides expertise in board, rack, and system-level integration. The companies envision Jalapeño as the foundational element of a multi-generation compute platform, with initial large-scale deployments slated for the end of 2026 in data centers operated by Microsoft and other key partners. This initiative is expected to expand full-stack platform capabilities and significantly lower inference costs for AI models. While Jalapeño is optimized for inference, the industry anticipates that OpenAI will continue to rely on Nvidia's GPUs for more computationally intensive tasks, such as large-scale model pre-training, for the foreseeable future. However, the introduction of Jalapeño signals a clear direction for OpenAI to gain greater control over its AI infrastructure and drive innovation in specialized hardware for generative AI.
#openai#broadcom#ai chip#llm#inference#machine learning infrastructure
Read original source