→ Back to Home
Generative AI

OpenAI Unveils "Jalapeno" Custom AI Chip to Reduce Nvidia Dependence

OpenAI, the creator of ChatGPT, has officially unveiled its inaugural custom-built artificial intelligence chip, dubbed "Jalapeno." Developed in partnership with Broadcom, this new semiconductor represents a pivotal strategic initiative for OpenAI to lessen its dependence on Nvidia's widely used graphics processing units (GPUs). The announcement underscores a broader industry trend where major tech companies, including Google, Amazon, and Microsoft, are pursuing their own custom chip designs to optimize performance and manage the escalating costs associated with running large-scale AI models. The "Jalapeno" chip is specifically engineered for AI inference, which involves running pre-trained AI models to generate responses and execute commands. This phase is particularly resource-intensive and costly for generative AI applications. According to OpenAI, preliminary testing results for Jalapeno demonstrate a substantial improvement in performance per watt when compared to existing cutting-edge solutions. This efficiency gain is crucial for enhancing the profitability and scalability of OpenAI's services. A remarkable aspect of Jalapeno's development is that OpenAI utilized its own AI models to accelerate the chip's design process, significantly shortening the development timeline to just nine months. This rapid iteration highlights the potential of AI to even design its own foundational hardware. The chip is not exclusively designed for OpenAI's proprietary models but is intended to be compatible with a wide array of large language models (LLMs). Deployment of these new chips is anticipated to begin in data centers operated by Microsoft and other partners starting in late 2026. By controlling its own chip design and production, OpenAI aims to achieve end-to-end technical competitiveness across its entire technology stack. This integrated approach, encompassing models, products, and underlying infrastructure, is geared towards delivering AI models that are faster, more reliable, and more cost-effective. While Jalapeno is specialized for inference, OpenAI is expected to continue relying on Nvidia GPUs for more demanding tasks like large-scale model pre-training for the foreseeable future. This move signifies OpenAI's ambition to solidify its position as a leader in the generative AI landscape amidst intensifying competition.
#openai#ai chip#hardware#generative ai#inference#broadcom
Read original source