Models
June 24, 2026

OpenAI and Broadcom unveil Jalapeño AI inference chip

OpenAI and Broadcom have introduced Jalapeño, a custom AI inference chip designed to improve performance, lower costs, and reduce reliance on third-party hardware for large-scale AI deployments.

OpenAI and Broadcom have announced Jalapeño, OpenAI's first custom AI inference chip built specifically for running large language models efficiently at scale. Designed for inference rather than model training, the chip will initially power workloads such as Codex and other customer-facing AI services.

OpenAI says Jalapeño is the first generation of a broader custom silicon roadmap aimed at improving performance, reducing operational costs, and decreasing dependence on NVIDIA hardware. Broadcom contributed its chip design expertise, while OpenAI provided insights from its AI research and infrastructure needs.

Deployment is expected to begin later this year.

#
OpenAI

Read Our Content

See All Blogs
LLM Models

Meta Muse Glimmer: Local agentic AI for enterprises

Vishesh Jain

August 13, 2026
Read more
LLM Models

GPT-5.6 benchmarks: The full testing breakdown

Deveshi Dabbawala

August 10, 2026
Read more