惯性聚合 高效追踪和阅读你感兴趣的博客、新闻、科技资讯
阅读原文 在惯性聚合中打开

推荐订阅源

奇客Solidot–传递最新科技情报
奇客Solidot–传递最新科技情报
Apple Machine Learning Research
Apple Machine Learning Research
aimingoo的专栏
aimingoo的专栏
H
Help Net Security
腾讯CDC
T
Tailwind CSS Blog
Hugging Face - Blog
Hugging Face - Blog
人人都是产品经理
人人都是产品经理
酷 壳 – CoolShell
酷 壳 – CoolShell
MongoDB | Blog
MongoDB | Blog
宝玉的分享
宝玉的分享
有赞技术团队
有赞技术团队
美团技术团队
雷峰网
雷峰网
让小产品的独立变现更简单 - ezindie.com
让小产品的独立变现更简单 - ezindie.com
博客园 - 司徒正美
博客园_首页
Recent Announcements
Recent Announcements
云风的 BLOG
云风的 BLOG
B
Blog RSS Feed
钛媒体:引领未来商业与生活新知
钛媒体:引领未来商业与生活新知
D
Docker
博客园 - Franky
Jina AI
Jina AI

Interesting Engineering

US firm to scale laser-based nuclear fusion ‘breakthrough’ with new partnership Military Archives - Interesting Engineering World’s first non-nuclear lead-cooled reactor to generate electricity begins installation US scientists devise new process to turn sewage sludge into 99% pure natural gas US firm unveils submarine-hunting drone with 9,200-mile-range, 35 mph top speed Military Archives - Interesting Engineering Supercomputer finds lithium-titanium tweak to boost sodium-ion batteries for grids Lockheed Martin demonstrates vertical launch missile system for mobile drone defense China’s 1116 MWe Taipingling Unit 1 reactor goes online, set to generate 9bn kWh yearly ChatGPT Images 2.0 update combines reasoning, research, and design with 2K output US Navy tests plug-and-play laser system on USS Bush carrier, downs drones at sea China’s CATL reveals 621-mile EV battery, under-7-minute charging to challenge BYD US uses world’s first exascale supercomputer to model supernovae, fusion reactors AI and Robotics Archives - Interesting Engineering First-in-human study confirms safety of graphene-based brain interface Tesla’s Optimus humanoid robot greets runners, poses for photos at Boston Marathon Interlocking materials offer high strength and flexibility for robotics, infrastructure US redeploys 100,000-ton nuclear-powered aircraft carrier in Red Sea after repairs US scientists unveil concept for ‘world’s first neutrino laser’ to unlock breakthroughs New military tech can maintain communication in contested electronic warfare environments Got a dark personality? Psychologists can help you choose your career wisely Humidity boosts performance of 3D-printed nanogenerator instead of degrading it China demonstrates microwave beam that recharges drones in flight, continues power delivery Scientists run compact free-electron laser for eight hours, cracks FEL stability problem China’s PLA considers to use minelaying underwater drones to enforce Taiwan blockade: Report 1-ton sharks may struggle for survival in waters exceeding 62.6°F, study suggests US firm’s thorium nuclear fuel bundles move to manufacturing for commercial reactors Tesla hits 0% charge in remote Chilean desert as YouTuber uses hood-mounted solar Humanoid robot surpasses human world record in Beijing half-marathon, clocking 50:26 mins New method extracts maximum work from unknown quantum states using symmetry tricks
OpenAI unveils Jalapeño chip for large-scale inference wo...
Neetika Walter · 2026-06-25 · via Interesting Engineering

OpenAI’s Jalapeño chip aims to speed up inference workloads while reducing compute costs.

Jalapeño was co-developed from initial design to manufacturing tape-out in just nine months

Jalapeño was co-developed from initial design to manufacturing tape-out in just nine months.

OpenAI has unveiled its first custom AI accelerator, called Jalapeño, marking the company’s move into chip design as it looks to reduce the cost and improve the efficiency of running large language models (LLMs).

Developed in partnership with Broadcom and Celestica, the chip is designed specifically for AI inference — the process of generating responses from trained AI models. OpenAI said early testing shows Jalapeño delivers significantly better performance per watt than current state-of-the-art AI accelerators, though detailed benchmarks will be released later.

The announcement expands OpenAI’s efforts to control more of the infrastructure behind its products. In addition to building models and applications such as ChatGPT and Codex, the company is now designing the hardware that powers them. Engineering samples of Jalapeño are already running machine learning workloads in the lab, including GPT-5.3-Codex-Spark, at production target frequency and power levels, according to the company.

Custom chip, faster AI

Unlike general-purpose AI accelerators adapted for multiple workloads, Jalapeño was built specifically for LLM inference. OpenAI said the architecture was designed around the compute, memory, networking, and serving requirements of modern AI models.

The company claims the chip reduces data movement while balancing compute, memory, and networking resources to improve hardware utilization. Broadcom contributed silicon implementation and networking technologies, including its Tomahawk networking platform.

“Jalapeño is part of our long-term full-stack infrastructure strategy to make compute more abundant, resulting in AI which is faster, more reliable, more affordable for people and businesses, and can be used to solve more important problems,” said Greg Brockman, President and Co-Founder of OpenAI.

Richard Ho, who leads OpenAI’s hardware program, said the accelerator was optimized around the workloads most important for frontier AI systems.

“Based on early testing, Jalapeño will efficiently execute our most important workloads close to the hardware’s theoretical limits,” Ho said. The chip is also intended to support future LLMs across the broader AI industry, not just OpenAI’s own models.

Nine-month design sprint

According to the companies, Jalapeño was developed from initial design to manufacturing tape-out in just nine months. OpenAI described the effort as potentially the fastest ASIC development cycle achieved for a high-performance advanced semiconductor.

The development process involved extensive software-hardware co-design between OpenAI and Broadcom engineers. OpenAI also said its own AI models were used to accelerate portions of the chip design and optimization workflow.

“Our collaboration with OpenAI represents a fundamental commitment to scaling the physical infrastructure required for the next decade of AI,” said Hock Tan, President and CEO of Broadcom.

The companies plan to deploy the accelerator at gigawatt-scale data centers beginning in 2026. Jalapeño is the first product in what OpenAI describes as a multi-generation compute platform that will combine OpenAI-designed accelerators with Broadcom networking and connectivity technologies and Celestica’s system integration expertise.

OpenAI said improvements in inference efficiency could translate into faster ChatGPT responses, lower AI operating costs, and more reliable access to advanced AI services as demand continues to grow.

Recommended Articles

The Blueprint

Get the latest in engineering, tech, space & science - delivered daily to your inbox.

With over a decade-long career in journalism, Neetika Walter has worked with The Economic Times, ANI, and Hindustan Times, covering politics, business, technology, and the clean energy sector. Passionate about contemporary culture, books, poetry, and storytelling, she brings depth and insight to her writing. When she isn’t chasing stories, she’s likely lost in a book or enjoying the company of her dogs.