惯性聚合 高效追踪和阅读你感兴趣的博客、新闻、科技资讯
阅读原文 在惯性聚合中打开

推荐订阅源

S
Secure Thoughts
B
Blog
MongoDB | Blog
MongoDB | Blog
GbyAI
GbyAI
博客园 - 【当耐特】
D
DataBreaches.Net
Apple Machine Learning Research
Apple Machine Learning Research
阮一峰的网络日志
阮一峰的网络日志
I
InfoQ
人人都是产品经理
人人都是产品经理
Microsoft Azure Blog
Microsoft Azure Blog
量子位
美团技术团队
Recent Announcements
Recent Announcements
freeCodeCamp Programming Tutorials: Python, JavaScript, Git & More
M
MIT News - Artificial intelligence
O
OpenAI News
SecWiki News
SecWiki News
A
About on SuperTechFans
J
Java Code Geeks
B
Blog RSS Feed
Y
Y Combinator Blog
L
LangChain Blog
Security Archives - TechRepublic
Security Archives - TechRepublic
Attack and Defense Labs
Attack and Defense Labs
小众软件
小众软件
CTFtime.org: upcoming CTF events
CTFtime.org: upcoming CTF events
Martin Fowler
Martin Fowler
博客园 - 聂微东
雷峰网
雷峰网
有赞技术团队
有赞技术团队
Google DeepMind News
Google DeepMind News
T
The Exploit Database - CXSecurity.com
N
News and Events Feed by Topic
OSCHINA 社区最新新闻
OSCHINA 社区最新新闻
Microsoft Security Blog
Microsoft Security Blog
Recorded Future
Recorded Future
P
Palo Alto Networks Blog
Blog — PlanetScale
Blog — PlanetScale
N
News | PayPal Newsroom
Scott Helme
Scott Helme
L
LINUX DO - 热门话题
F
Full Disclosure
cs.AI updates on arXiv.org
cs.AI updates on arXiv.org
The Hacker News
The Hacker News
cs.CL updates on arXiv.org
cs.CL updates on arXiv.org
U
Unit 42
博客园_首页
T
Tailwind CSS Blog
T
Tenable Blog

OfficeChai

These Are The 10 Cheapest AI Models In The World [June 2026] 18 Best AI Tools For English Speaking (With Examples) [2026] AI Impact? Vacancy Rates For US Office Properties Are Now Highest Since The 2008 Crisis KPMG Pulls Report Praising AI After It Was Found To Have Fake AI-Generated Citations India's Sarvam Raises $234 Million At $1.5 Billion Valuation After SpaceX Stock Pops 20%, Musk Has Made More Money In The Last 24 Hours Than Warren Buffett Made In His Entire Career OfficeChai Nobody Is Using AI Better Than Meta: NVIDIA CEO Jensen Huang 21 Best AI Tools For Animation (With Examples) [2026] 22 Best AI Tools For Architecture (With Examples) [2026] Datacenter Construction Spending Has Eclipsed Public Transportation Spending In The US China Scraps 12,000 Degree Courses, Mainly In Arts And Humanities, To Prepare For AI Age OfficeChai There Is No Job Loss With AI: David Friedberg Loop Between Human Capital And "Token Capital" Will Be The New IP For Firms, Says Satya Nadella How to Reduce Dependency on Key Employees 8 Google Index Checker Use Cases Beyond New Blog Posts Memory Squeeze? Smartphone Purchases Are Down Globally 21 Best AI Tools For Accounting (With Examples) [2026] AI For Voice Generation: 22 Best Options (With Examples) [2026] These Are The Most Popular Image Generation Models On OpenRouter [June 2026] Search Traffic For Websites Is Down 25% Over The Last Year Because Of AI: a16z Data Agentic Coding Has Led To A 50% Increase In Number Of Apps, But Most Are Finding Very Few Users: SimilarWeb Data OpenRouter Launches Fusion API, Which Uses A Combination Of Models To Achieve Fable-Like Performance At Half The Price Dario Amodei Refused To De-Deploy Or Fix Vulnerabilities In Fable Before US Export Controls, Says David Sacks 23 Best AI Tools For Notes Making (With Examples) [2026] 16 Best AI Tools For Astrology (With Examples) [2026] How Jensen Huang Once Had To Ask SEGA's CEO To Pay NVIDIA For A Technology That Didn't Work ChatGPT Already Has 11% Of The Search Market: OpenAI CFO Sarah Friar SpaceX Has Now Launched More Satellites Than Rest Of Humanity Combined Across History Globalization Is Dead, Time For India To Wake Up Says Sridhar Vembu After US Bans Anthropic Mythos And Fable Models For Foreign Users Elon Musk Becomes World's First Trillionaire After Record SpaceX IPO Anthropic Suspends Access To Mythos And Fable Models Following US Govt Directive Against Foreign Users 27 Best AI Tools For Market Research (With Examples) [2026] Why Jeff Bezos Makes Important Decisions Early in The Morning Education And Healthcare IT Have Been The Hardest Areas To Invest In: Peter Thiel Giving AI Long-Term Goals Could Lead To The Emergence Of Self-Preservation: Geoffrey Hinton Your Startup Doesn't Have a Hardware Problem. It Has an Accountability Problem Cyber Incidents Rarely Start With a Hacker: The Weak Links Businesses Overlook What Makes an App Worth Returning to Every Day? 21 Best AI Tools For Lead Generation (With Examples) [2026] How NBA Player Shaquille O'Neal Became An Early Investor In Ring AI For Kids Learning: 22 Best Options (With Examples) [2026] These Are The Most Popular AI Model Companies On OpenRouter [June 2026] Advanced Fintech and NeoBank Software Development Solutions: Building the Digital Banks of Tomorrow TRON Payments: Integrating AML Checks Into Business Workflows 18 Best AI Tools For Resume (With Examples) [2026] 16 Best AI Tools For UI Design (With Examples) [2026] These Are Top 10 Countries Generating The Most Internet Traffic How to Choose the Best Magento Agency for Your Store These Are The Best AI Models For Creative Writing [June 2026] AI For Managers: 28 Best Tools (With Examples) [2026] 17 AI Tools For Trading (With Examples) [2026] AI Has Led To An Explosion Of New Apps, But Nearly None Have Managed To Garner Significant Usage Cloudflare CEO Matthew Prince Says Vinod Khosla Asked Him To Fire His Co-founders For Him To Invest In His Company Australia’s AirTrunk To Invest $30 Billion To Develop Datacenters In India Anthropic Says That Their Employees Are Using AI To Write 8x More Code Compared To 18 Months Ago Anthropic Is Extremely Expensive, Many Are Urgently Looking For Alternatives: Microsoft AI CEO Mustafa Suleyman Sergey Tokarev on creating DIY “Beehives” and a free guidebook AI Crypto Price Prediction: How Accurate Are Machine Learning Models? Why Anthropic Could Find It Hard To Maintain Its $965 Billion Valuation Startup CEO Says They're Saving "Millions Of Dollars" By Replacing Anthropic Models With DeepSeek Ola Cabs' Valuation Falls 99% From Peak, Now Valued At Just $70 Million By Vanguard After TCS Case, Former Wipro Employee Alleges Attempt At Religious Conversion By Coworkers Bot Traffic Has Surpassed Human Traffic On The Internet For The First Time In History, Clouflare Says ChatGPT's Free Users Do 7 Queries Per Day, Those On $20 Plan Do 3x More: CFO Sarah Friar How Keith Rabois Had Been "Highly Skeptical" In 2023 That Anthropic Would Be Worth More Than $5 Billion In 10 Years How to Install AdGuard Home with Docker Step by Step We're Running Out Of Training Data, But Not Too Worried Because There Are Alternate Approaches: Google's Jeff Dean JioHotstar Is Hiring For 75 AI Roles Amid AI Content Push NVIDIA's Nemotron 3 Becomes Most Intelligent Open Weights Model From The US Hackers Allegedly Fooled Meta's AI To Take Over Accounts By Simply Asking It To Change User Emails Manchester Super Giants' AI Promotional Video Gets Panned As "Slop" For Glaring Cricketing Errors AI Reducing Jobs Is "Complete Nonsense": NVIDIA CEO Jensen Huang MiniMax Releases MiniMax M3, Is Competitive With Frontier Models On Many Benchmarks IIT Delhi-Incubated BotLab Dynamics Lights Up Skies With Lord Shiva Themed Drone Show During IPL Final NVIDIA Introduces RTX Spark, A New Chip Optimized For AI Agents For Windows Laptops And PCs NVIDIA Introduces Vera, A New CPU Chip For AI Agents That Is 80% Faster Than x86 CPUs OpenAI's Codex Reaches 5 Million Users, Resets Rate Limits For Users Key Factors That Influence Personal Loan Approval in India AI Is Allowing Me To Experiment And Try Crazier Things: Mathematician Terrance Tao Efficiency Of Human Learning Is Still A Thousand Times Better Than LLM Learning, Need Algorithmic Advances To Improve It: Jeff Dean San Francisco Home's Zillow Listing Says It'll Accept OpenAI Or Anthropic Stock As Payment Open-Source Models Currently Lag Proprietary Models By Just 4 Months: Epoch AI Self-Improvement Possible In AI Models Within A Year, Say Google's Top AI Leaders Digital Minds: Preparing for a Moral Challenge Before It Arrives Nearly 30% Of US-Based Y-Combinator Founders Are Of Indian Origin: SF Chronicle Data "A New Era Of PC": NVIDIA, Microsoft Windows Tease New Collaboration At Least 146,000 AI Hallucinated Citations In Papers Published In 2025, Finds Paper AI Doesn't Undergo Experiences, Has No Moral Conscience: Pope Leo XIV Claude Opus 4.8 Tops Artificial Analysis Intelligence Index, Edges Out GPT 5.5 With Score Of 61.4 Anthropic Says Its Annual Revenue Run-rate Has Now Touched $47 Billion Anthropic Raises $65 Billion At $965 Billion Valuation, Is Now Worth More Than OpenAI Claude Opus 4.8 Is Better Than Opus 4.7 But Not As Good As Mythos Preview, Says Anthropic Claude Opus 4.8 Beats GPT 5.5 On GDPval-AA Benchmark For Real World Tasks Anthropic Releases Claude Opus 4.8, Beats Opus 4.7, GPT-5.5 On Many Benchmarks GTM for Tech Startups Explained How to Use an AI Picture Generator to Create Professional Images Anthropic Is Now Generating 35% More Revenue Than OpenAI: The Information SK Hynix, Micron Join $1 Trillion Club Following AI-Led Memory Shortages
OpenAI Has Designed And Built Its First AI Chip Named Jalapeño In Partnership With Broadcom
OfficeChai Team · 2026-06-24 · via OfficeChai

OpenAI had thus far stayed in the AI models space, but it’s now taking steps into moving into the hardware direction.

The company has unveiled Jalapeño — its first custom AI chip — built in partnership with Broadcom. Designed specifically for LLM inference, the chip was conceived from scratch by OpenAI’s engineering teams and taken from design to manufacturing tape-out in just nine months, a timeline that Broadcom describes as potentially the fastest ASIC development cycle ever achieved in high-performance semiconductors. Celestica is also part of the picture, handling board, rack, and system integration to bring the platform to production.

openai chip jalapeno

Jalapeño is purpose-built around what OpenAI actually runs: ChatGPT, Codex, its API, and the growing range of agentic products the company is building toward. Rather than adapting a general-purpose accelerator to fit LLM workloads, OpenAI designed the chip around those workloads from the ground up. Engineering samples are already running GPT-5.3-Codex-Spark in the lab at production-target frequency and power. Early results suggest meaningfully better performance per watt than current state-of-the-art hardware, though OpenAI says a detailed technical report will follow in the coming months.

The architecture focuses on reducing data movement and better balancing compute, memory, and networking resources — the kind of tradeoffs that matter enormously at inference scale, where the bottleneck is often not raw compute but how efficiently data flows through the system.

The Infrastructure Play

For a company that has spent years buying Nvidia GPUs at scale, this is a meaningful shift. Inference — the process of serving a model’s response to a user query — represents a large and growing share of the compute bill for any company running at ChatGPT’s volume. A chip optimized specifically for that job, and tuned to OpenAI’s own models, gives the company a lever it previously lacked: the ability to improve inference economics without waiting for Nvidia’s next product cycle.

Broadcom CEO Hock Tan, who delivered a physical sample of the chip to Sam Altman and Greg Brockman, has long made this argument publicly — that companies serious about leading in AI need their own silicon. OpenAI’s Jalapeño is the company acting on that logic. Broadcom’s role isn’t just manufacturing; its Tomahawk networking silicon is woven into the platform, and the company is a critical part of how the chip scales to gigawatt-level deployments planned with Microsoft and other data center partners starting later in 2026.

Jalapeño is explicitly the first step in a multi-generation roadmap. The companies are targeting 10 gigawatts of compute powered by OpenAI-designed accelerators with Microsoft through 2029, which is the kind of commitment that signals this isn’t a one-off experiment.

Where the Rest of the Industry Stands

Google has been doing this for years. Its Tensor Processing Units (TPUs) — purpose-built accelerators that bypass general-purpose GPU architecture — have been in production since 2016. TPUs power much of Google’s AI infrastructure across Search, Translate, and DeepMind workloads. The advantage Google gained from designing silicon that matches its software is precisely what OpenAI is now trying to replicate.

Amazon has Trainium and Inferentia. Meta has MTIA. Microsoft has Maia. OpenAI, until now, was the notable exception among the major AI players — leaning heavily on Nvidia while the rest of the hyperscaler field was building inward.

Anthropic sits in a different position entirely. The company has no proprietary chip program and currently trains Claude across AWS Trainium, Google TPUs, and Nvidia GPUs. It has signed a deal with Google and Broadcom for next-generation TPU capacity, which means it’s still dependent on external silicon. Interestingly, Clive Chan — one of the earliest engineers on OpenAI’s chip team — recently left to join Anthropic, a sign that the company is accumulating hardware talent, but it’s a long road from hiring chip engineers to shipping your own silicon.

What This Means for Nvidia

OpenAI has relied on Nvidia almost exclusively since the beginning, and Jalapeño doesn’t change that overnight. Training workloads, in particular, still run on Nvidia hardware, and OpenAI has acknowledged it’s only exploring whether to extend its custom chip program into training. For now, Jalapeño is an inference play.

The harder question Jalapeño raises for Nvidia isn’t whether it loses OpenAI’s business outright. It’s whether the largest AI operators — Google, Amazon, Microsoft, Meta, and now OpenAI — continue accepting one default architecture for every job. Once a company has its own inference silicon tuned to its own models, the conversation about what to buy from Nvidia becomes more deliberate. That’s a different dynamic than simply waiting in line for the next GPU allocation.

Broadcom, for its part, emerges from this announcement with a stronger position. The company has quietly become the workshop of choice for hyperscalers building custom AI silicon — and landing OpenAI, arguably the highest-profile name in the space, cements that role.

OpenAI’s full-stack ambitions have been visible for a while. Jalapeño is the infrastructure layer becoming real.