惯性聚合 高效追踪和阅读你感兴趣的博客、新闻、科技资讯
阅读原文 在惯性聚合中打开

推荐订阅源

Cisco Talos Blog
Cisco Talos Blog
Latest news
Latest news
G
GRAHAM CLULEY
Security Latest
Security Latest
C
Cyber Attacks, Cyber Crime and Cyber Security
Security Archives - TechRepublic
Security Archives - TechRepublic
N
News and Events Feed by Topic
S
Secure Thoughts
Simon Willison's Weblog
Simon Willison's Weblog
Exploit-DB.com RSS Feed
Exploit-DB.com RSS Feed
Help Net Security
Help Net Security
S
Schneier on Security
Application and Cybersecurity Blog
Application and Cybersecurity Blog
博客园 - Franky
H
Heimdal Security Blog
K
KPMG report finds enterprise disconnect between AI and its ROI | CIO
The Cloudflare Blog
Recent Commits to openclaw:main
Recent Commits to openclaw:main
人人都是产品经理
人人都是产品经理
Threat Intelligence Blog | Flashpoint
Threat Intelligence Blog | Flashpoint
有赞技术团队
有赞技术团队
博客园 - 【当耐特】
雷峰网
雷峰网
Forbes - Security
Forbes - Security
IT之家
IT之家
The Hacker News
The Hacker News
P
Privacy & Cybersecurity Law Blog
www.infosecurity-magazine.com
www.infosecurity-magazine.com
N
News | PayPal Newsroom
AI
AI
博客园 - 三生石上(FineUI控件)
PCI Perspectives
PCI Perspectives
H
Hacker News: Front Page
Apple Machine Learning Research
Apple Machine Learning Research
N
News and Events Feed by Topic
WordPress大学
WordPress大学
P
Palo Alto Networks Blog
S
SegmentFault 最新的问题
Last Week in AI
Last Week in AI
C
CERT Recently Published Vulnerability Notes
Cyberwarzone
Cyberwarzone
阮一峰的网络日志
阮一峰的网络日志
K
Kaspersky official blog
L
LINUX DO - 最新话题
宝玉的分享
宝玉的分享
S
Security @ Cisco Blogs
Know Your Adversary
Know Your Adversary
小众软件
小众软件
OSCHINA 社区最新新闻
OSCHINA 社区最新新闻
博客园_首页

IEEE Spectrum

Why Some Coders Now Reach for GLM 5.2 Before Frontier AI Models New Soft Exoskeleton Outperforms Most Hip Assist Devices We’re Squandering LEDs’ Potential to Save Our Night Skies Balcony Solar Is Sneaking Onto Grids Before the Rules Are Ready Cortisol Could Be the Next Frontier for Wearables Million P Bit Machine Pushes Probabilistic Computing to New Scale Would You Let This Humanoid Robot Do Your Laparoscopic Surgery? How a Spinning Drone Exploits Your Eyes to Become Nearly Invisible Why Indonesia’s Fisheries Future Hinges On Data Integrity and Trust Inside the Race to Tame AI’s Wild Power Swings Stable Jobs Can Hide the Riskiest Move In Your Tech Career Inside ELIZA’s Source Code and Its Multiple Personalities Tiny Puerto Rican Island Tests Hydrogen to Slash Sky High Power Bills AI Turns DNA Into Tiny Dogs and Mona Lisa Nanostructures How Darth Vader Taught Me Card Counting and AI Security Got Weird The AI Arms Race in Technical Interviews Is Escalating Inside Nokia’s Race to Catch the iPhone and Android Wave Quantum Sensor Sniffs Out Radio Signals in 3D Two New Wheelchairs Reveal What “Smart” Really Means Today Video Friday: A World Cup for Robots Japan Pulls Off One of the Closest Asteroid Flybys Ever How Cheap Ground Robots Are Rewriting Frontline Warfare in Ukraine Nvidia’s NVLink Fusion Quietly Pushes Optics Inside the Rack Large Tabular Models Excel Where LLMs Fail Are Battery PoweredTrailers the Shortcut to Cleaner Long Haul Freight? The Hidden Overthinking Flaw That Could Drag AI Services Down Stacking Chips Sideways Gives AI More Memory There Independent Labs Crack Google Brain Inspired Camera Sensor Learns to See and Gently Forget Why Small AI Models Could Power Health Care Where Big Tech Cannot China’s Humanoid Army Pushes Japan to Rethink Its Robot Future NASA AI’s Wild Power Demands Are Quietly Rewriting Grid Rules Old EV Batteries Find a Second Life Backing Up the Grid UCLA’s Semiconductor Hub Is Rewiring Industry and Academia for AI Why Engineers Who Speak Up Build Stronger and Safer Careers The Orbital Data Center Hype Machine Is Already in Orbit What Emily Bender Really Meant by "Stochastic Parrots" The History and Mystery of Fireworks Poetry for Engineers: Nine Lives of Nikola Tesla Trump’s Quantum Orders Push Fault Tolerant Qubits Toward 2028 Underwater Tidal Kites Promise Steady Power for Remote Coasts How a Forgotten Wire Turned a Cheap Chip Into a Brainlike Neuron How the U.S. Engineered Its Sovereignty AI Model ConlangCrafter Dreams up Entire New Languages Weirdly Fascinating: Robotic Arm Crawls Using Its Three Fingers. Shadow-Free Augmented Reality Makes Illusions More Realistic How a Power Bank Can Turn Your AC Into a Grid Superhero Records Fall for 3D Chip Tech What it Means to Be a Mathematician When AI Does the Math Is This Stacked CFET Architecture The Ultimate CMOS Platform? Why 6 GHz Spectrum Could Make or Break Future Wi-Fi and 6G Plans Make an Origami Circuit Board AI Learns the "Dark Art" of RF Chip Design U.S. Regulator Aims to Cut Data Center Queues and Electricity Bills Home Broadband Is the Killer App 5G Was Never Designed For How Smarter Grids Could Save Americans $100 Billion On Power Can AI Learn to Read the Room? How Did Two Prompts Turn Into Potent Vibe Hacking Malware Is Europe Finally Ready to Take Back Control Of Its Tech Stack? New Device Can Take Photographs with a Single Atom War Taught this Ukrainian Entrepreneur the Value of Resilience Do Robots Need Legs? What If You Gave ChatGPT a Body? What Amazon’s Astro Taught Me About Giving Robots a Soul Optical Metasurface Sees a Sunny Future Can Sound-Driven Synapses Make AI Both Faster and Greener? Modos Color E‑Paper Monitor Pushes Open‑Source Displays Further Beat Biased Hiring By Owning Your Story In Every Interview Room How AI Attribution Could Finally Pay Musicians for Training Data How Liquid Cooling Let a Humanoid Robot Shatter Half Marathon Records Inside GM’s AI Push to Speed Up the Design of Cars and Moon Rovers Smart EV Charger Learns Your Battery’s Age to Let It Live Longer Phoenix Links IoT Chips to Save High‑Value Legacy Systems Phoenix Links IoT Chips to Save High‑Value Legacy Systems Tensordyne's Wild Log Math Aims to Leave Nvidia’s AI Chips In the Dust The Tiny Turbine That Kick-Started the U.S. Wind Industry Satellites Are Tracing Railroad Tracks Across SPHEREx’s Cosmic Map Are Emotion Reading Robots Still Missing What Matters Most? Watch This Humanoid Robot Move in Ways Your Hips Wouldn't Like The Real Cost Of Cooling GPUs In Space Might Shock You The Google DeepMind Spinoff Chasing Hidden Drug Targets We Are Crowd-Sourcing the Panopticon Gene Therapy and Sound Waves Team up to Steady Failing Hearts Save 14 Percent of Energy Used in LLM Training With This Trick The Real Tradeoffs Between Startups, Mid-Size Firms, and Giants When Does Job Hopping Stop Helping Your Engineering Future Why a Computer Science Degree Still Opens Hidden Doors AI Can Help Track the World’s Shrinking Glaciers Curiosity’s 13 Years of Software Hacks Keeps It Alive on Mars Fractal OS Lets Security Researchers See What Their CPUs Really Do Formula E DNA Helps the Cayenne Electric Bend Physics to Beat the Heat Moon’s Dark Craters Could Become the Most Precise Clocks in Space New Radio Giant in New Mexico Takes Its First Glimpse of the Cosmos Nvidia’s AI Hardware Comes to Windows in RTX Spark PCs Can Humanoid Robots Run Stairs Without Tripping? Do They Need Shoes? Inside the Compact Fusion Reactor Aiming to Power 280,000 Homes NSF X Labs Power Agile, High-Stakes Experiments "Hemopurifier" Could Help Fight Bundibugyo Ebola Strain Why Quantum Computers Need a ‘Healthy Chunk’ Of Classical Power
The Memory in Your Thumb Drive Could Fix AI's Big Problem
https://www.facebook.com/48576411181 · 2026-07-14 · via IEEE Spectrum

Large Language Models (LLMs) demand immense amounts of memory, and the more people use them, the more memory is required. Memory makers responded by accelerating plans to build new memory fabs, with a focus on high-bandwidth memory (HBM) and DRAM, the first of which is scheduled to start production in 2027. But the demand for memory may also provide an opportunity for new ideas to find footing.

One of these is a tricked-out version of the kind of memory that lives in an SD card or a thumb drive—High Bandwidth Flash (HBF). It essentially takes the ideas that made HBM successful—stacking multiple chips to increase capacity and bandwidth—and applies them to the NAND flash memory commonly used for data storage in SD cards, thumb drives, and smartphones, among many other devices.

“People ask, ‘How in the world does this make a grain of sense? Flash is enormously slow,’” says Jim Handy, general director at semiconductor market research firm Objective Analysis. He explains that while NAND flash is generally lacking in bandwidth, HBF will help alleviate that concern. “[Flash] is atrociously slow for writes, but for reads, it can be coaxed to go pretty fast. And High Bandwidth Flash is going to be coaxed to do that.”

What is High Bandwidth Flash?

NAND flash stores data as a trapped electric charge in arrays of floating-gate transistors, organized into blocks and pages rather than individually addressable bytes. It’s non-volatile, too, which means data persists without power.

These traits make flash a good choice for long-term storage. It can store more bytes in the same area than DRAM, and it doesn’t require power-hungry capacitors that need constant refreshing to hold their charge. But the mechanisms that make flash dense and non-volatile also make it slow to write to, as pushing charge into and out of an insulated gate takes longer than charging a capacitor.

The latest flash interface standard can support memory bandwidth up to 4.8GB/s per die. That’s not bad for many situations, and NAND is widely used in high-performance long-term storage, such as solid state drives. However, DDR5 provides bandwidth up to 70.4GB/s per DIMM (excluding overclocked memory), and HBM4E can reach up to 3.6TB/s per stack—a roughly 750-fold bandwidth advantage for HBM4E over flash.

Hoshik Kim, senior vice president of memory systems research at SK Hynix, says HBF improves bandwidth with packaging techniques similar to HBM. “By applying advanced 3D packaging and vertical stacking techniques to NAND flash, HBF can deliver vastly higher bandwidth than standard NVMe storage,” he says. Much as HBM stacks DRAM, HBF stacks NAND flash dies to create a memory-dense chip.

HBF is at least a year away from shipping, but flash memory manufacturer SanDisk has published fact sheets for its anticipated first-generation product. The company expects HBF to stack up to 16 NAND flash chips for a total capacity of up to 512GB per stack. It also projects memory read bandwidth up to 1.6TB/s. SanDisk’s HBF roadmap also projects a second and third generation with expected read bandwidth of 2TB/s and 3.2TB/s, respectively.

What is the purpose of HBF?

Though HBF has the potential to deliver a lot more bandwidth than earlier versions of flash, you might’ve noticed a wrinkle. It’s still a lot slower than the HBM used in high-performance GPUs. Why, then, is HBF promising?

The answer lies in key differences between AI training (teaching an LLM to predict tokens) and AI inference (serving the finished model).

A model is trained by presenting it with input tokens, seeing what the model predicts, checking if that prediction was correct, and then changing weights based on the error with a step called back-propagation. While this process is simple in summary, it involves calculations across billions or trillions of model weights. That means training is heavy on both reading and writing data, which makes flash a poor fit.

However, AI inference is different. The model weights are frozen and effectively read-only, which means flash’s poor write bandwidth is no longer an obstacle. “In an inference environment massive, read-heavy data, such as the static multi-billion parameter model weights or the precomputed KV cache, can be securely housed in the HBF tier,” says Kim. That would free up HBM to work as a “high-speed scratchpad.”

Handy says it’s a sensible way to target flash memory for inference workloads. “If you set that up right, you can get an awful lot of good performance out of that—that’s just basic caching. It’s one technology that I’m expecting to go places.”

What’s next for HBF?

Though it has potential, HBF is still early in development, and likely several years away from broad deployment.

On February 25, 2026, SanDisk and SK Hynix held a kickoff event launching a joint effort to standardize HBF under a dedicated workstream within the Open Compute Project (OCP)—the same kind of open-industry body that governs many data center hardware specs. While work on the standard is ongoing, a timeline for publishing the standard has not been set.

It might seem odd for memory manufacturers—and for SK Hynix, specifically—to put forth HBF as a less expensive alternative to HBM. After all, HBM is a higher-margin product that is currently leading SK Hynix to record revenues.

However, Kim frames HBF as a complementary tool rather than a rival technology. “By alleviating the severe capacity bottlenecks of HBM without sacrificing data delivery speeds, HBF has the potential to reduce the number of individual accelerators required to run large-scale models,” he says. Kim expects this will improve energy efficiency and lower costs, making it possible for data centers to further scale their AI inference hardware.