惯性聚合 高效追踪和阅读你感兴趣的博客、新闻、科技资讯
阅读原文 在惯性聚合中打开

推荐订阅源

钛媒体:引领未来商业与生活新知
钛媒体:引领未来商业与生活新知
WordPress大学
WordPress大学
T
Tailwind CSS Blog
V
Visual Studio Blog
月光博客
月光博客
Hugging Face - Blog
Hugging Face - Blog
小众软件
小众软件
奇客Solidot–传递最新科技情报
奇客Solidot–传递最新科技情报
博客园 - Franky
freeCodeCamp Programming Tutorials: Python, JavaScript, Git & More
Last Week in AI
Last Week in AI
阮一峰的网络日志
阮一峰的网络日志
量子位
有赞技术团队
有赞技术团队
酷 壳 – CoolShell
酷 壳 – CoolShell
Apple Machine Learning Research
Apple Machine Learning Research
博客园_首页
Jina AI
Jina AI
雷峰网
雷峰网
博客园 - 【当耐特】
博客园 - 叶小钗
美团技术团队
宝玉的分享
宝玉的分享
IT之家
IT之家

Latest from TechRadar

Quordle hints and answers for Monday, April 13 (game #1540) NYT Strands hints and answers for Monday, April 13 (game #771) NYT Connections hints and answers for Monday, April 13 (game #1037) Morbid Metal developer explains why he ditched an origami art direction in favor of gritty sci-fi — 'It worked, but it didn't really feel like me' '71% of US households get routers from ISPs': Why new FCC rules could leave millions stuck with outdated,… 'The CPU is the system’s executive layer': Intel joins SambaNova as both face existential threat from… ‘More bang for your buck’: 7 easy ways to boost your MacBook Neo’s performance for free DJI Romo P vs Roborock Saros 10R — which robot vacuum comes out on top when it comes to dodging obstacles? I put… I spent 6 hours with Genshin Impact on the Galaxy S26 Ultra, and I can't believe how far mobile gaming has come What is the release date for The Testaments episode 4 on Hulu and Disney+? I reviewed the LG G6 for 3 weeks, and it's a fantastic OLED TV that's the new best option for brighter rooms Is your bird feeder camera doing more harm than good? 3 tips for using it safely as RSPB issues urgent disease warning Chelsea vs Man City Live Streams: How to watch Premier League 2025/26 from anywhere in the world, team news How to watch Alcaraz vs Sinner for FREE: TV Channels for Monte-Carlo Masters Final Sunderland vs Tottenham Live Streams: How to watch Premier League 2025/26 from anywhere in the world, team news Are these the best-designed workout headphones ever? I used them for a month to find out How to watch Snooker 900 John Virgo online (it's free) – stream O'Sullivan vs Higgins anywhere I've only just discovered the Walk With Frodo app on Garmin's Connect IQ store — and as as a huge LOTR nerd, it's going to make the next 1,800 miles fly by 'Just not sustainable': Why your monthly £25 broadband internet bill could soon hit £45 How to watch Paris-Roubaix 2026: Free Streams & TV Info as Tadej Pogacar chases third Monument How to watch Euphoria season 3 online – stream Zendaya & Sydney Sweeney drama from anywhere today '$15K bill destroyed a solo developer’s startup': How hackers are using leaked Google API keys to… There's a sneaky way to watch UFC 327 really cheap... NYT Connections hints and answers for Sunday, April 12 (game #1036) NYT Strands hints and answers for Sunday, April 12 (game #770) Quordle hints and answers for Sunday, April 12 (game #1539) Amazon's Ring cameras are the perfect solution to secure your home on a budget — shop today's best deals… I've tested every iPhone since the iPhone 12, and Ceramic Shield 2 is the first iPhone glass I fully trust UFC 327 live stream: how to watch Procházka vs Ulberg, start time, preview, full card We're officially getting the DJI Pocket 4 on April 16, but here's how Insta360 could beat it
Tiny company steals AMD's thunder and challenges Nvidia w...
Efosa Udinmw · 2026-05-11 · via Latest from TechRadar

  • Skymizer claims giant AI models no longer need hyperscale GPU infrastructure
  • Old 28nm chips suddenly power massive language models at surprisingly low wattage
  • The HTX301 squeezes 384 GB of memory into a single PCIe accelerator card

A Taiwanese company called Skymizer has unveiled a PCIe AI accelerator that challenges both AMD and Nvidia using surprisingly old technology.

The HTX301 card can run language models with up to 700 billion parameters on a single device while consuming only 240 watts of power.

The card achieves this feat using older 28-nanometer chips and standard LPDDR4 and LPDDR5 memory instead of expensive HBM or GDDR solutions.

Old tech chip competes with modern AI accelerators

Skymizer claims its card delivers 30 tokens per second with just 0.5 TOPS at 100 GB per second bandwidth.

The HTX301 is built on Skymizer's HyperThought platform, which features next-generation LPU IP designed specifically for large language model workloads.

Each PCIe card contains six HTX301 chips working together, and the card offers up to 384 GB of total memory capacity.

The design uses efficient compression techniques for both weights and KV cache, outperforming open source llama.cpp by 9 to 17.8 percent.

Sign up to the TechRadar Pro newsletter to get all the top news, opinion, features and guidance your business needs to succeed!

Its power consumption sits at less than half of what leading PCIe AI accelerators from AMD and NVIDIA typically require.

The card supports agentic AI for coding, automation, and domain-specific workflows without needing hyperscale GPU clusters.

Running large language models in the cloud introduces privacy concerns and unpredictable costs that many organizations find unacceptable.

Upgrading on-premises infrastructure to support massive GPU accelerator platforms often requires expensive redesigns of data center power and cooling systems.

Skymizer's HTX301 offers enterprises a third option that fits into standard air-cooled servers without any infrastructure changes.

The company claims the era of needing hyperscale GPU clusters for ultra-large LLMs is over with their new technology.

The PCIe card form factor allows businesses to scale AI inference on premises while maintaining data sovereignty and predictable infrastructure costs.

Skymizer HTX301 awaits real-world testing

Skymizer will preview the HTX301 at Computex this year, allowing independent verification of its performance numbers.

The specifications of this chip look impressive on paper, but real-world testing will determine whether the card actually delivers 240 tokens per second on Llama2 7B workloads.

AMD recently launched its Instinct MI350P PCIe card with 144 GB of HBM3E memory and up to 4,600 peak TFLOPS at MXFP4 precision, yet it consumes considerably more power than Skymizer's offering.

Nvidia's RTX PRO 6000 Blackwell consumes roughly 600 watts, more than double what Skymizer's card requires for comparable inference tasks.

Should the HTX301 work as advertised, it could dramatically lower the barrier to entry for on-premises AI infrastructure.

Failure to deliver would place Skymizer among the many startups that could not back up their promises.

Via Wccftech


Google logo on a black background next to text reading 'Click to follow TechRadar'

Follow TechRadar on Google News and add us as a preferred source to get our expert news, reviews, and opinion in your feeds.