惯性聚合 高效追踪和阅读你感兴趣的博客、新闻、科技资讯
阅读原文 在惯性聚合中打开

推荐订阅源

V
Visual Studio Blog
I
InfoQ
freeCodeCamp Programming Tutorials: Python, JavaScript, Git & More
博客园 - 【当耐特】
小众软件
小众软件
B
Blog RSS Feed
大猫的无限游戏
大猫的无限游戏
博客园 - 三生石上(FineUI控件)
Engineering at Meta
Engineering at Meta
人人都是产品经理
人人都是产品经理
Microsoft Security Blog
Microsoft Security Blog
Last Week in AI
Last Week in AI
H
Help Net Security
爱范儿
爱范儿
云风的 BLOG
云风的 BLOG
博客园 - 司徒正美
Y
Y Combinator Blog
H
Hackread – Cybersecurity News, Data Breaches, AI and More
Microsoft Azure Blog
Microsoft Azure Blog
L
LangChain Blog
WordPress大学
WordPress大学
GbyAI
GbyAI
Google DeepMind News
Google DeepMind News
腾讯CDC

Latest from TechRadar in News

VodafoneThree gets Ofcom approval to bring satellite connectivity to your smartphone NYT Connections today – my hints and answers for April 16 (#1040) Quordle hints and answers for Thursday, April 16 (game #1543) NYT Strands hints and answers for Thursday, April 16 (game #774) Is this the tipping point for AI at work? New Gallup survey finds half of all US employees now use it in some way Allbirds — the shoe viral company — just pivoted into AI, and I wish this were an Onion headline 'Every Apple user needs to know about this nasty scam': Fake warnings tell users their iCloud data will be… 'Makes it even more disappointing': Microsoft backs fossil fuel big time with $7 billion deal in race for AI… 'Maybe it’s not science fiction': Solar panels are causing rainwater to fall in one of the driest places… Maine becomes first US state to pass data centre construction ban Dozens of WordPress plugins hijacked to target thousands of sites Drone-killing laser weapons greenlit for use in US airspace – FAA and Defense Department say high-energy weapons are ‘ready to protect all air travelers from illicit drone use’ despite airspace restrictions and friendly-fire incidents 'We are currently being extorted' — crypto giant Kraken says it is facing extortion attack, here's… McGraw Hill becomes latest to see its Salesforce data hacked Looking for a new PC? Now might be great time to upgrade, as Gartner figures claim shipments are rising — while… Farewell Surface Hub — Microsoft kills off its super-sized touchscreen displays, but you might still be able to get one if you act fast 'We have no interest in patient data in the UK': Palantir UK head defends record as criticisms rise Amazon’s new AI Bio Discovery tool can provide ‘every researcher’ with ‘lab-in-the-loop drug discovery’ – 40+ AI biology models can filter 300,000 novel antibody candidates down to the top results for testing in just weeks Over 100 Chrome Web Store extensions found stealing user data from thousands of accounts OpenAI reveals its Mythos rival designed for cybersecurity pros NYT Connections hints and answers for Tuesday, April 14 (game #1038) Forget Dr Doolittle, study finds animals might not only want to use tech, but they also want to talk to us with it… 'The decision is deeply troubling': Tesla gets a green light for Full Self-Driving in Europe — but not… OpenAI flags third-party data issue — all macOS users should update now Microsoft says Copilot is for ‘entertainment' not work, Meta’s Muse Spark and 7 other AI stories you… Man Utd vs Leeds Live Streams: How to watch Premier League 2025/26 from anywhere in the world, team news What is the release date for Invincible season 4 episode 7 on Prime Video? Linux rules on using AI-generated code - Copilot is OK, but humans must take 'full responsibility for the… The Lenovo Legion Go 2 handheld costs more than two Nvidia RTX 5080 GPUs — and that's genuinely absurd Secretlab is launching its first Diablo desk, with a design that 'traces the infernal history' of the series
Tiny company steals AMD's thunder and challenges Nvidia w...
Efosa Udinmw · 2026-05-11 · via Latest from TechRadar in News

  • Skymizer claims giant AI models no longer need hyperscale GPU infrastructure
  • Old 28nm chips suddenly power massive language models at surprisingly low wattage
  • The HTX301 squeezes 384 GB of memory into a single PCIe accelerator card

A Taiwanese company called Skymizer has unveiled a PCIe AI accelerator that challenges both AMD and Nvidia using surprisingly old technology.

The HTX301 card can run language models with up to 700 billion parameters on a single device while consuming only 240 watts of power.

The card achieves this feat using older 28-nanometer chips and standard LPDDR4 and LPDDR5 memory instead of expensive HBM or GDDR solutions.

Old tech chip competes with modern AI accelerators

Skymizer claims its card delivers 30 tokens per second with just 0.5 TOPS at 100 GB per second bandwidth.

The HTX301 is built on Skymizer's HyperThought platform, which features next-generation LPU IP designed specifically for large language model workloads.

Each PCIe card contains six HTX301 chips working together, and the card offers up to 384 GB of total memory capacity.

The design uses efficient compression techniques for both weights and KV cache, outperforming open source llama.cpp by 9 to 17.8 percent.

Sign up to the TechRadar Pro newsletter to get all the top news, opinion, features and guidance your business needs to succeed!

Its power consumption sits at less than half of what leading PCIe AI accelerators from AMD and NVIDIA typically require.

The card supports agentic AI for coding, automation, and domain-specific workflows without needing hyperscale GPU clusters.

Running large language models in the cloud introduces privacy concerns and unpredictable costs that many organizations find unacceptable.

Upgrading on-premises infrastructure to support massive GPU accelerator platforms often requires expensive redesigns of data center power and cooling systems.

Skymizer's HTX301 offers enterprises a third option that fits into standard air-cooled servers without any infrastructure changes.

The company claims the era of needing hyperscale GPU clusters for ultra-large LLMs is over with their new technology.

The PCIe card form factor allows businesses to scale AI inference on premises while maintaining data sovereignty and predictable infrastructure costs.

Skymizer HTX301 awaits real-world testing

Skymizer will preview the HTX301 at Computex this year, allowing independent verification of its performance numbers.

The specifications of this chip look impressive on paper, but real-world testing will determine whether the card actually delivers 240 tokens per second on Llama2 7B workloads.

AMD recently launched its Instinct MI350P PCIe card with 144 GB of HBM3E memory and up to 4,600 peak TFLOPS at MXFP4 precision, yet it consumes considerably more power than Skymizer's offering.

Nvidia's RTX PRO 6000 Blackwell consumes roughly 600 watts, more than double what Skymizer's card requires for comparable inference tasks.

Should the HTX301 work as advertised, it could dramatically lower the barrier to entry for on-premises AI infrastructure.

Failure to deliver would place Skymizer among the many startups that could not back up their promises.

Via Wccftech


Google logo on a black background next to text reading 'Click to follow TechRadar'

Follow TechRadar on Google News and add us as a preferred source to get our expert news, reviews, and opinion in your feeds.