惯性聚合 高效追踪和阅读你感兴趣的博客、新闻、科技资讯
阅读原文 在惯性聚合中打开

推荐订阅源

Google DeepMind News
Google DeepMind News
人人都是产品经理
人人都是产品经理
S
Securelist
P
Proofpoint News Feed
H
Help Net Security
S
Schneier on Security
T
Tenable Blog
C
Cisco Blogs
S
Security @ Cisco Blogs
博客园 - 司徒正美
博客园 - 叶小钗
Cisco Talos Blog
Cisco Talos Blog
Google DeepMind News
Google DeepMind News
C
Cybersecurity and Infrastructure Security Agency CISA
Google Online Security Blog
Google Online Security Blog
Exploit-DB.com RSS Feed
Exploit-DB.com RSS Feed
Hacker News: Ask HN
Hacker News: Ask HN
NISL@THU
NISL@THU
云风的 BLOG
云风的 BLOG
V
Vulnerabilities – Threatpost
T
The Blog of Author Tim Ferriss
aimingoo的专栏
aimingoo的专栏
W
WeLiveSecurity
www.infosecurity-magazine.com
www.infosecurity-magazine.com
Jina AI
Jina AI
腾讯CDC
WordPress大学
WordPress大学
Simon Willison's Weblog
Simon Willison's Weblog
Vercel News
Vercel News
小众软件
小众软件
N
Netflix TechBlog - Medium
有赞技术团队
有赞技术团队
AWS News Blog
AWS News Blog
雷峰网
雷峰网
Forbes - Security
Forbes - Security
The Hacker News
The Hacker News
博客园 - 聂微东
F
Full Disclosure
量子位
Scott Helme
Scott Helme
宝玉的分享
宝玉的分享
A
About on SuperTechFans
钛媒体:引领未来商业与生活新知
钛媒体:引领未来商业与生活新知
Schneier on Security
Schneier on Security
让小产品的独立变现更简单 - ezindie.com
让小产品的独立变现更简单 - ezindie.com
K
Kaspersky official blog
AI
AI
SecWiki News
SecWiki News
Webroot Blog
Webroot Blog
Martin Fowler
Martin Fowler

IEEE Spectrum

New Soft Exoskeleton Outperforms Most Hip Assist Devices We’re Squandering LEDs’ Potential to Save Our Night Skies Balcony Solar Is Sneaking Onto Grids Before the Rules Are Ready Cortisol Could Be the Next Frontier for Wearables Million P Bit Machine Pushes Probabilistic Computing to New Scale Would You Let This Humanoid Robot Do Your Laparoscopic Surgery? How a Spinning Drone Exploits Your Eyes to Become Nearly Invisible Why Indonesia’s Fisheries Future Hinges On Data Integrity and Trust Inside the Race to Tame AI’s Wild Power Swings Stable Jobs Can Hide the Riskiest Move In Your Tech Career Inside ELIZA’s Source Code and Its Multiple Personalities Tiny Puerto Rican Island Tests Hydrogen to Slash Sky High Power Bills AI Turns DNA Into Tiny Dogs and Mona Lisa Nanostructures How Darth Vader Taught Me Card Counting and AI Security Got Weird The Memory in Your Thumb Drive Could Fix AI's Big Problem The AI Arms Race in Technical Interviews Is Escalating Inside Nokia’s Race to Catch the iPhone and Android Wave Quantum Sensor Sniffs Out Radio Signals in 3D Two New Wheelchairs Reveal What “Smart” Really Means Today Video Friday: A World Cup for Robots Japan Pulls Off One of the Closest Asteroid Flybys Ever How Cheap Ground Robots Are Rewriting Frontline Warfare in Ukraine Nvidia’s NVLink Fusion Quietly Pushes Optics Inside the Rack Large Tabular Models Excel Where LLMs Fail Are Battery PoweredTrailers the Shortcut to Cleaner Long Haul Freight? The Hidden Overthinking Flaw That Could Drag AI Services Down Stacking Chips Sideways Gives AI More Memory There Independent Labs Crack Google Brain Inspired Camera Sensor Learns to See and Gently Forget Why Small AI Models Could Power Health Care Where Big Tech Cannot China’s Humanoid Army Pushes Japan to Rethink Its Robot Future NASA AI’s Wild Power Demands Are Quietly Rewriting Grid Rules Old EV Batteries Find a Second Life Backing Up the Grid UCLA’s Semiconductor Hub Is Rewiring Industry and Academia for AI Why Engineers Who Speak Up Build Stronger and Safer Careers The Orbital Data Center Hype Machine Is Already in Orbit What Emily Bender Really Meant by "Stochastic Parrots" The History and Mystery of Fireworks Poetry for Engineers: Nine Lives of Nikola Tesla Trump’s Quantum Orders Push Fault Tolerant Qubits Toward 2028 Underwater Tidal Kites Promise Steady Power for Remote Coasts How a Forgotten Wire Turned a Cheap Chip Into a Brainlike Neuron How the U.S. Engineered Its Sovereignty AI Model ConlangCrafter Dreams up Entire New Languages Weirdly Fascinating: Robotic Arm Crawls Using Its Three Fingers. Shadow-Free Augmented Reality Makes Illusions More Realistic How a Power Bank Can Turn Your AC Into a Grid Superhero Records Fall for 3D Chip Tech What it Means to Be a Mathematician When AI Does the Math Is This Stacked CFET Architecture The Ultimate CMOS Platform? Why 6 GHz Spectrum Could Make or Break Future Wi-Fi and 6G Plans Make an Origami Circuit Board AI Learns the "Dark Art" of RF Chip Design U.S. Regulator Aims to Cut Data Center Queues and Electricity Bills Home Broadband Is the Killer App 5G Was Never Designed For How Smarter Grids Could Save Americans $100 Billion On Power Can AI Learn to Read the Room? How Did Two Prompts Turn Into Potent Vibe Hacking Malware Is Europe Finally Ready to Take Back Control Of Its Tech Stack? New Device Can Take Photographs with a Single Atom War Taught this Ukrainian Entrepreneur the Value of Resilience Do Robots Need Legs? What If You Gave ChatGPT a Body? What Amazon’s Astro Taught Me About Giving Robots a Soul Optical Metasurface Sees a Sunny Future Can Sound-Driven Synapses Make AI Both Faster and Greener? Modos Color E‑Paper Monitor Pushes Open‑Source Displays Further Beat Biased Hiring By Owning Your Story In Every Interview Room How AI Attribution Could Finally Pay Musicians for Training Data How Liquid Cooling Let a Humanoid Robot Shatter Half Marathon Records Inside GM’s AI Push to Speed Up the Design of Cars and Moon Rovers Smart EV Charger Learns Your Battery’s Age to Let It Live Longer Phoenix Links IoT Chips to Save High‑Value Legacy Systems Phoenix Links IoT Chips to Save High‑Value Legacy Systems Tensordyne's Wild Log Math Aims to Leave Nvidia’s AI Chips In the Dust The Tiny Turbine That Kick-Started the U.S. Wind Industry Satellites Are Tracing Railroad Tracks Across SPHEREx’s Cosmic Map Are Emotion Reading Robots Still Missing What Matters Most? Watch This Humanoid Robot Move in Ways Your Hips Wouldn't Like The Real Cost Of Cooling GPUs In Space Might Shock You The Google DeepMind Spinoff Chasing Hidden Drug Targets We Are Crowd-Sourcing the Panopticon Gene Therapy and Sound Waves Team up to Steady Failing Hearts Save 14 Percent of Energy Used in LLM Training With This Trick The Real Tradeoffs Between Startups, Mid-Size Firms, and Giants When Does Job Hopping Stop Helping Your Engineering Future Why a Computer Science Degree Still Opens Hidden Doors AI Can Help Track the World’s Shrinking Glaciers Curiosity’s 13 Years of Software Hacks Keeps It Alive on Mars Fractal OS Lets Security Researchers See What Their CPUs Really Do Formula E DNA Helps the Cayenne Electric Bend Physics to Beat the Heat Moon’s Dark Craters Could Become the Most Precise Clocks in Space New Radio Giant in New Mexico Takes Its First Glimpse of the Cosmos Nvidia’s AI Hardware Comes to Windows in RTX Spark PCs Can Humanoid Robots Run Stairs Without Tripping? Do They Need Shoes? Inside the Compact Fusion Reactor Aiming to Power 280,000 Homes NSF X Labs Power Agile, High-Stakes Experiments "Hemopurifier" Could Help Fight Bundibugyo Ebola Strain Why Quantum Computers Need a ‘Healthy Chunk’ Of Classical Power
Why Some Coders Now Reach for GLM 5.2 Before Frontier AI Models
https://www.facebook.com/48576411181 · 2026-07-21 · via IEEE Spectrum

Zain Hasan, an AI engineer at Together AI, has taught himself to use AI coding assistants while still keeping an eye on cost. He directs difficult problems to a frontier model, meaning one near the current state of the art in reasoning and capability, such as Anthropic’s Fable. But if the task that Hasan is outsourcing is more straightforward, he directs it to a less capable—and less expensive—language model.

Right now, the cheaper model, for him, tends to be GLM 5.2. Released on 16 June by the Beijing-based lab Z.ai, GLM 5.2 is an open-weights model, meaning any organization with sufficient hardware can download and host the model for free.

Those that pay Z.ai for GLM access still can save money, because the company’s API costs $4.40 per million output tokens. That’s less than a fifth of the comparable price for access to Anthropic’s Opus 4.8 model, and a tenth the price of Anthropic’s Fable coding model. An output token is the basic unit of text a model generates in response to a prompt.

Yet many software engineers around the world, Hasan said, aren’t yet fully mindful of the net AI pricetag for a given coding project.

“A lot of companies right now, they’re still trying to figure this technology out, and so there isn’t really a token budget,” said Hasan. And when someone else is paying, the rational move for many software engineers is to skip tabulating costs entirely. “The easiest thing is to pick the most powerful model.”

That price-be-damned habit, reinforced by loose token budgets in software companies today, may now be the widest moat protecting the U.S. frontier AI labs.

Z.ai narrows benchmark gap with U.S. rivals

Z.ai’s GLM 5.2 is an AI large language model (LLM) with 753 billion parameters, though it has only 40 billion parameters active at once—an optimization that improves the speed at which a model can respond. Z.ai released the model under an MIT open-source license, which means anyone can distribute, copy, modify, and use it.

GLM 5.2’s release added to fears that U.S. AI companies could lose their competitive edge. The model nearly ties Opus 4.8’s score on some agentic coding benchmarks, such as FrontierSWE and PostTrainBench. Cybersecurity researchers have also found that GLM 5.2 scores well in cybersecurity benchmarks, a capability that spurred comparisons to Anthropic’s Mythos.

Z.ai arrives amid a broader trend. According to Stanford’s AI Index (an annual, 300-plus-page survey of AI trends) Chinese companies produced just over half as many “notable” AI models in 2025 as their U.S. counterparts. That’s up from roughly a third in 2023, and a fifth in 2020.

GLM 5.2 caused hand-wringing among some U.S. observers due to its outstanding benchmark scores, which set new records for both open-weights models and Chinese-developed models generally. The model’s Chinese origin also complicates its use for U.S. and other companies wary of routing sensitive data through Chinese-linked infrastructure.

However, the model’s open weights provide an out. Any organization worried about where their data is being sent can instead host the model on their own hardware. This stands in contrast to most frontier-level models, which are gated behind an API with no self-hosting option.

Z.ai backed up GLM 5.2’s release with the company’s own numbers, publishing a research report the same day as GLM 5.2’s launch.

The report doesn’t mention Anthropic’s Mythos or Fable, which were announced but not yet publicly available at the time of its release.The report instead focuses on Anthropic’s Opus 4.8 and OpenAI’s GPT-5.5. And while GLM 5.2 often performs almost as well as Opus 4.8 in benchmarks, the report only claims a win in two less-difficult reasoning benchmarks—and none in coding.

Many of the report’s benchmarks place GLM 5.2 behind Opus 4.8 (and, at times, OpenAI’s GPT-5.5) in agentic coding. For example, GLM 5.2 completed just 13 percent of tasks in SWE-Marathon, a difficult long-duration agentic coding benchmark. Claude Opus 4.8 doubled GLM 5.2’s score in this benchmark. Opus 4.8 also notched wins of 10 percent or more in the coding benchmarks NL2Repo, DeepSWE, and Tool-Decathlon.

How do coders use GLM 5.2?

Software engineers who’ve pitted GLM 5.2 against their own workflows report a wide range of results.

“The main thing that I realized with [GLM 5.2], was that it can do more long-horizon tasks,” said Hasan, whose company hosts GLM 5.2 on North American infrastructure. Earlier open-weights models, he said, often lost the thread after around five to fifteen back-and-forth exchanges. “This one, I noticed that I could be using it for hours, and it would still have a coherent train of thought.”

David Nix, a principal software engineer at the Denver-based MetaRouter, puts LLMs to work at both his day job and for personal side projects. (Nix also operates a jobs board of AI engineers.) Nix said GLM 5.2 comes “really close” to frontier models like Anthropic’s Opus and OpenAI’s GPT-5.5—close enough to earn a permanent spot in his rotation.

“It’s pretty great at front-end development, for example, where I don’t need to always go to Opus or Fable for those things,” said Nix. He estimates that GLM 5.2 handles 10 to 20 percent of the work he sends to an LLM on a given day, and it’s now his first stop for some specific tasks, such as front-end design.

Others reported the same strength. Hasan said GLM 5.2 has “really good taste” in web design. Kacper Michalik, a software engineer at Kraków, Poland-based Screen Studio, received good results while using GLM 5.2 to create forms for use on a website.

On the other hand, Sai Kiran Myadaram, a software engineer at Bengaluru-based Indhic AI, reports less positive results with Z.ai. He signed up for Z.ai’s subscription plan the week GLM 5.2 launched and found the model burned through its token allotment quickly. “The weekly quota that Z.ai provides has been exhausted for me in less than two to three days,” he said. Michalik, who also accessed Z.ai directly, had no significant issues with the model’s quality but occasionally bumped into rate limits, though in his case he stuck to the free plan.

In addition to rate limits, Myadaram experienced problems with model hallucinations and over-planning when asked to tackle minor front-end fixes. “It’s messing up my code base,” said Myadaram. He’s since drifted back to OpenAI’s Codex.