惯性聚合 高效追踪和阅读你感兴趣的博客、新闻、科技资讯
阅读原文 在惯性聚合中打开

推荐订阅源

S
Secure Thoughts
B
Blog
MongoDB | Blog
MongoDB | Blog
GbyAI
GbyAI
博客园 - 【当耐特】
D
DataBreaches.Net
Apple Machine Learning Research
Apple Machine Learning Research
阮一峰的网络日志
阮一峰的网络日志
I
InfoQ
人人都是产品经理
人人都是产品经理
Microsoft Azure Blog
Microsoft Azure Blog
量子位
美团技术团队
Recent Announcements
Recent Announcements
freeCodeCamp Programming Tutorials: Python, JavaScript, Git & More
M
MIT News - Artificial intelligence
O
OpenAI News
SecWiki News
SecWiki News
A
About on SuperTechFans
J
Java Code Geeks
B
Blog RSS Feed
Y
Y Combinator Blog
L
LangChain Blog
Security Archives - TechRepublic
Security Archives - TechRepublic
Attack and Defense Labs
Attack and Defense Labs
小众软件
小众软件
CTFtime.org: upcoming CTF events
CTFtime.org: upcoming CTF events
Martin Fowler
Martin Fowler
博客园 - 聂微东
雷峰网
雷峰网
有赞技术团队
有赞技术团队
Google DeepMind News
Google DeepMind News
T
The Exploit Database - CXSecurity.com
N
News and Events Feed by Topic
OSCHINA 社区最新新闻
OSCHINA 社区最新新闻
Microsoft Security Blog
Microsoft Security Blog
Recorded Future
Recorded Future
P
Palo Alto Networks Blog
Blog — PlanetScale
Blog — PlanetScale
N
News | PayPal Newsroom
Scott Helme
Scott Helme
L
LINUX DO - 热门话题
F
Full Disclosure
cs.AI updates on arXiv.org
cs.AI updates on arXiv.org
The Hacker News
The Hacker News
cs.CL updates on arXiv.org
cs.CL updates on arXiv.org
U
Unit 42
博客园_首页
T
Tailwind CSS Blog
T
Tenable Blog

Hacker News

Introducing Claude Opus 4.7 Qwen Studio The Future of Everything is Lies, I Guess: Where Do We Go From Here? GitHub - SeanFDZ/macmind: Single-layer transformer in HyperTalk for the classic Macintosh Show HN: Agent-cache – Multi-tier LLM/tool/session caching for Valkey and Redis Bonsai 1-bit WebGPU - a Hugging Face Space by webml-community Moving a large-scale metrics pipeline from StatsD to OpenTelemetry / Prometheus GitHub - Nightmare-Eclipse/RedSun: The Red Sun vulnerability repository GitHub - SethPyle376/hiraeth: Local AWS emulator focused on fast integration testing, with SQS support, SQLite-backed state, and a debug-friendly web UI. GitHub - macOS26/Agent: Any AI, replaces Claude Code, Cursor, OpenClaw. Over 18 LLM providers (Claude, OpenAI, Gemini, Ollama, Zai, HF, Qwen) wired into a native Mac app that writes code, builds Xcode projects, bumps versions, manages git, automates Safari, use AppleScript, JS or Accessibility, extend Agent! w/ MCP Servers, run tasks from your iPhone via Messages. YouTube now lets you turn off Shorts I Made a Terminal Pager Burgers | マクドナルド公式 Commands — HackerNews CLI documentation ChatGPT for Excel PiCore - Raspberry Pi Port of Tiny Core Linux Live Nation illegally monopolized ticketing market, jury finds Google Broke Its Promise to Me. Now ICE Has My Data. Founding Engineer at Adaptional | Y Combinator CRISPR takes important step toward silencing Down syndrome’s extra chromosome GitHub - saffron-health/libretto: The AI toolkit for building reliable browser automations US v. Heppner (S.D.N.Y. 2026) no attorney-client privilege for AI chats [pdf] Retrofitting JIT Compilers into C Interpreters IPv6 – Google The Accursèd Alphabetical Clock Cybersecurity Looks Like Proof of Work Now Fragments: April 14 Cal.com Goes Closed Source: Why AI Security Is Forcing Our Decision | Cal.com - Scheduling Software for Online Bookings Laravel raised money and now injects ads directly into your agent When moving fast, talking is the first thing to break Too much Discussion of the XOR swap trick – Heather Cafe Introduction to Spherical Harmonics for Graphics Programmers The Grand Line Building a Z-Machine in the worst possible language High-Level Rust: Getting 80% of the Benefits with 20% of the Pain GitHub - duguyue100/midnight-captain: Inspired by Midnight Commander, tailored to my taste. How to build a `git diff` driver · Jamie Tanna | Software Engineer Center for Responsible, Decentralized Intelligence at Berkeley The Local Universe’s Expansion Rate Is Clearer Than Ever, but Still Doesn’t Add Up - A new synthesis of astronomical measurements confirms a persistent mismatch that could point to physics beyond current models The air throughout our homes is infused with microplastics. But there are things you can do to breathe less of them The disturbing white paper Red Hat is trying to erase from the internet – OSnews The Future of Everything is Lies, I Guess: Annoyances ‘Abhorrent’: the inside story of the Polymarket gamblers betting millions on war Productive procrastination — Max van IJsselmuiden maps, territory and LMs 447 Terabytes per Square Centimetre at Zero Retention Energy: Non-Volatile Memory at the Atomic Scale on Fluorographane Show HN: Pardonned.com – A searchable database of US Pardons 20 Years on AWS and Never Not My Job The Seasons are Wrong Artemis II crew splashes down near San Diego after historic moon mission We gave an AI a 3 year retail lease in SF and asked it to make a profit | Andon Labs How a dancer with ALS used brainwaves to perform live On filing the corners off my MacBooks Installing every* Firefox extension OpenClaw’s memory is unreliable, and you don’t know when it will break Steve Blank Nowhere Is Safe Chimpanzees in Uganda locked in vicious 'civil war', say researchers watgo - a WebAssembly Toolkit for Go linux/Documentation/process/coding-assistants.rst at master · torvalds/linux GitHub - callumlocke/json-formatter: Makes JSON easy to read. Founding Product Engineer at Bild AI | Y Combinator A compelling title that is cryptic enough to get you to take action on it GitHub - Keychron/Keychron-Keyboards-Hardware-Design: Industrial design files for Keychron keyboards and mice. 100+ models with CAD assets in STEP, DXF, DWG, and PDF. Source-available, with commercial use allowed for original compatible accessories within the license terms. [ANNOUNCE] WireGuardNT v0.11 and WireGuard for Windows v0.6 Released 1D-Chess Helium Is Hard to Replace Cooperative Vectors Introduction | Evolve Keeping a Postgres queue healthy — PlanetScale Our response to the Axios developer tool compromise Do Americans read print books, e-books or audiobooks more? The Zettelkasten Method in Obsidian: A Practical Setup Guide Artemis II Is Competency Porn and We Are Starving For It WeakC4 Flight Viz — Cockpit View A Mexican surveillance giant you’ve never heard of is now watching the U.S. border Surelock: Deadlock-Free Mutexes for Rust RISC-V 101 – what is it and what does it mean for Canonical? | Ubuntu The Problem That Built an Industry How Much Linear Memory Access Is Enough? | Solidean Investigating Split Locks on x86-64 Simplest hash functions Sybilproof reputation mechanisms (2005) [pdf] What is a property? How Complex is my Code? Static code analysis in Kotlin — tools overview Toffoli gates are all you need PGLite evangelism dcmake: a new CMake debugger UI Clojure on Fennel part one: Persistent Data Structures Fragments: April 2 Python Release Python install manager 26.1 The Life and Death of the Book Review - Liberties Introducing Database Traffic Control — PlanetScale Bitcoin miners are losing $19,000 on every BTC produced as difficulty drops 7.8% God sleeps in the minerals Building slogbox Apple Silicon and Virtual Machines: Beating the 2 VM Limit Who was “Not Even Wrong” first? Pokemon Evolution Vs Darwinian Evolution The APL Programming Language Source Code
rars in Rust, bro
davidsong · 2026-05-14 · via Hacker News

I’ve done a few different reverse-engineering projects with LLMs, and figured it’s time to push the clankers to their limits.

A RAR compressor for every version of RAR ought to have taken about 5 years, which is why nobody has ever bothered. Today, it takes 5 weeks of evenings and weekends, clanking OpenAI Codex 5.5 and Claude Opus 4.7, and cost roughly £40 in (heavily subsidised) tokens.

Yes it’s 55k lines of slop, no it’s not that fast, and it almost earned me an OpenAI ban. But it works.


SPECIF~1.RAR

RAR was originally an LZSS compressor for DOS, which peaked in popularity as the warez scene’s format of choice. Fighting with WinZip for feature parity and supremacy, WinRAR boasted multi-volume support, recovery records and even an internal VM, but its USP was always superior compression. It’s a middle-aged format that never stopped growing up, it’s as big as a house.

unrar comes with source code but that code is not actually free, and somewhat ironically RAR’s author Eugene Roshal isn’t a big fan of piracy. So ideally I’d need to implement my version from spec, which doesn’t really exist.

The monstrous task of creating one involved pulling code from free decompressor sources in the wild - unar, libarchive, UNRARLIB, plus random web pages and folk lore.

I then set Claude to work documenting as much as it could. After each pass, I quizzed it on missing features and maintained an ongoing gaps doc containing the hard-to-know stuff. This persisted between context resets, which were needed to flow the tokens into the gaps. It took 2 weeks of cooking, going back and forth until we had most of the reader side documented. The writer side, however, remained a mix of confabulation and conjecture.

So next I grabbed the RAR binaries for DOS and Windows, and set to work making test fixtures, hex-dumping and doing passes in Ghidra and DOSBox-x to get some idea of how they were packed. Another week or two of work and the gaps started to close up.

Now I had something that might be useful; spec docs for every version of the RAR file format:


Building something

Being confidently wrong enough to start, Codex, Claude and I set off building a (precariously) compatible Rust CLI. The workflow was shaped something like this:

Working from spec

Opus is great, but it tends to enthusiastically generate code while missing the bigger picture. Claude requires remedial passes, refactoring and a short leash, but is great for a chat about strategy or architecture. Gippity 5.5 stays on target when left alone, but will rabbit hole you hard if you chat with it too much. I could give Codex the spec docs and basically tell it to just get on with it. Very refreshing.

While working from spec, Codex would randomly stop due to cyber violations, and I’d need to manually compact to continue. Eventually I had to get verified by OpenAI to stop it from happening. Well, it turned out that at some time during spec investigation, Claude needed to understand authenticity verification which is a paid feature. With a context full of reverse engineering tools it cracked WinRAR and bypassed product registration, then dutifully documented its crimes in the spec. The docs, when viewed, triggered OpenAI’s alarms and stopped it dead in its tracks. I squashed this out of the git history, and decided not to implement the feature at all.

One foot on the brake

You’ve gotta keep an eye on the bots and interrupt when things start to smell bad. If you don’t then they’ll special-case their way out of every problem and around every test, ugly patterns will propagate through your code, and you’ll need an expensive refactor later. I should have done more, but I didn’t, and for that I paid the price later on. The tokens were subsidised, but it was a waste of my time.

For the last 15 months or so my hobby has been shouting at Claude, so I’m getting good at interventions. I enjoy it, even if it has damaged my personality. I tend to swear at Codex far less, maybe because it’s faster or less of a grinning idiot, but probably because it’s bland and professional. This may be a good thing, but I’m not sure yet.

Tests, for science

Way too many tests. Fragile tests, coverage that doesn’t matter, excessively_long_test_names_that_fill_your_screen, these are vital when working on something this size. They provide a statistical mass that warps text generation, pulling the bots back on track when they go off-piste or try to cut corners. So, reams of unit tests and as much coverage as is possible please. We can always remove them later, right? Right?…

So the tests keep it in shape, but actually running the code is what aligns it with reality. So the real work is about fixtures, oracles, and updating the spec where wrong. In doing this, Codex cleared up autofill bullshit (“hallucinations”) that had previously passed at least ten rounds of review.

So it turns out that empirically grinding against reality is the best source of signal, and with enough time the spec was honed into something close to Truth. Science.

Cross cutting context

Periodically, I had Claude generate a full review of the code. This helps nudge the codebase away from an intolerable slop and more towards a tolerable one, which is the best we can hope for in May 2026.

The problem with review agents though, is enthusiasm. They generate laser focused nitpickings that quibble over things that don’t matter, so you need a filter.

My filter is, I .gitignore a review.md and instruct codex to group reviews into batches. I then add them to a plan.md by functional area, and the plan drives development tasks. Claude being aware of previous reviews invites compounding blind spots, telling Codex which things I don’t care about pollutes the context - as all text does. Selective context management, switching between sessions, rm’ing the review doc, and running on different machines provides variety that helps the work flow into all the gaps.

After a while I had Claude act as a UAT tester, orchestrate compatibility test suites and test against archives in the wild too, producing fresh review.md’s that fed into the plan by the usual route. It was a serial process, but not too much of a bottleneck.

A first release

By the time I reached RAR 2.9 I was getting a bit bored of reading machine-generated spew. So I switched to some other projects for a bit. Having a working CLI for version 1.3 and 1.4 of RAR, which very few tools can even open, I regenerated this into a new dir and pushed it up to crates.io.

I figured it’d be useful for archivists, even if I gave up. So here it is:


Scoring a /goal

I picked it back up late last week when OpenAI released the /goal feature. This is essentially a Ralph loop that allows the bot to grind on at a task indefinitely, compacting and picking up after filling its meagre context limit. Running only one session means it doesn’t even hit the 5 hour usage limits, so it ran multiple times for 6+ hours while transcribing the rest of the spec into code, and once for a solid 16 hours before I interrupted it and demanded a refactor. It smashed through the bulk of the work this way, flood-filling around 40,000 lines, doing recovery records, encryption, multi-volume support and tons of spec work that I’m still barely aware of.

While Codex worked, Claude and I found more RAR files and set up a compression benchmarking and a compatibility regression test suite. Giving Codex the task of optimizing compression worked surprisingly well, it was able to apply well-known techniques from other compressors to optimize LZSS to around 5-10% worse than WinRAR, and beat RAR on some of our test data. WinRAR was optimized for decades by a skilled and obsessive Russian hacker, I used the median of the distribution with brute force and ignorance. Coming so close feels like a huge win given the effort involved. And since I don’t want to read the code too much, it’ll have to do.

Performance was a different story. Codex is happy using valgrind and hyperfine to find hot spots, and it gobbled up the low hanging fruit without problems. It fell short on finding the sort of novel performance hacks that a seasoned C dev would use to squeeze the hot loops, and ended up multiple times slower. Or it could just be that idiomatic, safe Rust code is slow. I suspect it’s both.

The latest codex model generates code with barely any comments, which I’m a fan of because comments aren’t executed and quickly rot. But comment-aversion plus compaction equalled a stream of functional regressions in RAR 1.4 compatibility, mostly with the DOS version of RAR, ones that tripped it up on at least 3 occasions. Putting a comment in the source would have prevented this.

Another thing, Claude’s reviews, specially the UAT reviews, got deep into the details but missed the most obvious thing - that the UX was a terrible blob of machine readable noise. It caught inconsistencies and errors, but until I specifically said “tell me why this is shit UX”, it didn’t recommend anything to do with that despite UX being the main goal. This applies to other areas too, they have blind spots by default that can probably be solved with agent skills or a bit of “uwotm8” in the prompts.


What did we learn?

So the TL;DR is:

  1. Working from spec actually works.
  2. Modern models are very good at Rust.
  3. Autonomous research is extremely powerful.
  4. Tests, docs and comments shape the context through mass and steering.
  5. Bring your own architecture, or pay a refactoring fee.
  6. Don’t expect decent performance or novel insights just yet.
  7. Bots can overlook the obvious.

🦀 rars

So here it is, a Rust implementation of RAR. It’s sloppy, it’s slow, it’s almost two megabytes in size and somewhat worse than WinRAR on compression.

But, it works, and the world now has a free software RAR implementation. So that was worth the effort.

You can install rars like so:

cargo install rars-cli

And the links are here: