惯性聚合 高效追踪和阅读你感兴趣的博客、新闻、科技资讯
阅读原文 在惯性聚合中打开

推荐订阅源

S
Secure Thoughts
B
Blog
MongoDB | Blog
MongoDB | Blog
GbyAI
GbyAI
博客园 - 【当耐特】
D
DataBreaches.Net
Apple Machine Learning Research
Apple Machine Learning Research
阮一峰的网络日志
阮一峰的网络日志
I
InfoQ
人人都是产品经理
人人都是产品经理
Microsoft Azure Blog
Microsoft Azure Blog
量子位
美团技术团队
Recent Announcements
Recent Announcements
freeCodeCamp Programming Tutorials: Python, JavaScript, Git & More
M
MIT News - Artificial intelligence
O
OpenAI News
SecWiki News
SecWiki News
A
About on SuperTechFans
J
Java Code Geeks
B
Blog RSS Feed
Y
Y Combinator Blog
L
LangChain Blog
Security Archives - TechRepublic
Security Archives - TechRepublic
Attack and Defense Labs
Attack and Defense Labs
小众软件
小众软件
CTFtime.org: upcoming CTF events
CTFtime.org: upcoming CTF events
Martin Fowler
Martin Fowler
博客园 - 聂微东
雷峰网
雷峰网
有赞技术团队
有赞技术团队
Google DeepMind News
Google DeepMind News
T
The Exploit Database - CXSecurity.com
N
News and Events Feed by Topic
OSCHINA 社区最新新闻
OSCHINA 社区最新新闻
Microsoft Security Blog
Microsoft Security Blog
Recorded Future
Recorded Future
P
Palo Alto Networks Blog
Blog — PlanetScale
Blog — PlanetScale
N
News | PayPal Newsroom
Scott Helme
Scott Helme
L
LINUX DO - 热门话题
F
Full Disclosure
cs.AI updates on arXiv.org
cs.AI updates on arXiv.org
The Hacker News
The Hacker News
cs.CL updates on arXiv.org
cs.CL updates on arXiv.org
U
Unit 42
博客园_首页
T
Tailwind CSS Blog
T
Tenable Blog

Hacker News

Introducing Claude Opus 4.7 Qwen Studio The Future of Everything is Lies, I Guess: Where Do We Go From Here? GitHub - SeanFDZ/macmind: Single-layer transformer in HyperTalk for the classic Macintosh Show HN: Agent-cache – Multi-tier LLM/tool/session caching for Valkey and Redis Bonsai 1-bit WebGPU - a Hugging Face Space by webml-community Moving a large-scale metrics pipeline from StatsD to OpenTelemetry / Prometheus GitHub - Nightmare-Eclipse/RedSun: The Red Sun vulnerability repository GitHub - SethPyle376/hiraeth: Local AWS emulator focused on fast integration testing, with SQS support, SQLite-backed state, and a debug-friendly web UI. GitHub - macOS26/Agent: Any AI, replaces Claude Code, Cursor, OpenClaw. Over 18 LLM providers (Claude, OpenAI, Gemini, Ollama, Zai, HF, Qwen) wired into a native Mac app that writes code, builds Xcode projects, bumps versions, manages git, automates Safari, use AppleScript, JS or Accessibility, extend Agent! w/ MCP Servers, run tasks from your iPhone via Messages. YouTube now lets you turn off Shorts I Made a Terminal Pager Burgers | マクドナルド公式 Commands — HackerNews CLI documentation ChatGPT for Excel PiCore - Raspberry Pi Port of Tiny Core Linux Live Nation illegally monopolized ticketing market, jury finds Google Broke Its Promise to Me. Now ICE Has My Data. Founding Engineer at Adaptional | Y Combinator CRISPR takes important step toward silencing Down syndrome’s extra chromosome GitHub - saffron-health/libretto: The AI toolkit for building reliable browser automations US v. Heppner (S.D.N.Y. 2026) no attorney-client privilege for AI chats [pdf] Retrofitting JIT Compilers into C Interpreters IPv6 – Google The Accursèd Alphabetical Clock Cybersecurity Looks Like Proof of Work Now Fragments: April 14 Cal.com Goes Closed Source: Why AI Security Is Forcing Our Decision | Cal.com - Scheduling Software for Online Bookings Laravel raised money and now injects ads directly into your agent When moving fast, talking is the first thing to break Too much Discussion of the XOR swap trick – Heather Cafe Introduction to Spherical Harmonics for Graphics Programmers The Grand Line Building a Z-Machine in the worst possible language High-Level Rust: Getting 80% of the Benefits with 20% of the Pain GitHub - duguyue100/midnight-captain: Inspired by Midnight Commander, tailored to my taste. How to build a `git diff` driver · Jamie Tanna | Software Engineer Center for Responsible, Decentralized Intelligence at Berkeley The Local Universe’s Expansion Rate Is Clearer Than Ever, but Still Doesn’t Add Up - A new synthesis of astronomical measurements confirms a persistent mismatch that could point to physics beyond current models The air throughout our homes is infused with microplastics. But there are things you can do to breathe less of them The disturbing white paper Red Hat is trying to erase from the internet – OSnews The Future of Everything is Lies, I Guess: Annoyances ‘Abhorrent’: the inside story of the Polymarket gamblers betting millions on war Productive procrastination — Max van IJsselmuiden maps, territory and LMs 447 Terabytes per Square Centimetre at Zero Retention Energy: Non-Volatile Memory at the Atomic Scale on Fluorographane Show HN: Pardonned.com – A searchable database of US Pardons 20 Years on AWS and Never Not My Job The Seasons are Wrong Artemis II crew splashes down near San Diego after historic moon mission We gave an AI a 3 year retail lease in SF and asked it to make a profit | Andon Labs How a dancer with ALS used brainwaves to perform live On filing the corners off my MacBooks Installing every* Firefox extension OpenClaw’s memory is unreliable, and you don’t know when it will break Steve Blank Nowhere Is Safe Chimpanzees in Uganda locked in vicious 'civil war', say researchers watgo - a WebAssembly Toolkit for Go linux/Documentation/process/coding-assistants.rst at master · torvalds/linux GitHub - callumlocke/json-formatter: Makes JSON easy to read. Founding Product Engineer at Bild AI | Y Combinator A compelling title that is cryptic enough to get you to take action on it GitHub - Keychron/Keychron-Keyboards-Hardware-Design: Industrial design files for Keychron keyboards and mice. 100+ models with CAD assets in STEP, DXF, DWG, and PDF. Source-available, with commercial use allowed for original compatible accessories within the license terms. [ANNOUNCE] WireGuardNT v0.11 and WireGuard for Windows v0.6 Released 1D-Chess Helium Is Hard to Replace Cooperative Vectors Introduction | Evolve Keeping a Postgres queue healthy — PlanetScale Our response to the Axios developer tool compromise Do Americans read print books, e-books or audiobooks more? The Zettelkasten Method in Obsidian: A Practical Setup Guide Artemis II Is Competency Porn and We Are Starving For It WeakC4 Flight Viz — Cockpit View A Mexican surveillance giant you’ve never heard of is now watching the U.S. border Surelock: Deadlock-Free Mutexes for Rust RISC-V 101 – what is it and what does it mean for Canonical? | Ubuntu The Problem That Built an Industry How Much Linear Memory Access Is Enough? | Solidean Investigating Split Locks on x86-64 Simplest hash functions Sybilproof reputation mechanisms (2005) [pdf] What is a property? How Complex is my Code? Static code analysis in Kotlin — tools overview Toffoli gates are all you need PGLite evangelism dcmake: a new CMake debugger UI Clojure on Fennel part one: Persistent Data Structures Fragments: April 2 Python Release Python install manager 26.1 The Life and Death of the Book Review - Liberties Introducing Database Traffic Control — PlanetScale Bitcoin miners are losing $19,000 on every BTC produced as difficulty drops 7.8% God sleeps in the minerals Building slogbox Apple Silicon and Virtual Machines: Beating the 2 VM Limit Who was “Not Even Wrong” first? Pokemon Evolution Vs Darwinian Evolution The APL Programming Language Source Code
It's Not Just X. It's Y.
Eryk Salvaggio · 2026-05-31 · via Hacker News
A colorful but noisy image of a man playing a trumpet without a shirt.
A recent experiment testing the limits of noise-generation in diffusion models paired with verbs in the kinetic identification dataset, which has nothing to do with the topic of this post.

Against the Quantification of Integrity

When the measure of language becomes its target, it ceases to be good language.

💡

Nerd Rating: 1/5. I discuss the origins of certain linguistic tics in LLMs and what it means for writing, student assessment, and thinking.

"It's not x, it's y."

Large Language Models gravitate toward this type of construction, called negative parallelism. It has its uses: it sets up a contrast. It's useful, especially, for reframing assumptions: "You think it's like that, but it's really like this."

It's all over social media, especially on LinkedIn, and the construction has sparked a backlash amid an ongoing war against automated language production. If you use em-dashes – you might be a bot. If you describe things that delve, quietly, or genuinely (or create lists of three, like that one), you might be a bot.  

Recent overuse by language models has led many to declare it bad writing. I'm not so sure. Nobody called JFK a lazy writer when he said, "ask not what your country can do for you – ask what you can do for your country." Negative parallelism is a rhetorical device, and any rhetorical device is only as lazy or inspired as what it contains.

Automated Language Production

Now, we have AI detectors that claim to protect you from the witch hunt by looking for these patterns. You take your own writing and you run it through Grammarly, which will analyze word patterns that AI detectors might flag. Then it offers ideas for how to change them, which a) gives Grammarly the power to write for you and b) makes your writing lose any sense of rhythm or intent.

Grammarly's review of this section has flagged 27 examples of text I should change to avoid the accusation that I am a machine. For example, Grammarly identified the above phrase – "automated language production" – as 11 times more likely to be AI. It suggests that a human would be "against mechanized language synthesis" instead. The simple two-word combo, "align with" was flagged as 43x more likely to be AI-generated. Real humans say "corresponds." These are small suggestions that add up until the result resembles nothing I chose. The human voice replaced by a machine trying to sound human.

As a result, I just paid Pangram – another AI-detection company – $20 to verify that a recently submitted journal article wasn't AI-generated before submission. It wasn't, and I knew it wasn't. It agreed. That's what I paid for: not to learn whether I wrote it, but to be told it wouldn't flag me. Because if Pangram's AI system found me guilty, that's the end of my career. That's literally extortion.

And if it had flagged it, then what? It would give me a score (four valuations: high, very likely, somewhat likely, human) to assign my integrity a category. In the ecosystem we're all building, I'd have to use Grammarly to rephrase everything: using a machine to write for me to prove that I didn't use a different machine to write for me.

A Culture Hostile to Reason

Our instinct in making sense of these machines is to examine the training data. That training data is no longer "just the Web." The web is the raw meat, but this sausage is heavily pre- and post-processed. Post-training optimizes the model for whatever it's designed to do. This includes techniques such as RLHF (reinforcement learning with human feedback) and RLVR (reinforcement learning through verified rewards). RLHF has humans rank replies, then the system emphasizes those kinds of replies.

RLVR is weirder, and I suspect it's why we see "It's not X, it's Y" so often. Dismissing negative parallelism as lazy gets in the way of understanding why it's showing up everywhere. This type of language is such a powerful framework for thinking that we mistake it for a model's capacity for thought. We credit computation for the work that's done by language.

Weird Dogs

RLVR isn't a structure that watches for words and triggers some sub-process. Instead, you train a model, like you would any model. When that model is done, it predicts tokens. Lots of people are still in denial about this. Token prediction involves producing a list of candidates based on their mathematical distribution in the training data, ranking them by their likelihood given the previous words in the prompt or sequence.

RLVR intervenes by having the model solve math problems by writing their way to a solution, reproducing the language we would use when thinking out loud about how to solve it. When the model arrives at the correct answer, the language it used most often to get there is then emphasized in the finished model. This is (partly) what the industry calls reasoning.

What day was it that we saw that weird dog?

So, think of it like this: You are sitting with a friend. Your phones are dead. Your friend asks: what day was it that we saw that weird dog? You start by saying, "It was Thursday." Your friend says: "No, it wasn't Thursday, because Thursday I was out of town." So you say that's right, so it must have been Wednesday, because Wednesday was your mutual friend's birthday, and you both went to the party, and you saw the dog on the way to the party. Your friend says: "That's right, except, Wednesday was our friend's birthday but the party was on Friday. So we must have seen the dog on Friday."

The two of you have articulated your way to the answer, a verifiable one: you could pop on your phones and check your photos and see that yes, the weird dog picture was taken on Friday. In dehumanizing terms, your gut instinct ("it's Thursday") is what a model might spit out at first guess, and that's where models used to stop.

But you didn't. Your friend countered: "It wasn't [Thursday], it was [Wednesday]." There are more words, which narrow the window of possible answers, and then you arrive, through "its-not-x-its-y-ing," at the correct date. The two of you had actual memories and visceral experiences to work with. Language was the vessel through which these experiences were communicated and conflicts were resolved. The model, by contrast, extends language in longer and longer bursts, replicating the pattern of reasoning you two just engaged in. These longer runs re-enact that deliberation within language rather than through it.

Other high-entropy states get filled by words like "suppose..." which triggers longer speculative passages. "Because," "consider," "alternatively," even "wait" can occupy these positions. These are words that lead to language that brings contrast, exceptions, and abstraction along for the ride. If they get to a correct answer on a math problem, they get pushed to occur more often.

The Reason We Reason

When we talk about a weird dog or have conversations like it, the point of the question was not to identify the date on the calendar when the dog was encountered. It was an opening for a reminiscence. It was posed to reconstruct the memory, to revel in its surrounding context, and to deepen a connection between friends through a shared experience.

Defining reasoning this way assumes that the point of asking a question is to get an answer, that answers can be verified, and that nothing is lost in immediate closure.

Defining reasoning the way it has been used in LLMs assumes that the point of asking a question is to get an answer, that answers can be verified, and that nothing is lost in immediate closure. This has real effects on writing, and the openness to doubt is something we lose in the rapid prototyping of thought that occurs with a language model. Ambiguity, doubt, and uncertainty matter more to some ways of thinking than any immediate answer. The inner life grows in the spaces between the industrial complexes that harness every remnant of our externalized thought.

Nonetheless, the language we use in these states is the same. When AI detectors flag text as AI-generated, is it because it follows a certain structural pattern of that reasoning? Pangram and reasoning models both detect structural patterns based on how humans reason when writing. Pangram's model is trained on pre-2021 data; it then inserts AI-generated versions of the same text into its training.

So, if we publicly shame people whose text looks like it might have been written by a machine – because it mimics the language used for human reasoning – and people stop writing in ways that they internalize as "AI writing" out of fear of false detection, it sends a signal that your language for reasoning must be policed, or you too could be held up to public scrutiny.

In the end, shaming people for writing that gets flagged as AI can lead people to sidestep structures the model has learned from us: structures that are effective tools for argumentation. We take the tools of critical thinking out of the kit at the time we most need them.

For Good Measure

There's another angle to this. An AI-based essay assessment tool was tested in the UK against human graders. The system rewarded writing structures that I can't help notice look a lot like RLVR-based reasoning: "giving out higher marks based on essay length, vocabulary range and sentence complexity, which are often unrelated to academic standards," all of which are hallmarks of AI reasoning.

In other words, the LLM grades humans based on the criteria engineers use to assess the LLM.

The LLM grades humans based on the criteria engineers use to assess the LLM.

There's this old adage from economics called Goodhart's law. The econo-nese version of it is that "any observed statistical regularity will tend to collapse once pressure is placed upon it for control purposes." Or: when a measure becomes a target, it ceases to be a good measure. It could be tweaked to apply to large language models: "when the measure of language becomes its target, it ceases to be good language."

There is danger in evaluating for language patterns over its content, and both generation and detection incentivize this. Automated grading is somewhere between the two: rewarding students for employing the form of reason over the act of reasoning will only make them more tempting and more common. And yet, punishing the form risks punishing reason. Ultimately, we have to think critically in all cases, instead of deferring to the judgments of machines.

Against Automatic Thinking

I'm not convinced by the old "if you haven't done anything wrong, you don't have anything to worry about" line. I've seen 99.8% cited as a measure of accuracy in automated surveillance systems since 2018. As Arvind Narayanan has noted, that is on a per-paper basis, which compounds every time we use it. So up to 10% of college students could be falsely accused. If we collectively run every bit of text through an AI model to check whether it is AI-generated, we will generate false positives on an even larger scale.

These models concentrate real authority; companies promise they will reason on our behalf. We normalize something dangerous when we run every two-line phrase through an AI interpreter, post the result online, and say "see? They're plagiarists!"

We create a culture of self-censorship and AI-detector-pressured rewriting and paraphrasing as people strive to avoid these witch hunts. That is the opposite of protecting human expression. We should resist normalizing a trust in any machine's ability to determine matters of guilt. If using AI to write is, at its worst, an industrialization of the mind, then AI detection, at its worst, becomes a surveillance system for thought.


Monthly, for the Second Week in a Row.

Thanks for reading! As mentioned last week, I am only a sporadic poster these days, aiming for once a month. If you're paying for the newsletter and would like to calibrate your donations (or would like to start supporting it!) you are very welcome to set up or change your subscription here.