惯性聚合 高效追踪和阅读你感兴趣的博客、新闻、科技资讯
阅读原文 在惯性聚合中打开

推荐订阅源

D
DataBreaches.Net
IT之家
IT之家
The Cloudflare Blog
Apple Machine Learning Research
Apple Machine Learning Research
WordPress大学
WordPress大学
N
Netflix TechBlog - Medium
阮一峰的网络日志
阮一峰的网络日志
P
Proofpoint News Feed
L
LangChain Blog
博客园 - Franky
美团技术团队
J
Java Code Geeks
Microsoft Security Blog
Microsoft Security Blog
博客园 - 叶小钗
小众软件
小众软件
Y
Y Combinator Blog
B
Blog RSS Feed
Cyber Security Advisories - MS-ISAC
Cyber Security Advisories - MS-ISAC
D
Docker
Hugging Face - Blog
Hugging Face - Blog
Jina AI
Jina AI
罗磊的独立博客
大猫的无限游戏
大猫的无限游戏
Vercel News
Vercel News

Hacker News: Show HN

PurrrrrFocus: Pomodoro Timer App - App Store Workflow Engine — Multi-Step Orchestration for Bun RapidPhoto: Pro Photo Editor App - App Store GitHub - DheerG/swarms: Achieve extraordinary results with claude code across a variety of tasks SPICE simulation → oscilloscope → verification with Claude Code — Lucas Gerads Show HN: VCoding – A 5 MB native Windows IDE with no dynamic dependencies Show HN: LLMs don't hallucinate because they're bad at math, it's the format GitHub - Agent-FM/agentfm-core: AgentFM is a peer-to-peer network that turns everyday computers into a decentralized AI supercomputer. AgentFM lets you run massive AI workloads directly across a global mesh of idle CPUs and GPUs. Show HN: Tracking Top US Science Olympiad Alumni over Last 25 Years GitHub - Potarix/agent-hub: One place to talk to all your agents Show HN: Runtime security for AI agents(injection,tool abuse, data exfiltration) GitHub - dubeyKartikay/lazyspotify: Terminal Spotify client for macOS and Linux GitHub - the-banana-tool/king-louie: Easy to use GUI Personal AI Assistant. Win/Linux/Mac. Show HN I made my vacation rental bookable by AI agents–no Airbnb, 0% commission GitHub - basteez/jsf-autoreload: maven plugin to enable hot reload on jsf projects uvm32/hosts/host-gdbstub at main · ringtailsoftware/uvm32 GitHub - labsai/EDDI: Config-driven engine that turns JSON into production-grade AI agents. Multi-agent orchestration, 12+ LLM providers, MCP/A2A protocols, RAG, persistent memory, and enterprise compliance (EU AI Act, GDPR, HIPAA). Built on Quarkus. GitHub - glitchnsec/fortyone-oss: AI Executive Assistant Platform Quickstart | Alien GitHub - muxshed/shed: One stream in, or many. Every destination, simultaneously. No cloud middleman, no per-channel fees, no limits. GitHub - ocrbase-hq/ocrbase: 📄 PDF/IMG ->.MD/JSON Document OCR API for PaddleOCR and GLMOCR. Self-hostable. GitHub - impactjo/home-memory: MCP server that lets your AI assistant remember everything about your home. GitHub - Sets88/dbcls: DbCls is a powerful terminal database client that supports various databases GitHub - neptun2000/heor-agent-mcp GitHub - SeanFDZ/macmind: Single-layer transformer in HyperTalk for the classic Macintosh RollQuation: Math Puzzles - Apps on Google Play GitHub - dropbox/witchcraft Show HN: Agent-cache – Multi-tier LLM/tool/session caching for Valkey and Redis GitHub - opentalon/opentalon: OpenTalon is an open-source platform built from the ground up in Go as a robust alternative to OpenClaw LinkedIn™ 职位抓取工具 - Chrome 应用商店
Who
heterodoxjed · 2026-06-23 · via Hacker News: Show HN

A small experiment in language-model memory

When a language model is trained, what it absorbs about the world gets baked into its weights — the billions of numbers that hold everything it knows. A person is in the weights if a model can recall them on its own, without looking them up — but that's rarely all-or-nothing, so intheweights.com asks 13 models how sure each is, scoring its confidence from 0 to 100. (Combine a person's 13 confidences and the site gives them one strength score; we work with the 13 underlying numbers.) We ran 291 real people through it — from household names to the genuinely obscure — and looked for patterns.

How the 13 models score one person

Each model's confidence runs from 0 ("never heard of them") to 100 ("sure who they are"). It's not yes-or-no, and it's not one verdict — for the same person, the 13 answers can land anywhere from 0 to 100. Here's one person — a Georgian judoka:

The bars are the 13 models' confidence. Some are sure; several have no idea who he is — pooled together, those make up his strength. Now do this for 291 people.

More looked up → better known, but loosely

Each dot is one of 185 people. Across the bottom: how often they're looked up — their monthly Wikipedia pageviews. Up the side: how well the 13 models know them, their 13 confidences averaged into one score. It climbs — household names sit near the top — but loosely: among the rarely-looked-up, the models know some people surprisingly well and draw a blank on others just as obscure.

Pageviews are on a log scale — each step right means about 10× more monthly views, so the famous and the obscure fit on one chart. Hover any dot for who it is, its score, and a Wikipedia preview; click through to the article.

Some models are far more confident than others

Average each model's confidence across the people we ran, and the averages run from one model that's sure of almost everyone (about 90 out of 100, top) down to one that's blank on almost everyone (about 18, bottom). So how high a person scores depends a lot on which model you ask, not just on who the person is.

Bar length = that model's average confidence across those people, 0–100. Hover a model's name for its size and knowledge cutoff.

But they mostly rank people the same way

Scoring people high and ranking the same people high are two different things — a model can hand out higher numbers across the board but still put the same people on top. So set the overall levels aside and compare orderings: line up each model's people from best-known to least-known. By that test the models match moderately well: a typical pair ranks the people about 0.65 alike, where 1 would mean identical orderings and 0 means no relation at all. So they mostly agree on who is better-known than whom, even where they disagree on the exact scores. The clear exception is the smallest model, whose ordering barely matches the others.

Your line of work barely matters — except for athletes

You might expect some kinds of people to live in the weights more than others. Mostly they don't: split by occupation, recognition is strikingly flat — six of the seven groups sit within a few points of each other, at broadly similar fame. The one clear exception is athletes, who lag the rest. So apart from sport, what you did for a living barely changes whether the models know you.

Average recognition (0–100) per occupation, across the people we ran; their fame is broadly similar across groups, so this isn't just a fame gap. Athletes (orange) stand out.

A shared name costs you a little

Does sharing your name with other notable people hurt? We compared German footballers whose full name is theirs alone on Wikidata to ones whose exact name is shared by several other notable people, matched on fame — so the only real difference is the name:

The short version

Whether a model "knows" a person tracks how often the world looks them up — loosely. A shared name hurts a bit. And while the 13 models broadly agree on who ranks as better- or lesser-known, they differ enormously in how confident they are overall — so whether a borderline person counts as "in the weights" still comes down to which model you ask.

Every number here describes these 291 people — a deliberately wide spread from famous to obscure, not a random or representative sample. Read them as comparisons, not rates.