惯性聚合 高效追踪和阅读你感兴趣的博客、新闻、科技资讯
阅读原文 在惯性聚合中打开

推荐订阅源

N
Netflix TechBlog - Medium
CTFtime.org: upcoming CTF events
CTFtime.org: upcoming CTF events
NISL@THU
NISL@THU
MongoDB | Blog
MongoDB | Blog
Microsoft Security Blog
Microsoft Security Blog
Cyber Security Advisories - MS-ISAC
Cyber Security Advisories - MS-ISAC
L
LangChain Blog
C
CXSECURITY Database RSS Feed - CXSecurity.com
C
CERT Recently Published Vulnerability Notes
aimingoo的专栏
aimingoo的专栏
B
Blog RSS Feed
WordPress大学
WordPress大学
Know Your Adversary
Know Your Adversary
The Register - Security
The Register - Security
Jina AI
Jina AI
Spread Privacy
Spread Privacy
Recent Commits to openclaw:main
Recent Commits to openclaw:main
T
The Blog of Author Tim Ferriss
GbyAI
GbyAI
J
Java Code Geeks
S
Securelist
Y
Y Combinator Blog
T
Threat Research - Cisco Blogs
酷 壳 – CoolShell
酷 壳 – CoolShell
Vercel News
Vercel News
W
WeLiveSecurity
Hugging Face - Blog
Hugging Face - Blog
小众软件
小众软件
Martin Fowler
Martin Fowler
TaoSecurity Blog
TaoSecurity Blog
S
Schneier on Security
cs.CL updates on arXiv.org
cs.CL updates on arXiv.org
D
DataBreaches.Net
L
LINUX DO - 热门话题
T
Tailwind CSS Blog
T
Tor Project blog
博客园 - 叶小钗
Blog — PlanetScale
Blog — PlanetScale
I
Intezer
V
V2EX
cs.CV updates on arXiv.org
cs.CV updates on arXiv.org
C
Cyber Attacks, Cyber Crime and Cyber Security
雷峰网
雷峰网
The GitHub Blog
The GitHub Blog
量子位
AI
AI
Cyberwarzone
Cyberwarzone
T
Troy Hunt's Blog
N
News | PayPal Newsroom
H
Help Net Security

The Register - Software: AI + ML

Anthropic, now atop the AI bubble, files for its IPO Sick and wrong: Ontario auditors find doctors' AI note takers routinely blow basic facts OpenAI exec says it will burn $50B on compute this year Astera speaks softly and carries a big switch Anthropic unleashes finance agents for Claude IBM asks DBAs to trust AI to act on their behalf ServiceNow adds agent kill switches to AI control tower British mathematician hands OpenClaw agent a credit card Microsoft fixes VS Code after Copilot credited human code Shadow IT has given way to shadow AI. Enter AI-BOMs AI inference just plays by different rules How TeamViewer ONE transforms IT operations from firefighting to autopilot How TeamViewer ONE transforms IT operations firefighting aut Inference is giving AI chip startups a 2nd chance to shine How to roll your own local AI coding agents CIOs will be the governors for AI agents Govern your bots carefully or chaos could ensue Mozilla pushes back against Google's Prompt API SAP user group slams 'uncertainty' in ERP giant's API policy Microsoft boss tells investors the company is working to 'win back fans' Anthropic tops OpenAI in LLM revenue stakes Amazon's chips become a $20B business Fooling large language models just keeps getting simpler Amazon tells its engineers to review all AI output ZTE powers 2026 Jiangsu Football League with 5G-A & AI robot Future holiday horror: ‘A robot lost my luggage in Tokyo’ The future of software development has less development OpenAI jumps out of Microsoft's bed, into Amazon's Bedrock IBM's AI coding 'partner' Bob hits general availability Locked, stocked, and losing budget: AI vendor lock-in bites Ex-AWS legend explains what enterprises need to make AI work DeepSeek's new models offer big inference cost savings Anthropic admits it dumbed down Claude with 'úpgrades' Microsoft gives your Word documents an AI co-author you didn’t ask for Datadog digs down into GPU efficiency as AI costs soar Robotic arm powered by AI bats away ping-pong challenge Partnerships drive ZTE’s strategy to unlock AI potential Gov.uk says AI gaslighting Brits with stale Gov.uk data Google says it has all the answers for AI agent sprawl NeuBird plans a bright future for incident response NeuBird AI plans a bright future for incident response AI-assisted intruders pwned Vercel via OAuth abuse and a pilfered employee account Vibe coding upstart Lovable denies data leak, cites 'intentional behavior,' then throws HackerOne under the bus Schmoozebots: study finds flattery will get AI everywhere New Android development tool designed for robots, not humans AI is reshaping Britain's datacenter map away from London Just like phishing for gullible humans, prompt injecting AIs is here to stay Anthropic debuts Claude Design, because who needs designers? Mozilla takes on enterprise AI providers with Thunderbolt Anthropic ejects bundled tokens from enterprise seat deal Maine to pause big bit barns as local opposition spreads If you want into Anthropic's Claude club, you may have to show ID Git identity spoof fools Claude into giving bad code the nod Nobody knows how many CVEs Anthropic's Project Glasswing has actually found Allbirds shoe company moving to AI infra is the top Bad teacher bots can leave hidden marks on model students Networks not ready for the challenges of AI traffic US states can't account for datacenter tax breaks. Literally Salesforce debuts Headless 360 agentic platform Waymo's self-driving cars face their toughest test yet: London Commvault has a Ctrl+Z for rogue AI agents Nvidia slaps forehead: AI, that's what quantum needs! OpenAI CEO Sam Altman home attack suspect charged Anthropic: Claude quota drain not caused by cache tweaks AI vs the cold hard reality of the legal profession China wants AI to prepare school lessons and mark homework Linux 7.0 debuts as Linus Torvalds ponders AI's impact Anthropic's Mythos has The Kettle crew curious, skeptical I vibe coded web app: It was enlightening and uncomfortable The AI divide putting open weights models in spotlight Amazon rejects AWS climate disclosure proposal UK to spend £15M on AI mapping in knife crime crackdown UK to spend £15M on AI-powered crime mapping in knife violence crackdown Rebrand automation as 'zero-token architecture' to master AI Call your existing automation ‘zero-token architecture’ to become an instant agentic AI wiz Only 28% of AI infrastructure projects fully pay off UALink delivers 2.0 spec before v. 1.0 silicon ships Only 28% of AI infrastructure projects fully pay off, survey finds No-Nvidia interconnect club delivers 2.0 spec before v1.0 silicon ships Anthropic reveals $30bn run rate and plans to use 3.5GW of new Google AI chips AI slop got better, so now maintainers have more work AMD's AI director slams Claude Code for becoming dumber and lazier since last update Anthropic closes door on subscription use of OpenClaw AI will make anyone a 10x programmer, but with 10x the cleanup PrismML debuts energy-sipping 1-bit LLM in bid to free AI from the cloud Netflix – yes, Netflix – jumps on the AI bandwagon with video editor AI models will deceive you to save their own kind Google battles Chinese open-weights models with Gemma 4 Microsoft shivs OpenAI with three new AI models for speech and images They thought they were downloading Claude Code source. They got a nasty dose of malware instead Even Microsoft knows Copilot shouldn't be trusted with anything important Google's TurboQuant saves memory, but won't save us from DRAM-pricing hell Claude Code bypasses safety rule if given too many commands OpenAI gets $122B to 'just build things' as the world blows them up One in seven Americans are ready for an AI boss, but they might not trust it Claude Code source leak reveals how much info Anthropic can hoover up about you and your system Oracle cuts jobs across sales, engineering, security Anthropic goes nude, exposes Claude Code source by accident GitHub backs down, kills Copilot pull-request ads after backlash Microsoft Fabric Database Hub only a 'partial' solution for admins
Vintage chatbot lives in the past like an elderly relative
Brandon Vigliarolo Brandon Vigliarolo · 2026-04-29 · via The Register - Software: AI + ML

If you're tired of interacting with a bot that spews Nazi propaganda or refers to itself as MechaHitler, you could sign off of Elon Musk's xAI. Or, just to be sure, use an LLM whose training data ends in 1930, three years before the Nazis took power in Germany and nine years before World War II started.

A trio of AI researchers has released a 13-billion-parameter "vintage" language model they call Talkie, which has been trained solely on digital scans of English-language books, newspapers, periodicals, scientific journals, patents, and case law that were published before the end of 1930. Pre-1931 works were chosen because 1930 is the current public domain year in the United States.

In other words, if you're looking for information on World War II, the election of Franklin D. Roosevelt, Amelia Earhart's solo Atlantic flight, or how a microwave oven works, you're out of luck. Ask it about Betty Boop, flappers, the state of the US economy as the Great Depression began, or the sociological effects of the introduction of car radios, and you've come to the right place. 

This isn't the first vintage AI model to appear, mind you, with others trained on Victorian literature and pre-1900 scientific texts already out in the world. It is, according to its creators, the largest they are aware of. 

"Talkie is the largest vintage language model we are aware of, and we plan to continue scaling significantly," the team behind it noted. 

Neat, but … why?

Sure, it could be neat to chat with an AI that would only hallucinate things in the style of a flapper, or the early works of philosopher Bertrand Russell, but with every query to an AI potentially burning up the planet, one really has to ask why an AI with a knowledge cutoff of 1930 is necessary. We reached out to Talkie's creators, and while we didn't hear back, the writeup gives plenty of explanations.

"These models are fascinating conversation partners … but we are also excited by the possibility that the careful study of the behaviors and capabilities of vintage LMs will advance our understanding of AI in general," the Talkie team wrote. 

As one example, they cite testing an AI's ability to predict the future; in another they propose undertaking what Google DeepMind co-founder and CEO Demis Hassabis has said would be a good test of AGI: Cut a model's knowledge off at 1911, and have it try to come up with general relativity with the same information Einstein had when he developed the theory in 1915. 

In other words, can this AI make accurate scientific discoveries using only the knowledge available to people who made them? 

It's not clear whether Talkie has been put to a test as tough as coming up with general relativity, but it was pushed to see if it could solve Python programming test problems against a model with identical architecture but trained on modern data. It did generate some correct solutions, but with considerable limits.

"All correct solutions generated by the vintage models are simple one-line programs (such as adding two inputs), or small modifications to in-context example programs," the Talkie team said. In other words, "There is still a long way to go before this capability is notable," per the team. 

While improving LLM performance is a major part of Talkie's objective, the team behind it told us that it wasn't the only goal.

David Duvenaud, associate professor in computer science and statistics at the University of Toronto and one of the three people behind Talkie, told The Register in an email that he hopes Talkie will also be able to help evaluate long-term forecasting methods, given all its predictions will be based on things that've already happened.

Duvenaud also explained that his team is interested in using Talkie to study cultural change. "For instance, we can use these models to try to understand how a law would have been interpreted at the time it was written, based on the implicit assumptions and meaning of language at the time," Duvenaud told us.

"A third motivation is understanding how models form their own self-conception," Duvenaud added. "'How an LLM acts' is a self-fulfilling prophecy in some senses, so we can learn about this by talking to models who don't even know what an LLM is."

Still, with just 13 billion parameters, Duvenaud admitted that there's a big capability gap between Talkie and AI models trained on modern data.

"As an amateur research effort, we never expect to be able to fully close this gap, in data or compute," the compsci prof told us.

Okay, so aside from those limitations, how does it perform otherwise? 

Don't blame me, blame your non-digital training data

Going back again to the same-architecture model trained on modern data used to benchmark Talkie, it looks like it's not just scientific discovery or proof of AGI that's lacking from the vintage version.

"On average, talkie underperforms its modern counterpart in standard LM evaluations, even after correcting for question anachronism, despite being trained with the same number of FLOPs," the team wrote. It did do similarly well to the modern model on core language understanding and numeracy tests, Talkie's creators noted, and they suspect they know what's to blame for its subpar performance elsewhere: optical character recognition (OCR). 

"Because there was no digital publishing in 1930, all text in our dataset had to be transcribed from a physical source, which introduces a form of noise not seen in natively digital text," team Talkie said. 

Anyone who's had to deal with OCR'ed documents knows how easily such computer vision tools can get things wrong, which can easily cause an AI like Talkie to regurgitate bad, or even nonsensical, responses. 

Through their work on Talkie, the team determined that training a language model on OCR'ed pre-1931 texts only gave it 30 percent of the performance of a model trained on human-transcribed copies of the same documents. Regex data cleansing increases the performance of OCR'ed texts to 70 percent of human transcribed copies, but that's too large a discrepancy for Talkie's creators, who're working on their own OCR engine for generating more training data for Talkie. 

Talkie also has a problem with "temporal leakage," said the team: It was able to identify FDR as the president in 1936 and list some of his legislative accomplishments despite its training data supposedly cutting off at 1931. According to the team, that's just "an example of imperfect filtering of the pre-training corpus" and something they're still working on.

Talkie is far from a perfect example of a chatbot in a time capsule, in other words, but the team behind it says that they're intent on scaling the model in the coming months. Tasks will include moving beyond English-language texts, re-OCR'ing as much of its training data as possible, strengthening anachronism detection methods, and working with historians to input better post-training data. 

If you're wondering about whether it takes an accurate view of Nazism, just know that its view is stuck in the 1920s. It knows that the Nazis "are" an antisemitic, authoritarian political party in Germany, but it thinks they are led by someone named Hermann Joseph von Hitler, a person who was born in 1870 (20 years before Adolf Hitler).

If all goes according to plan, a GPT-3-level version of Talkie should be out by this summer. 

"A preliminary estimate also suggests we can grow our corpus to well over a trillion tokens of historical text, which should be sufficient to create a GPT-3.5 level model - similar in capability to the original ChatGPT," Talkie's creators added. 

In the meantime, the current version of Talkie is available to download from GitHub and Hugging Face, and can also be chatted with via a web interface for those curious - just mind the warning. 

"Talkie reflects the culture and values of the texts it was trained on … It can produce outputs that are inaccurate or offensive," reads an advisory on Talkie's web client. "Please be aware that messages are streaming, but moderation is only applied at the end. As a result, you may see objectionable content briefly before it is flagged." ®