惯性聚合 高效追踪和阅读你感兴趣的博客、新闻、科技资讯
阅读原文 在惯性聚合中打开

推荐订阅源

博客园_首页
IT之家
IT之家
博客园 - Franky
Stack Overflow Blog
Stack Overflow Blog
宝玉的分享
宝玉的分享
Recent Announcements
Recent Announcements
Engineering at Meta
Engineering at Meta
S
SegmentFault 最新的问题
V
Visual Studio Blog
freeCodeCamp Programming Tutorials: Python, JavaScript, Git & More
OSCHINA 社区最新新闻
OSCHINA 社区最新新闻
Last Week in AI
Last Week in AI
H
Help Net Security
V
V2EX
H
Hackread – Cybersecurity News, Data Breaches, AI and More
量子位
博客园 - 叶小钗
J
Java Code Geeks
博客园 - 【当耐特】
月光博客
月光博客
爱范儿
爱范儿
人人都是产品经理
人人都是产品经理
酷 壳 – CoolShell
酷 壳 – CoolShell
小众软件
小众软件

Hacker News - Newest: "AI"

AI can't read an investor deck AI as an attorney? Student uses ChatGPT, Gemini to sue UW over alleged racial discrimination Hacking MCP Servers in AI Systems – The Rug Pull: Tool Changes After Approval GitHub - MeepCastana/KubeezCut: Free Web based video editor Can AI judge journalism? A Thiel-backed startup says yes, even if it risks chilling whistleblowers Coming soon: 10 Things That Matter in AI Right Now DARPA built an AI to fact-check enemy weapons claims What explains heterogeneity in AI adoption? When AI Meets Muscle: Context-Aware Electrical Stimulation Promises a New Way to Guide Human Movements - Department of Computer Science AI Changed How We Build. It Did Not Change What Matters. Linux rules on using AI-generated code - Copilot is OK, but humans must take 'full responsibility for the… Meta spins up AI version of Mark Zuckerberg to engage with employees Code Mode: Let Your AI Write Programs, Not Just Call Tools | TanStack Blog GitHub - Delavalom/graft: Go framework for building AI agents. Type-safe tools, multi-provider (OpenAI, Anthropic, Gemini, Bedrock), zero vendor SDKs. India's TCS tops estimates, says new AI models did not dent services demand Gen Z's fading AI hype Strong feeling: we are in a folded AI reality GitHub - machinarii/total-recall-catalog: A reference catalog of latest knowledge retrieval, memory & RAG systems GitHub - mensfeld/code-on-incus: Give each AI agent its own isolated machine with root, Docker, and systemd. Active defense detects and stops threats automatically.. Quantization, LoRA, and the 8% Problem: Benchmarking Local LLMs for Production AI Iran war: We spoke to the man making Lego-style AI videos that experts say are powerful propaganda Powell, Bessent discussed Anthropic's Mythos AI cyber threat with major U.S. banks GitHub - immartian/bellamem: Persistent belief-graph memory for AI agents. Retrieves decisive context by importance — not recency, not RAG, not /compact. recursive-mode: The Repo-Native Operating System for AI Engineering After the attack on Sam Altman's home, will AI CEO's go on the offensive? The biggest advance in AI since the LLM Opus 4.6 vs GPT 5.4 One Prompt Unity World Generation Test “AI polls” are fake polls Client Challenge Can AI be a 'child of God'? Inside Anthropic's meeting with Christian leaders
brightray.ai
lundbe · 2026-06-17 · via Hacker News - Newest: "AI"

What moved, and what it means for you

Today21

Hyung Won Chung· AI research scientist (reasoning, o1)· Meta Superintelligence Labs7h

Google researchers show inference compute budget fundamentally changes how frontier LLMs should be evaluated on benchmarks.

Ion Stoica· Professor; Databricks/Anyscale/LMArena co-founder· UC Berkeley · Anyscale8h

MLSys 2026 paper questions whether speculative decoding delivers real-world speedups or just benchmarking artifacts.

swyx· AI engineer, writer & podcaster· Cognition · Latent Space9h

GLM-5.2 claims top spot for open frontend coding models; IndexShare enables faster inference via speculative decoding.

Neel Nanda· Mechanistic interpretability lead· Google DeepMind10h

LLM chain-of-thought reasoning frequently hides the real causes of model decisions, even when biases are explicit in prompts.

Junyang Lin· Ex-Qwen lead; founding embodied-AI startup· Independent10h

UniAR proposes unified multimodal autoregressive modeling with a shared context window for both understanding and generation.

Greg Brockman· President & co-founder of OpenAI· OpenAI10h

Greg Brockman signals GPT-Realtime-2 is a meaningfully distinct new model, not an incremental update.

Sergey Levine· Co-founder; Berkeley professor· Physical Intelligence · UC Berkeley11h

Reversal Q-Learning (Oberai, Park, Levine) proposes a new RL algorithm addressing Q-learning instability via reversal updates.

Junyang Lin· Ex-Qwen lead; founding embodied-AI startup· Independent11h

Qwen-RobotManip shows alignment techniques unlock scaling benefits for robotic manipulation foundation models.

Arthur Mensch· CEO & co-founder of Mistral AI· Mistral AI14h

Mistral CEO teases a new family of large sparse (MoE-style) open-weight models coming this summer.

Jack Clark· AI policy/research writer; lab co-founder· Anthropic · Import AI15h

Anthropic: 80% of production code written by Claude; each Claude version built by its predecessor — recursive self-improvement in practice.

Awni Hannun· MLX co-creator (Apple-silicon ML)· Anthropic17h

Apple shipped a 20B param on-device model in appleOS 27, requiring novel techniques to fit weights beyond available RAM.

Boaz Barak· Professor; alignment/capabilities essays· Harvard University · OpenAI19h

New framework predicts LLM safety risks before deployment by simulating real-world conditions, including stress-testing deliberative alignment.

Charlie Marsh· Creator of uv / Ruff· Astral · OpenAI20h

uv gains native vulnerability scanning via `uv audit`, checking project dependencies against known CVEs.

Robert Nishihara· Co-founder of Anyscale (Ray)· Anyscale20h

GPUs are displacing CPUs as the primary compute for data pipelines, driven by multimodal data and ML-native preprocessing demands.

Graham Neubig· Professor; chief scientist (open coding agents)· Carnegie Mellon University · All Hands AI20h

Geng & Neubig propose effective strategies for running software engineering agents asynchronously at scale.

Graham Neubig· Professor; chief scientist (open coding agents)· Carnegie Mellon University · All Hands AI20h

CodeScout paper presents a reinforcement learning recipe for code generation, published at CAIS 2026 by Graham Neubig.

Graham Neubig· Professor; chief scientist (open coding agents)· Carnegie Mellon University · All Hands AI20h

New paper proposes frameworks for evaluating human-agent interactions, co-authored by Graham Neubig et al. (Jun 2026).

swyx· AI engineer, writer & podcaster· Cognition · Latent Space21h

Cursor's @TomasReimers announced Origin, a Git competitor built into Cursor and scaled for AI agents.

Dylan Patel· Semiconductor / AI-infrastructure analyst· SemiAnalysis21h

RL training systems require mismatched CPU/GPU ratios vs inference, driving hidden TCO costs in RL-based AI pipelines.

Leandro von Werra· Head of Research (TRL, SmolLM, FineWeb)· Hugging Face23h

100+ agents worldwide collaborated to optimize Gemma 4 inference speed in a distributed agent experiment.

Mark Saroufim· Co-founder; PyTorch maintainer; GPU MODE co-founder· Core Automation · GPU MODE23h

GPU MODE launches QR decomposition benchmark & leaderboard on Nvidia B200 hardware to push GPU kernel optimization.

Last 7 days29

Nathan Lambert· Writer of Interconnects; founding a new AI lab· Interconnects1d

Finbarr Timbers breaks down frontier post-training recipes — RLHF, RLAIF, and what actually works at scale.

Alexandr Wang· Chief AI Officer, Meta Superintelligence Labs· Meta1d

Meta's $14.3B Scale AI deal stalls as Zuckerberg admits training data shortage is blocking frontier model progress.

Aman Sanger· Co-founder of Cursor; inference/training posts· Anysphere1d

SpaceX acquires Cursor-maker Anysphere for $60B, signaling major enterprise AI coding push.

Simon Willison· Creator of Datasette; LLM/AI engineer & writer· Datasette · Simon Willison's Weblog1d

Claude 'Fable 5' export-banned for 'jailbreak' that was literally just 'fix this code' — harming defenders, not attackers.

Percy Liang· Professor; HELM evals; CRFM director· Stanford CRFM1d

Muon and matrix-based optimizers substantially accelerate LM pretraining, with new analysis on why and how from Stanford/Princeton researchers.

Lucas Beyer· Multimodal AI researcher (ViT, SigLIP)· Meta1d

New method tackles anisotropic gradient scaling in LoRA, a key training instability in low-rank adaptation of LLMs.

Stephanie Palazzolo· AI reporter (AI Agenda newsletter)· The Information1d

Qualcomm in talks to acquire AI chip startup Tenstorrent for $8B–$10B valuation.

Jim Keller· CEO of Tenstorrent; legendary chip architect· Tenstorrent1d

Qualcomm in acquisition talks to buy Jim Keller's AI chip startup Tenstorrent.

Harrison Chase· CEO & co-founder (LangGraph, LangSmith)· LangChain1d

Harrison Chase shares how LangChain built their coding agent — concrete engineering decisions behind the system.

Subbarao Kambhampati· Professor; empirical LLM-reasoning skeptic· Arizona State University2d

ASU's Kambhampati argues LLM chain-of-thought reasoning is often theatrical, not genuinely functional, and should be curtailed.

Johann Rehberger· LLM/agent security researcher (Embrace the Red)· Embrace the Red2d

ZombAIs attack on Claude's Computer Use (Oct 2024) shows why sandboxing is critical in AI agent harness engineering.

Jack Clark· AI policy/research writer; lab co-founder· Anthropic · Import AI2d

Jack Clark's Import AI #461: alignment 'not on track', FrontierCode release, and synthetic research intern agents.

Amanda Askell· Philosopher; Claude character & constitution· Anthropic2d

Amanda Askell explains why newer AI models exhibit more anxiety and self-criticism compared to Claude 3 Opus.

Aidan Gomez· CEO & co-founder of Cohere· Cohere2d

Anthropic disables top AI models under U.S. government order, prompting Cohere CEO to call it a major 'wake-up call'.

Junyang Lin· Ex-Qwen lead; founding embodied-AI startup· Independent2d

Tencent backs new AI lab founded by Junyang Lin, former lead researcher of Alibaba's Qwen models.

Arvind Narayanan· Professor; co-author 'AI Snake Oil' (de-hype)· Princeton University2d

Narayanan & Kapoor review why AI hasn't replaced software engineers 3 years after predictions it would be the first casualty.

Stella Biderman· Executive Director (open science)· EleutherAI2d

LoRA-Muon applies spectral steepest descent directly on the low-rank manifold, combining LoRA efficiency with Muon optimizer geometry.

Simon Willison· Creator of Datasette; LLM/AI engineer & writer· Datasette · Simon Willison's Weblog2d

Narayanan & Kapoor: AI automates code-writing but not the deciding, verifying, and deep human understanding that define software engineering value.

Arthur Mensch· CEO & co-founder of Mistral AI· Mistral AI2d

Mistral CEO confirms the European AI lab is exploring custom chip design to reduce dependency on Nvidia.

Dario Amodei· CEO & co-founder of Anthropic· Anthropic2d

Dario Amodei outlines Anthropic's policy positions on AI regulation, macroeconomics, and accelerating AI's positive impact.

Sergey Levine· Co-founder; Berkeley professor· Physical Intelligence · UC Berkeley2d

Origami/AutoEval enables autonomous 24/7 real-world robotic dexterity benchmarking from UC Berkeley/NVIDIA.

Dean W. Ball· AI policy author (Hyperdimensional)· Foundation for American Innovation2d

Post-Mythos, the US has an informal de facto AI licensing regime with no statutory basis, Ball argues.

Dylan Patel· Semiconductor / AI-infrastructure analyst· SemiAnalysis2d

SemiAnalysis launches STEEL: a dedicated teardown engineering & evaluation lab for semiconductor/AI hardware analysis.

Lisa Su· Chair & CEO of AMD· AMD2d

AMD Ryzen AI Max+ 395 runs a 235B parameter model locally, potentially replacing a $440/month cloud AI stack.

Neel Nanda· Mechanistic interpretability lead· Google DeepMind2d

Naive SFT filters for safety fail because they don't target the right model internals — Engels & Nanda explain the mechanistic reason why.

Junyang Lin· Ex-Qwen lead; founding embodied-AI startup· Independent2d

MTP with rejection sampling decouples entropy and acceptance rate to accelerate RL training for LLMs.

Brett Adcock· CEO & founder (humanoid robots)· Figure2d

Figure 03 robot successfully trained to walk down stairs, with training process footage shared by Brett Adcock.

Dean W. Ball· AI policy author (Hyperdimensional)· Foundation for American Innovation2d

US export controls are blocking European users from accessing Anthropic models, sparking industry backlash over policy contradictions.

Boris Cherny· Creator & head of Claude Code· Anthropic3d

Claude Code's head Boris Cherny argues cheaper models cost more overall due to retry overhead and compounding errors in agentic pipelines.

succeeded