惯性聚合 高效追踪和阅读你感兴趣的博客、新闻、科技资讯
阅读原文 在惯性聚合中打开

推荐订阅源

酷 壳 – CoolShell
酷 壳 – CoolShell
量子位
V2EX - 技术
V2EX - 技术
K
Kaspersky official blog
Know Your Adversary
Know Your Adversary
Hacker News - Newest:
Hacker News - Newest: "LLM"
cs.AI updates on arXiv.org
cs.AI updates on arXiv.org
I
Intezer
H
Heimdal Security Blog
奇客Solidot–传递最新科技情报
奇客Solidot–传递最新科技情报
S
SegmentFault 最新的问题
阮一峰的网络日志
阮一峰的网络日志
博客园_首页
博客园 - Franky
GbyAI
GbyAI
T
The Blog of Author Tim Ferriss
Recorded Future
Recorded Future
freeCodeCamp Programming Tutorials: Python, JavaScript, Git & More
Vercel News
Vercel News
Apple Machine Learning Research
Apple Machine Learning Research
The Hacker News
The Hacker News
T
Tenable Blog
Recent Commits to openclaw:main
Recent Commits to openclaw:main
雷峰网
雷峰网
WordPress大学
WordPress大学
Blog — PlanetScale
Blog — PlanetScale
Application and Cybersecurity Blog
Application and Cybersecurity Blog
Webroot Blog
Webroot Blog
L
LangChain Blog
C
Check Point Blog
N
News | PayPal Newsroom
L
LINUX DO - 热门话题
T
Tor Project blog
V
Visual Studio Blog
Microsoft Security Blog
Microsoft Security Blog
S
Security Affairs
Schneier on Security
Schneier on Security
Hacker News: Ask HN
Hacker News: Ask HN
Stack Overflow Blog
Stack Overflow Blog
Y
Y Combinator Blog
Cyber Security Advisories - MS-ISAC
Cyber Security Advisories - MS-ISAC
Security Latest
Security Latest
MyScale Blog
MyScale Blog
Cyberwarzone
Cyberwarzone
N
Netflix TechBlog - Medium
Scott Helme
Scott Helme
PCI Perspectives
PCI Perspectives
The Last Watchdog
The Last Watchdog
人人都是产品经理
人人都是产品经理
W
WeLiveSecurity

Hacker News - Newest: "AI"

AI can't read an investor deck AI as an attorney? Student uses ChatGPT, Gemini to sue UW over alleged racial discrimination Hacking MCP Servers in AI Systems – The Rug Pull: Tool Changes After Approval GitHub - MeepCastana/KubeezCut: Free Web based video editor GitHub - GenAI-Gurus/awesome-eu-ai-act: Curated tools, official sources, OSS, templates, and guides for EU AI Act compliance. Can AI judge journalism? A Thiel-backed startup says yes, even if it risks chilling whistleblowers Coming soon: 10 Things That Matter in AI Right Now DARPA built an AI to fact-check enemy weapons claims What explains heterogeneity in AI adoption? When AI Meets Muscle: Context-Aware Electrical Stimulation Promises a New Way to Guide Human Movements - Department of Computer Science AI Changed How We Build. It Did Not Change What Matters. Linux rules on using AI-generated code - Copilot is OK, but humans must take 'full responsibility for the… Meta spins up AI version of Mark Zuckerberg to engage with employees Code Mode: Let Your AI Write Programs, Not Just Call Tools | TanStack Blog GitHub - Delavalom/graft: Go framework for building AI agents. Type-safe tools, multi-provider (OpenAI, Anthropic, Gemini, Bedrock), zero vendor SDKs. India's TCS tops estimates, says new AI models did not dent services demand Gen Z's fading AI hype Strong feeling: we are in a folded AI reality GitHub - machinarii/total-recall-catalog: A reference catalog of latest knowledge retrieval, memory & RAG systems GitHub - mensfeld/code-on-incus: Give each AI agent its own isolated machine with root, Docker, and systemd. Active defense detects and stops threats automatically.. Quantization, LoRA, and the 8% Problem: Benchmarking Local LLMs for Production AI Iran war: We spoke to the man making Lego-style AI videos that experts say are powerful propaganda Powell, Bessent discussed Anthropic's Mythos AI cyber threat with major U.S. banks GitHub - immartian/bellamem: Persistent belief-graph memory for AI agents. Retrieves decisive context by importance — not recency, not RAG, not /compact. recursive-mode: The Repo-Native Operating System for AI Engineering After the attack on Sam Altman's home, will AI CEO's go on the offensive? The biggest advance in AI since the LLM Opus 4.6 vs GPT 5.4 One Prompt Unity World Generation Test “AI polls” are fake polls Client Challenge Can AI be a 'child of God'? Inside Anthropic's meeting with Christian leaders How to Switch AI Chatbots and Why You Might Want To GitHub - MattMessinger1/agentic_refund_guardrail: Safe refund policy layer for AI agents — Python + TypeScript. Same behavior, shared tests. Adam/papers/emergent_values_whitepaper.md at master · strangeadvancedmarketing/Adam Ask HN: How do you stop playing 20 questions with your AI coding tools How far can automation and AI support psychotherapy? - @theU GitHub - stagas/rtdiff: realtime git diff gui and AI-assisted commits A Mac Studio for Local AI — 6 Months Later A History of the Early Years of AI at the University of Edinburgh Why AI Coding Tools Still Feel Stuck on Localhost MSN AI Datacenters Are Becoming Strategic Targets twitter.com Penn Researchers Use AI to Surface Unreported GLP-1 Side Effects in Reddit Posts Show HN: MoodSense AI (ML and FastAPI and Gradio, Deployed on Hugging Face) Moodsense Ai - a Hugging Face Space by aman179102 AI models are terrible at betting on soccer—especially xAI Grok GitHub - xialeistudio/echoic GitHub - HimashaHerath/github-dev-wrapped: AI-powered weekly GitHub activity reports deployed to GitHub Pages GitHub - alejandrobalderas/claude-code-from-source: Architecture, patterns & internals of Anthropic's AI coding agent — reverse-engineered from source maps AI and Tech brief: Ireland ascendant GitHub - Titovilal/context0: Context0 - Never Surrender Training for a Marathon with an AI Coach: What Worked and What Didn't Cyber Pulse: Agentic Intel - Apps on Google Play I Built an AI PR Reviewer That Catches Bugs by Not Looking for Bugs Gen Z workers are so fearful AI will take their job they’re intentionally sabotaging their company’s AI rollout | Fortune How AI Is Reimagining the Game of Golf–For Both Players and Courses GitHub - nattergabriel/reseed: A CLI tool for managing and distributing agent skills across projects Is SVG the final frontier? My AI workflow evolved from prompts to a near-autonomous workflow MLSharp Help - 3DGS Viewer & Generator I put my cognitive field based AI's runtime on GitHub Is Numble the first AI-proof game? A3: Kubernetes for autonomous AI agent fleets | Emergent Principles Deepali Vyas ("The Elite Recruiter") GitHub - msmarkgu/RelayFreeLLM: A restful API designed to route user prompts to various AI model providers. Unionized ProPublica staff are on strike over AI, layoffs, and wages Unleashing the Advantage of Quantum AI We're heading for an AI-fueled 'dementia crisis,' brain scientist warns The AI-Assisted Breach of Mexico's Government Infrastructure [pdf] GitHub - stef41/lmscan: 🔍 Detect AI-generated text and fingerprint which LLM wrote it. Open-source GPTZero alternative. Zero dependencies, works offline. MSN GitHub - visionscaper/collabmem: Enabling long-term collaboration with Agentic AI - building up episodic and world model memory over time with in-context awareness We gave an AI a 3 year retail lease in SF and asked it to make a profit | Andon Labs AI Code is Hollowing Out Open Source, and Maintainers are Looking the Other Way What leaked "SteamGPT" files could mean for the PC gaming platform's use of AI AI is the boss at this retail store. What could go wrong? GitHub - Wuzu11517/agentic-proxy: Local proxy meant to help reduce With Drones, Geophysics and ArtificiaI Intelligence, Researchers Prepare to Do Battle Against Land Mines A Single Operator, Two AI Platforms, Nine Government Agencies: The Full Technical Report 在 Steam 上购买 FriedrichAI: Offline AI 立省 10% GitHub - inevolin/resume-cli: Hit Claude usage limits? Resume any AI coding session elsewhere. Switch tools at zero friction. GitHub - atripati/ark: AI Runtime Kernel — a context operating system for AI agents. Eliminates tool bloat, loads only what’s needed, and gives LLMs their reasoning space back. How to Build a Secure AI PR Reviewer with Claude, GitHub Actions, and JavaScript This Startup Wants You to Pay Up to Talk With AI Versions of Human Experts Intel Arc Pro B70 Brings 32GB VRAM to Local AI for $949 WordPress 7.0: The Good, the AI, and the Still Missing AI on the couch: Anthropic gives Claude 20 hours of psychiatry IatroBench: Pre-Registered Evidence of Iatrogenic Harm from AI Safety Measures AI Agents Know About Supabase. They Don't Always Use It Right. The history and future of AI at Google, with Sundar Pichai Inside an AI‑enabled device code phishing campaign How Meta Used AI to Map Tribal Knowledge in Large-Scale Data Pipelines AI for Systems: Using LLMs to Optimize Database Query Execution Forecasting the Economic Effects of AI Introducing Tinker: Play with AI, bring your ideas to life AI sheds light on an ancient gaming mystery People really hate AI but not as much as Iran—or Democrats | Fortune What is an AI Product Engineer? Phoebe Gates wants her $185 million AI startup to succeed with 'no ties to my privilege or my last name': 'I have a chip on my shoulder' | Fortune
“You can't vibe code scale”: What the AI hype gets wrong about software engineering
Eira May · 2026-05-18 · via Hacker News - Newest: "AI"

In some quarters, there’s a sense that AI has democratized software creation to the point where deep engineering expertise is becoming somehow optional. Vibe coding—in theory, at least—lets anyone describe what they want and watch AI build it. It’s not that there’s nothing to this. Prototyping is faster, junior engineering tasks are more accessible to folks without extensive formal training, and iteration cycles have been compressed. Again in theory, this should give developers more time and mental energy for complex, higher-order work.

But (you knew there would be a “but”), there’s a big difference between building software and running software at scale. AI hasn’t closed that gap. As Braze cofounder and CTO Jon Hyman and Stack Overflow CPTO Jody Bailey discussed on last week’s episode of Leaders of Code, the AI explosion actually makes senior engineering judgment more, not less, valuable. Because someone still has to own the consequences of what gets built and whether it can function at scale.

Let’s give the Pollyanna case its due. AI really has lowered the barrier to building functional software. Product managers can spin up interactive mockups to inform their thinking and communicate the functionality they need to engineers. Designers can resolve UX issues quickly and independently. Small teams that couldn’t afford to staff for moonshot projects can start building right now. All of these breakthroughs are real and worth noting. But as useful as vibe coding can be, it has a ceiling.

Building software and operating software are two different disciplines. That distinction gets lost in most conversations about AI and productivity, and it's where the vibe coding narrative starts to break down.

A prototype doesn't have users, traffic spikes, cascading failures, or data pipelines that degrade quietly under load before anyone notices. It doesn't have the accumulated weight of three years of architectural decisions, some brilliant and some regrettable, all of which constrain what you can do next. The prototype is the easy part. What comes after—running software reliably, at scale, for real customers with real expectations—is where engineering judgment becomes absolutely indispensable. As AI takes on more of the execution layer, the gap between what a model can generate and what it can understand becomes ever-more consequential.

Scale has specific, unforgiving demands. Distributed systems fail in ways that are rarely obvious and almost never reproducible in a local environment. Latency compounds across service boundaries in ways that don't show up until the stakes are uncomfortably high. A database schema decision made in week two becomes a migration nightmare in year three. An architectural pattern that works elegantly for ten thousand users can collapse like a house of cards at ten million. These kinds of challenges are the day-to-day reality of engineering teams running productive systems. Solving for them demands knowledge that’s deeply contextual—hard-won through experience and distinctly human.

Jon Hyman, CTO of Braze, put it plainly: "You can't vibe code scale... Being able to run that at high complexity, high scale, high down use cases is something that requires a deep understanding of what you're doing, the business problem that you're solving, and then how all the systems work together."

AI models are getting remarkably good at reading code, but that’s not the same as understanding a system. Even with a million-token context window, a model doesn't contain your business processes, customer use cases, organizational constraints, or the reasoning behind decisions made long ago. AI sees the what; it rarely has access to the why. (That’s where Stack Internal comes in!)

A new technology that makes everyone more productive? Of course some people are going to look at it as a cost-cutting opportunity. If your engineers can do twice as much, you need half as many engineers. Right?

Not so fast. Think about how the competitive advantage actually works. AI productivity gains aren't proprietary. Every company in your market got access to the same models, the same tools, and roughly the same multiplier on engineering output at roughly the same time. So if you use that multiplier to reduce headcount and hold output steady, you haven't improved your competitive position. You've just spent less money to stay in exactly the same place. Meanwhile, your competitors are using their multiplier to build more and ship it faster.

As Jon framed it: "Everyone instantly, globally, got this stepwise increase in productivity... if you had 100 engineers and all of a sudden now you have the output of, let's call it 180 engineers, is the first thing you do, is it to go and build Salesforce? Because all of your competitors also went from having 100 engineers of output to 180 engineers of output. And they're working on their roadmaps."

In fast-moving markets, the ability to ship meaningful features consistently—and faster than the competition—is itself a differentiator. Customers notice; deals are won and lost on it. Cutting engineering capacity at the exact moment that execution velocity becomes more achievable and more competitively important is a strange way to deploy a productivity windfall.

AI didn't give everyone an advantage; it raised the floor for everybody. What you build on top of that floor—how ambitiously you use the newly available capacity, how clearly you prioritize your roadmap, how well your engineering culture is positioned to move fast without breaking things—is still a human decision. That’s not (just) a limitation of the technology, I’d argue: It’s a good touch-grass reminder that strategy isn’t technology’s job in the first place.

The gist is that AI doesn’t make senior engineers redundant. But it can—and probably already has—changed what they spend their time on. If you’re one of those engineers, you probably spend less time on boilerplate, scaffolding, and repetitive work, which we hope gives you more time for system design, architectural decision-making, and other kinds of context-dependent problems for which your human brain is an unconditional requirement.

This raises the ceiling on what a small but experienced engineering team can do. As autonomous agents take over more of the execution involved in engineering, a new responsibility emerges that sits squarely with senior engineers: codifying what they know.

Agents can only work effectively within the boundaries of what they've been given. Right now, much of that knowledge lives exclusively between experienced engineers' ears. Getting it out of those heads and into a form that agents can actually use will be among the more consequential engineering tasks of the next few years—because it determines how much of your AI investment actually compounds over time.

If you understand that AI can increase the value of human judgment while absorbing low-complexity execution, then the implications for how you should manage your team, assess your budget, and set your expectations are significant. Here are a few places to start.

AI should be handling work that doesn't require the human judgment of senior engineers. If that’s not the case, you have a leadership conversation, not a technology problem. Engineering managers should be actively identifying the categories of work that can be offloaded (e.g., boilerplate, routine testing, basic debugging) and setting a clear expectation for where engineers should spend their time going forward. The bar for what a team can deliver in a sprint has risen.

This is a low-hanging fruit of a metric: easy to focus on, but probably the least useful one for a team with genuine roadmap ambitions. Some better questions are:

  • What can we build now that we couldn't before?
  • How has our cycle time on features changed?
  • Do your engineers report less burnout and frustration at work? Are they excited about the new things they can build?
  • Are we resolving more UX debt, shipping more experiments, or responding faster to customer feedback?
  • What's the ratio of inference cost to meaningful output? Is it improving?

This is the one most organizations are behind on and will feel the consequences of soonest. If your agents are producing generic, pattern-inconsistent output, it's usually because they don't have access to the context your senior engineers carry in their heads. Fixing that means documenting coding standards, testing expectations, and architectural patterns in a form models can actually use. It entails building a process for capturing decisions and the reasoning behind them (not just what was decided, but why). For this work to happen, leadership must assign ownership to senior engineers, rather than waiting for it to happen organically.

At the risk of stating the obvious, running models at scale across an engineering organization gets expensive fast. Leaders who aren't already tracking inference spend per engineer, not to mention thinking about what efficient usage actually looks like, are likely to face an uncomfortable budget conversation in the next planning cycle. You can ahead of it by building cost-awareness into how AI usage is discussed and measured on the team.

Letting engineers explore different tools and workflows is how organizations learn what actually works. But there's a point at which the variation becomes its own inefficiency. Inconsistent patterns, duplicated effort, and agents operating without shared context are all significant inefficiencies that arise when you let everybody play in the sandbox at once. You should be moving from experimental to intentional: keeping what's working, discarding what isn't, and building the shared infrastructure that lets the whole team benefit.

The vibe coding narrative is seductive for a simple reason: it focuses on inputs. Look how easy it is to generate code. Look how fast a prototype comes together. Look how much a single engineer can produce in an afternoon with the right model. These things matter. At the same time, we have to recognize that they're measurements of what goes in, not what comes out. What comes out is still software that has to run reliably, scale under pressure, serve real customers, and hold up against competitors who are moving just as fast as you are.

The tools have gotten better; the execution layer has gotten faster and cheaper. But the human judgment required to build systems that actually work at scale, over time, in intensely competitive markets, is as important as ever. Given the volume of output that human judgement now has to oversee, you might even say human judgment is more valuable than ever.

The question for engineering leaders isn’t, “How many engineers do we need?” It’s, “What kind of engineering culture do we want to build?” Do you want an engineering culture that uses AI to do the same stuff, only faster and cheaper? Or do you want engineers who use AI to do things that weren’t previously possible?