惯性聚合 高效追踪和阅读你感兴趣的博客、新闻、科技资讯
阅读原文 在惯性聚合中打开

推荐订阅源

C
Cybersecurity and Infrastructure Security Agency CISA
N
News and Events Feed by Topic
S
Securelist
Threat Intelligence Blog | Flashpoint
Threat Intelligence Blog | Flashpoint
Spread Privacy
Spread Privacy
T
Threat Research - Cisco Blogs
T
Tor Project blog
C
Cyber Attacks, Cyber Crime and Cyber Security
K
Kaspersky official blog
L
LINUX DO - 热门话题
T
The Exploit Database - CXSecurity.com
S
Schneier on Security
A
Arctic Wolf
Security Latest
Security Latest
T
Threatpost
P
Palo Alto Networks Blog
Simon Willison's Weblog
Simon Willison's Weblog
AWS News Blog
AWS News Blog
Cyberwarzone
Cyberwarzone
L
Lohrmann on Cybersecurity
P
Privacy International News Feed
V
Vulnerabilities – Threatpost
D
Darknet – Hacking Tools, Hacker News & Cyber Security
Cisco Talos Blog
Cisco Talos Blog
C
CXSECURITY Database RSS Feed - CXSecurity.com
G
GRAHAM CLULEY
The Hacker News
The Hacker News
C
CERT Recently Published Vulnerability Notes
Know Your Adversary
Know Your Adversary
I
Intezer
Scott Helme
Scott Helme
T
Tenable Blog
NISL@THU
NISL@THU
cs.CL updates on arXiv.org
cs.CL updates on arXiv.org
C
Cisco Blogs
N
News and Events Feed by Topic
P
Proofpoint News Feed
P
Privacy & Cybersecurity Law Blog
Project Zero
Project Zero
Latest news
Latest news
Hacker News: Ask HN
Hacker News: Ask HN
Recent Commits to openclaw:main
Recent Commits to openclaw:main
Forbes - Security
Forbes - Security
Security Archives - TechRepublic
Security Archives - TechRepublic
AI
AI
S
Security Affairs
cs.AI updates on arXiv.org
cs.AI updates on arXiv.org
D
Docker
P
Proofpoint News Feed
博客园 - Franky

Hacker News - Newest: "AI"

AI can't read an investor deck AI as an attorney? Student uses ChatGPT, Gemini to sue UW over alleged racial discrimination Hacking MCP Servers in AI Systems – The Rug Pull: Tool Changes After Approval GitHub - MeepCastana/KubeezCut: Free Web based video editor GitHub - GenAI-Gurus/awesome-eu-ai-act: Curated tools, official sources, OSS, templates, and guides for EU AI Act compliance. Can AI judge journalism? A Thiel-backed startup says yes, even if it risks chilling whistleblowers Coming soon: 10 Things That Matter in AI Right Now DARPA built an AI to fact-check enemy weapons claims What explains heterogeneity in AI adoption? When AI Meets Muscle: Context-Aware Electrical Stimulation Promises a New Way to Guide Human Movements - Department of Computer Science AI Changed How We Build. It Did Not Change What Matters. Linux rules on using AI-generated code - Copilot is OK, but humans must take 'full responsibility for the… Meta spins up AI version of Mark Zuckerberg to engage with employees Code Mode: Let Your AI Write Programs, Not Just Call Tools | TanStack Blog GitHub - Delavalom/graft: Go framework for building AI agents. Type-safe tools, multi-provider (OpenAI, Anthropic, Gemini, Bedrock), zero vendor SDKs. India's TCS tops estimates, says new AI models did not dent services demand Gen Z's fading AI hype Strong feeling: we are in a folded AI reality GitHub - machinarii/total-recall-catalog: A reference catalog of latest knowledge retrieval, memory & RAG systems GitHub - mensfeld/code-on-incus: Give each AI agent its own isolated machine with root, Docker, and systemd. Active defense detects and stops threats automatically.. Quantization, LoRA, and the 8% Problem: Benchmarking Local LLMs for Production AI Iran war: We spoke to the man making Lego-style AI videos that experts say are powerful propaganda Powell, Bessent discussed Anthropic's Mythos AI cyber threat with major U.S. banks GitHub - immartian/bellamem: Persistent belief-graph memory for AI agents. Retrieves decisive context by importance — not recency, not RAG, not /compact. recursive-mode: The Repo-Native Operating System for AI Engineering After the attack on Sam Altman's home, will AI CEO's go on the offensive? The biggest advance in AI since the LLM Opus 4.6 vs GPT 5.4 One Prompt Unity World Generation Test “AI polls” are fake polls Client Challenge Can AI be a 'child of God'? Inside Anthropic's meeting with Christian leaders How to Switch AI Chatbots and Why You Might Want To GitHub - MattMessinger1/agentic_refund_guardrail: Safe refund policy layer for AI agents — Python + TypeScript. Same behavior, shared tests. Adam/papers/emergent_values_whitepaper.md at master · strangeadvancedmarketing/Adam Ask HN: How do you stop playing 20 questions with your AI coding tools How far can automation and AI support psychotherapy? - @theU GitHub - stagas/rtdiff: realtime git diff gui and AI-assisted commits A Mac Studio for Local AI — 6 Months Later A History of the Early Years of AI at the University of Edinburgh Why AI Coding Tools Still Feel Stuck on Localhost MSN AI Datacenters Are Becoming Strategic Targets twitter.com Penn Researchers Use AI to Surface Unreported GLP-1 Side Effects in Reddit Posts Show HN: MoodSense AI (ML and FastAPI and Gradio, Deployed on Hugging Face) Moodsense Ai - a Hugging Face Space by aman179102 AI models are terrible at betting on soccer—especially xAI Grok GitHub - xialeistudio/echoic GitHub - HimashaHerath/github-dev-wrapped: AI-powered weekly GitHub activity reports deployed to GitHub Pages GitHub - alejandrobalderas/claude-code-from-source: Architecture, patterns & internals of Anthropic's AI coding agent — reverse-engineered from source maps AI and Tech brief: Ireland ascendant GitHub - Titovilal/context0: Context0 - Never Surrender Training for a Marathon with an AI Coach: What Worked and What Didn't Cyber Pulse: Agentic Intel - Apps on Google Play I Built an AI PR Reviewer That Catches Bugs by Not Looking for Bugs Gen Z workers are so fearful AI will take their job they’re intentionally sabotaging their company’s AI rollout | Fortune How AI Is Reimagining the Game of Golf–For Both Players and Courses GitHub - nattergabriel/reseed: A CLI tool for managing and distributing agent skills across projects Is SVG the final frontier? My AI workflow evolved from prompts to a near-autonomous workflow MLSharp Help - 3DGS Viewer & Generator I put my cognitive field based AI's runtime on GitHub Is Numble the first AI-proof game? A3: Kubernetes for autonomous AI agent fleets | Emergent Principles Deepali Vyas ("The Elite Recruiter") GitHub - msmarkgu/RelayFreeLLM: A restful API designed to route user prompts to various AI model providers. Unionized ProPublica staff are on strike over AI, layoffs, and wages Unleashing the Advantage of Quantum AI We're heading for an AI-fueled 'dementia crisis,' brain scientist warns The AI-Assisted Breach of Mexico's Government Infrastructure [pdf] GitHub - stef41/lmscan: 🔍 Detect AI-generated text and fingerprint which LLM wrote it. Open-source GPTZero alternative. Zero dependencies, works offline. MSN GitHub - visionscaper/collabmem: Enabling long-term collaboration with Agentic AI - building up episodic and world model memory over time with in-context awareness We gave an AI a 3 year retail lease in SF and asked it to make a profit | Andon Labs AI Code is Hollowing Out Open Source, and Maintainers are Looking the Other Way What leaked "SteamGPT" files could mean for the PC gaming platform's use of AI AI is the boss at this retail store. What could go wrong? GitHub - Wuzu11517/agentic-proxy: Local proxy meant to help reduce With Drones, Geophysics and ArtificiaI Intelligence, Researchers Prepare to Do Battle Against Land Mines A Single Operator, Two AI Platforms, Nine Government Agencies: The Full Technical Report 在 Steam 上购买 FriedrichAI: Offline AI 立省 10% GitHub - inevolin/resume-cli: Hit Claude usage limits? Resume any AI coding session elsewhere. Switch tools at zero friction. GitHub - atripati/ark: AI Runtime Kernel — a context operating system for AI agents. Eliminates tool bloat, loads only what’s needed, and gives LLMs their reasoning space back. How to Build a Secure AI PR Reviewer with Claude, GitHub Actions, and JavaScript This Startup Wants You to Pay Up to Talk With AI Versions of Human Experts Intel Arc Pro B70 Brings 32GB VRAM to Local AI for $949 WordPress 7.0: The Good, the AI, and the Still Missing AI on the couch: Anthropic gives Claude 20 hours of psychiatry IatroBench: Pre-Registered Evidence of Iatrogenic Harm from AI Safety Measures AI Agents Know About Supabase. They Don't Always Use It Right. The history and future of AI at Google, with Sundar Pichai Inside an AI‑enabled device code phishing campaign How Meta Used AI to Map Tribal Knowledge in Large-Scale Data Pipelines AI for Systems: Using LLMs to Optimize Database Query Execution Forecasting the Economic Effects of AI Introducing Tinker: Play with AI, bring your ideas to life AI sheds light on an ancient gaming mystery People really hate AI but not as much as Iran—or Democrats | Fortune What is an AI Product Engineer? Phoebe Gates wants her $185 million AI startup to succeed with 'no ties to my privilege or my last name': 'I have a chip on my shoulder' | Fortune
Navigating AI with paper maps
adamsurg · 2026-05-22 · via Hacker News - Newest: "AI"

The advice economy around LLMs is such a strange thing to watch. Somewhere on LinkedIn this morning, someone with a job title that didn't exist eighteen months ago is sharing their hard-won secrets for better LLM output with forty thousand followers. They're not wrong, exactly. The techniques worked — for the model they had, when they found them. The trouble is that "what works" has a shelf life measured in model releases. The post keeps circulating long after it's relevant; the conference talk is accepted months before anyone delivers it.

Even the vendors' own training material runs a release or two behind their own models. Anthropic's official course videos were, as I wrote this, recorded on Claude Opus 3.x — which, in a fast-moving field, might as well be on punched cards. The fundamentals don't shift, but the models do and quickly.

This is the state of the AI conversation in 2026: a great deal of energy spent optimising out last year's bottlenecks, and very little spent on durable solutions. What's missing is AI solution architecture — knowing how to structure systems so they're durable, do the job and don't eat the budget. That's not in the courses, because it doesn't fit on a slide.

Cutting the steak up

Pick almost any prompt-engineering case study and run it through the only test that matters: are a couple of the right tools better than a swiss army knife?

The hackneyed meal-plan is an example. A user describes their dietary needs, peculiarities, and goals; the system produces a week of meals. The standard treatment is to grind out a multi-thousand-token prompt, build a training set, define a fitness function, and tune until the output stops embarrassing you. Days of work. A six-figure salary line item, PowerPoints, applause.

But what is it actually doing? Three things. Parsing misspelled text into structured parameters, solving a constraint-satisfaction problem and serving up the result as digestible verbiage.

The first part is a good candidate for an ML solution that even the browser can run; the selection part is good old-fashioned programming; and the last is where LLMs earn their keep. Doing all three with one prompt is easier at first flush, but it's not the right solution for much of the problem and doesn't scale.

The decomposition is just engineering. It only feels like a revelation because so much of the surrounding conversation hasn't caught up with it. And when you push back on it even a little, the very model that was helping with the prompt is more than capable of doing this too — you can even treat one stage as a working prototype for the next.

And you can just ask it

There's a content economy of people explaining, in 23-minute videos fronted by startled faces, how to use LLMs well. There are no likes to be had explaining that the model itself is better at this than they are.

Seriously, just ask the model, or several and compare. The advice will be current, specific to your stack, ready to answer follow-ups and doesn't ask you to smash the like button.

Renting someone else's brain

Effective prompt engineering is iterating over a string that seeds a particular version of a generalised model running on someone else's neural network until you get plausible answers to enough of your scenarios that you're happy to write the answer down in ink.

When the model changes, no problem — just repeat the whole process again, and hope you can do it before anyone notices the old prompt is broken. It's less of a system, more of a wish.

There are places it's the right call — prototypes, internal tools. But in production, where reliability and cost matter, the architecture is where humans come to rescue the robots, with or without lasers.

Solving the previous problem

Then there's the whole RAG apparatus. Vector databases, chunking strategies, reranking pipelines, embedding-model selection, retrieval evaluation harnesses. Conference talks, certifications, consultancies. This was a sensible response to a real constraint circa 2023: context windows were small, models forgot things, and you needed clever plumbing to get the right facts in front of the model at the right moment.

The constraint has eased. Context windows that would have seemed absurd two years ago are now routine, and caching takes the sting out of paying for them. What's dying isn't retrieval — it's the assumption that you must do the retrieving up front. Increasingly the better pattern is to give the model the tools to fetch its own context: a search call, a database query, a file read, behind a caching layer so it's cheap to do repeatedly. The model decides what it needs and goes and gets it, rather than you guessing in advance and stuffing a prompt.

That's the shift. Naive embedding-similarity over chunked PDFs is on the way out; the requirement underneath — the right context, in the right shape, at the right moment — is as central as ever. It's just moved from a pipeline you build to a capability you expose. Which is a tooling problem, and a much more interesting one.

Workflowing away

When models were weaker, you got better results by choreographing them carefully. A whole generation of libraries grew up to express that choreography, and some of them are genuinely well made. Capable models now build that workflow themselves, often better than the hand-rolled version, because they can see the whole problem at once.

This is the part that gets misread: it does not mean turning a capable model loose on production and heading to the pub. It means the workflow itself — the decomposition, the sequencing, the choice of approach — is no longer the thing worth obsessing over. The guardrails are more important: what tools the model can call, with which arguments, at what permission scope, validation, audit trail, halting conditions, and with what escalation when it all goes sideways.

The old approach scripted the workflow in detail and treated the guardrails as an afterthought — "and then the LLM does the right thing." The newer approach lets the model handle the workflow and spends the engineering effort on the edges of what it's allowed to do. The choreography is free now. The fence is where the work is.

Where’s the action

So: if you're an engineer wondering where to spend your finite professional-development hours, the answer is increasingly boring and increasingly right.

Learn the APIs properly, a working grasp of streaming, tool use, batching, caching, error modes, and cost behaviour. The people who understand the monthly bill keep their budgets.

Learn MCP: the right tools, schemas, and permission so that the model does less — cheaper and harder to break. It's the durable surface where capable models meet your systems, and it doesn't need a manicure every time a new model lands.

Learn hooks and interception. Where you validate, check, redirect, halt. The plumbing of guardrails. Unglamorous. Compounds.

Learn evaluation that matters to your problem: does your system fail safely on the inputs you actually care about? No leaderboard is trying to answer that for you.

Learn integration patterns that survive a model change. Timeouts, retries, idempotency, audit trails, observability — the dull discipline that keeps every other distributed system upright.

It’s been emotional

None of this is trending, and it's firmly a minority view. But read the people actually building, or just ask the model what works, and the same thing keeps surfacing: once the hype burns off, the new world looks a lot like the old one. The fundamentals didn't move. What changed is that we can build faster, for less — and hand the robots the dull work, which leaves the interesting part to us. That was always the good bit.

No posts