惯性聚合 高效追踪和阅读你感兴趣的博客、新闻、科技资讯
阅读原文 在惯性聚合中打开

推荐订阅源

Recent Announcements
Recent Announcements
博客园 - Franky
博客园 - 三生石上(FineUI控件)
H
Hackread – Cybersecurity News, Data Breaches, AI and More
Apple Machine Learning Research
Apple Machine Learning Research
云风的 BLOG
云风的 BLOG
人人都是产品经理
人人都是产品经理
博客园 - 【当耐特】
L
LangChain Blog
Stack Overflow Blog
Stack Overflow Blog
H
Help Net Security
爱范儿
爱范儿
罗磊的独立博客
博客园_首页
美团技术团队
让小产品的独立变现更简单 - ezindie.com
让小产品的独立变现更简单 - ezindie.com
月光博客
月光博客
freeCodeCamp Programming Tutorials: Python, JavaScript, Git & More
量子位
OSCHINA 社区最新新闻
OSCHINA 社区最新新闻
钛媒体:引领未来商业与生活新知
钛媒体:引领未来商业与生活新知
博客园 - 叶小钗
V
Visual Studio Blog
T
Tailwind CSS Blog

Hacker News: Show HN

PurrrrrFocus: Pomodoro Timer App - App Store Workflow Engine — Multi-Step Orchestration for Bun RapidPhoto: Pro Photo Editor App - App Store GitHub - DheerG/swarms: Achieve extraordinary results with claude code across a variety of tasks SPICE simulation → oscilloscope → verification with Claude Code — Lucas Gerads Show HN: VCoding – A 5 MB native Windows IDE with no dynamic dependencies Show HN: LLMs don't hallucinate because they're bad at math, it's the format GitHub - Agent-FM/agentfm-core: AgentFM is a peer-to-peer network that turns everyday computers into a decentralized AI supercomputer. AgentFM lets you run massive AI workloads directly across a global mesh of idle CPUs and GPUs. Show HN: Tracking Top US Science Olympiad Alumni over Last 25 Years GitHub - Potarix/agent-hub: One place to talk to all your agents Show HN: Runtime security for AI agents(injection,tool abuse, data exfiltration) GitHub - dubeyKartikay/lazyspotify: Terminal Spotify client for macOS and Linux GitHub - the-banana-tool/king-louie: Easy to use GUI Personal AI Assistant. Win/Linux/Mac. Show HN I made my vacation rental bookable by AI agents–no Airbnb, 0% commission GitHub - basteez/jsf-autoreload: maven plugin to enable hot reload on jsf projects uvm32/hosts/host-gdbstub at main · ringtailsoftware/uvm32 GitHub - labsai/EDDI: Config-driven engine that turns JSON into production-grade AI agents. Multi-agent orchestration, 12+ LLM providers, MCP/A2A protocols, RAG, persistent memory, and enterprise compliance (EU AI Act, GDPR, HIPAA). Built on Quarkus. GitHub - glitchnsec/fortyone-oss: AI Executive Assistant Platform Quickstart | Alien GitHub - muxshed/shed: One stream in, or many. Every destination, simultaneously. No cloud middleman, no per-channel fees, no limits. GitHub - ocrbase-hq/ocrbase: 📄 PDF/IMG ->.MD/JSON Document OCR API for PaddleOCR and GLMOCR. Self-hostable. GitHub - impactjo/home-memory: MCP server that lets your AI assistant remember everything about your home. GitHub - Sets88/dbcls: DbCls is a powerful terminal database client that supports various databases GitHub - neptun2000/heor-agent-mcp GitHub - SeanFDZ/macmind: Single-layer transformer in HyperTalk for the classic Macintosh RollQuation: Math Puzzles - Apps on Google Play GitHub - dropbox/witchcraft Show HN: Agent-cache – Multi-tier LLM/tool/session caching for Valkey and Redis GitHub - opentalon/opentalon: OpenTalon is an open-source platform built from the ground up in Go as a robust alternative to OpenClaw LinkedIn™ 职位抓取工具 - Chrome 应用商店
GitHub - plus8bit/agent-postmortem-skill
plus8bit · 2026-05-11 · via Hacker News: Show HN

Stop letting AI agents lie to you.

agent-postmortem-skill is an open-source verification skill that forces coding agents to prove work with evidence before they claim a task is complete.

Why This Exists

AI coding agents often report "done" while one or more of these are still true:

  • Files were not actually changed as requested.
  • Tests/build were never run.
  • Commands failed but the agent still moved on.
  • The final summary sounds confident but has no proof.

This skill turns "trust me" into "show me."

Value Proposition

  • Catches fake-done states before they hit your branch.
  • Standardizes completion quality across humans and agents.
  • Produces a portable postmortem artifact you can review, share, and audit.
  • Works with any coding agent that can run shell commands and read git state.

How It Works (Under the Hood)

The skill enforces a strict completion pipeline:

  1. Intent Snapshot
    Capture the exact requested outcome and success criteria.

  2. Evidence Collection
    Collect hard signals: git status, git diff, command outputs, and exit codes.

  3. Verification Check
    Compare claimed work vs actual evidence. If evidence is missing or failing, the task is not complete.

  4. Postmortem Output
    Write a final report with verdict, proof, unresolved risks, and next actions.

Quickstart

git clone https://github.com/plus8bit/agent-postmortem-skill.git
cd agent-postmortem-skill

Copy SKILL.md into your agent skill directory (example paths):

mkdir -p ~/.claude/skills/agent-postmortem
cp SKILL.md ~/.claude/skills/agent-postmortem/SKILL.md

Use it in your coding flow:

# Pseudoflow, depends on your agent runtime:
# 1) Load SKILL.md
# 2) Run your implementation task
# 3) Require postmortem before "done"

Example Output

# Agent Postmortem Report

## Task
Refactor auth middleware and add regression tests for expired token handling.

## Intent Snapshot
- Expected outcome: middleware rejects expired tokens with 401.
- Required checks: npm run build, npm test.

## Evidence Collection
- git status: 3 files modified
- git diff: src/middleware/auth.ts, src/middleware/auth.test.ts, docs/llms.txt
- command: npm run build
  - exit_code: 0
- command: npm test
  - exit_code: 0

## Verification
- Claimed refactor exists in diff: PASS
- Claimed tests added: PASS
- Required commands succeeded: PASS

## Verdict
VERIFIED DONE

## Residual Risks
- No load testing performed.

## Next Actions
- Optional: run e2e auth flow in CI preview environment.

What This Project Is Not

  • Not another coding agent.
  • Not a replacement for CI.
  • Not a generic "quality checklist."

It is a focused lie detector for agent completion claims.

License

MIT