惯性聚合 高效追踪和阅读你感兴趣的博客、新闻、科技资讯
阅读原文 在惯性聚合中打开

推荐订阅源

WordPress大学
WordPress大学
Microsoft Azure Blog
Microsoft Azure Blog
aimingoo的专栏
aimingoo的专栏
Vercel News
Vercel News
U
Unit 42
L
LangChain Blog
J
Java Code Geeks
钛媒体:引领未来商业与生活新知
钛媒体:引领未来商业与生活新知
The Cloudflare Blog
F
Fortinet All Blogs
小众软件
小众软件
I
InfoQ
P
Proofpoint News Feed
D
DataBreaches.Net
Martin Fowler
Martin Fowler
H
Help Net Security
T
Tailwind CSS Blog
N
Netflix TechBlog - Medium
有赞技术团队
有赞技术团队
Y
Y Combinator Blog
Recent Announcements
Recent Announcements
B
Blog RSS Feed
酷 壳 – CoolShell
酷 壳 – CoolShell
B
Blog

Hacker News: Show HN

PurrrrrFocus: Pomodoro Timer App - App Store Workflow Engine — Multi-Step Orchestration for Bun RapidPhoto: Pro Photo Editor App - App Store GitHub - DheerG/swarms: Achieve extraordinary results with claude code across a variety of tasks SPICE simulation → oscilloscope → verification with Claude Code — Lucas Gerads Show HN: VCoding – A 5 MB native Windows IDE with no dynamic dependencies Show HN: LLMs don't hallucinate because they're bad at math, it's the format GitHub - Agent-FM/agentfm-core: AgentFM is a peer-to-peer network that turns everyday computers into a decentralized AI supercomputer. AgentFM lets you run massive AI workloads directly across a global mesh of idle CPUs and GPUs. Show HN: Tracking Top US Science Olympiad Alumni over Last 25 Years GitHub - Potarix/agent-hub: One place to talk to all your agents Show HN: Runtime security for AI agents(injection,tool abuse, data exfiltration) GitHub - dubeyKartikay/lazyspotify: Terminal Spotify client for macOS and Linux GitHub - the-banana-tool/king-louie: Easy to use GUI Personal AI Assistant. Win/Linux/Mac. Show HN I made my vacation rental bookable by AI agents–no Airbnb, 0% commission GitHub - basteez/jsf-autoreload: maven plugin to enable hot reload on jsf projects uvm32/hosts/host-gdbstub at main · ringtailsoftware/uvm32 GitHub - labsai/EDDI: Config-driven engine that turns JSON into production-grade AI agents. Multi-agent orchestration, 12+ LLM providers, MCP/A2A protocols, RAG, persistent memory, and enterprise compliance (EU AI Act, GDPR, HIPAA). Built on Quarkus. GitHub - glitchnsec/fortyone-oss: AI Executive Assistant Platform Quickstart | Alien GitHub - muxshed/shed: One stream in, or many. Every destination, simultaneously. No cloud middleman, no per-channel fees, no limits. GitHub - ocrbase-hq/ocrbase: 📄 PDF/IMG ->.MD/JSON Document OCR API for PaddleOCR and GLMOCR. Self-hostable. GitHub - impactjo/home-memory: MCP server that lets your AI assistant remember everything about your home. GitHub - Sets88/dbcls: DbCls is a powerful terminal database client that supports various databases GitHub - neptun2000/heor-agent-mcp GitHub - SeanFDZ/macmind: Single-layer transformer in HyperTalk for the classic Macintosh RollQuation: Math Puzzles - Apps on Google Play GitHub - dropbox/witchcraft Show HN: Agent-cache – Multi-tier LLM/tool/session caching for Valkey and Redis GitHub - opentalon/opentalon: OpenTalon is an open-source platform built from the ground up in Go as a robust alternative to OpenClaw LinkedIn™ 职位抓取工具 - Chrome 应用商店
Show HN: When the LLM Accidentally
lucid-dev · 2026-05-08 · via Hacker News: Show HN

When the LLM accidentally... outputs some high-level abstraction of "thinking" into it's direct response. See text block at end.

What else have you seen the LLM accidentally do?

This isn't jailbreaking on my end - just normal use - with GPT5.4 in this case, reasoning and verbosity both set to "high".. (on /completions)

Point is, this block (plus lots more) is at the top of the response - then the "actual output" or response later on... but it's kind of cool, it helps me see how much time is being spent inferring things improperly and how I could change my instructions..

The real question is, would there be any reliable way of obtaining this same sort of internal "data" in a reliable way about the mid-points of generation?? And what is really actually occurring here?

It comes up rarely - maybe I've seen it twice in thousands of prompts with GPT5.4 and input context windows ranging from 150k-200k

> Need perhaps use terminal/filesearch? We can use add_document only; actual contents will be in next prompt, so in this response we probably can't edit them since not visible yet. But maybe we can leverage terminal instruction to inspect them now? However code changes based on terminal output may violate DMP source-of-truth? DMP said docs are authoritative. Safer to first just curate DMP and update task/checklist to note next sub-slice inspect registry/persistence insertion points. But user asked after plan do current tasks etc and code can continue. Since now auto loop wants continue task. Yet maybe one response can add docs and update task with exact next coding slice. But autonomous checklist says no routine pause. However it's acceptable to do "narrow codebase re-entry" if necessary. Current DMP visible docs maybe insufficient. So response can be instruction blocks only + doc updates? But need code/doc changes to current task? We likely should update current task with inspection step? But it already says immediate next move. Could update checklist/current task not necessary if same. Maybe just DMP add/hide and maybe filesearch/terminal to inspect? Since abilities available and can be in same response. Terminal can inspect file content though maybe not after add_document. But likely easier to add docs and maybe use terminal_execute cat/sed these files? Yet next prompt will have docs added and also terminal output likely in re-prompt. Could then implement. Because machine will process instruction blocks. Should we use terminal_execute to inspect hidden files now without waiting? That output will come in next prompt. Good. We can combine with add_document/hide and maybe filesearch search relevant terms. Great.

> Need summary to user. No code changes directly because not enough source files. So use instruction and maybe maybe update current task? Could perhaps not necessary but maybe mention next move. But automation expects code changes maybe not always. However from step 2 "always update ... current task doc and 1-2 codebase documents" but if doing discovery maybe maybe not. Yet current task already updated. Maybe okay. But maybe should update current task noting "first runtime registry seam inspection on core/persistence stores". However whole task is already adequate. Since not major change. We can leave.

> Need construct instruction blocks according to developer. For DMP use instruction blocks as lucid-meta. Need maintain special title. Use one block toggle_hide maybe mixed? only disk docs. Another add_document for 8 files. Another maybe filesearch_search? Another terminal_execute. Ensure correct JSON. Shell cwd /mnt/c... maybe relative to project root. Use sed -n maybe command allowed with restricted? likely yes. We can do python? better bash sed/cat fine. Need perhaps use filesearch_search for "IPaneCapability" nonexistent. Search "PaneDefinitionStore" and "SeededPaneCatalog". Maybe terminal_execute with rg is maybe restricted? maybe okay, but filesearch_search better for search. Use add_document not too many maybe 6 docs. Let's think exact hidden files needed: