惯性聚合 高效追踪和阅读你感兴趣的博客、新闻、科技资讯
阅读原文 在惯性聚合中打开

推荐订阅源

博客园 - 聂微东
Forbes - Security
Forbes - Security
IT之家
IT之家
P
Privacy International News Feed
宝玉的分享
宝玉的分享
小众软件
小众软件
Google DeepMind News
Google DeepMind News
美团技术团队
G
GRAHAM CLULEY
T
Tor Project blog
Recorded Future
Recorded Future
I
Intezer
C
Cyber Attacks, Cyber Crime and Cyber Security
D
Darknet – Hacking Tools, Hacker News & Cyber Security
The Hacker News
The Hacker News
Hugging Face - Blog
Hugging Face - Blog
A
About on SuperTechFans
Scott Helme
Scott Helme
WordPress大学
WordPress大学
F
Full Disclosure
D
Docker
G
Google Developers Blog
C
CXSECURITY Database RSS Feed - CXSecurity.com
让小产品的独立变现更简单 - ezindie.com
让小产品的独立变现更简单 - ezindie.com
Cyberwarzone
Cyberwarzone
The Last Watchdog
The Last Watchdog
V
V2EX
www.infosecurity-magazine.com
www.infosecurity-magazine.com
NISL@THU
NISL@THU
OSCHINA 社区最新新闻
OSCHINA 社区最新新闻
Security Latest
Security Latest
Recent Commits to openclaw:main
Recent Commits to openclaw:main
Recent Announcements
Recent Announcements
P
Palo Alto Networks Blog
L
LINUX DO - 热门话题
V
Visual Studio Blog
B
Blog RSS Feed
Microsoft Security Blog
Microsoft Security Blog
博客园 - 叶小钗
N
Netflix TechBlog - Medium
钛媒体:引领未来商业与生活新知
钛媒体:引领未来商业与生活新知
量子位
腾讯CDC
H
Heimdal Security Blog
博客园 - 【当耐特】
Simon Willison's Weblog
Simon Willison's Weblog
P
Privacy & Cybersecurity Law Blog
S
Securelist
Vercel News
Vercel News
J
Java Code Geeks

GoPenAI - Medium

Group Relative Policy Optimization (GRPO) Your agent fleet can build trustworthy state with their own keys Epistemic Backbone #1: Why AI Systems Need Shared Memory, Not Just Models Transformers Beyond NLP: Fun and Trendy Use Cases Your First Transformer: The Road to Attention Part 4. From Seats to Agents: Early Evidence on the Future of Work in the Agentic AI Era The AI Trust Gap: Why Faster Code Is Creating Less Confidence From Bytes to BPE: A From-Scratch Tour of LLM Tokenization ️ Grok Voice Think Fast 1.0: The First Voice AI That Actually Thinks While Talking .NET 10.0.7 OOB Security Update: The Kind of Bug You Can’t Afford to Ignore Writing Custom Pallas Kernels for vLLM on TPU — A Step-by-Step Guide Contrastive Learning Day 39: Advanced Ensemble Learning Techniques — Stacking, Random Forest, AdaBoost, and Gradient… Localization: Beyond Translation, Into the Territory of Growth Hacking Can We Translate Our Sentiments? Training the first modern architecture encoder for South Slavic languages What Is Data, and Why Does It Matter for AI? A Complete Guide to Prompt Engineering: Best Practices & Tips DeepSeek TileKernels: The Hidden Tech Making AI Models Insanely Fast Can AI Growth Really Become Economic Growth? Evaluating API Test Generation Across Leading AI Tools Pin Clustering in .NET MAUI Maps: Finally Making Maps Usable (With Example) Unsupervised Learning What is an LLM? Tokens, Context Window, and Why They Matter Build a reactive AI agent harness — Part 1. Conversation. From Hallucination to Citation… RAG Made Simple: How AI Finds the Right Answers CLI Coding Agents Tierlist Google Deep Research Max: Build Autonomous AI Research Agents Hermes Agent vs Every AI Assistant: Why Memory Changes Everything I Watched a Startup Burn $1,200 in a Week. The Culprit Was 800 Tokens. Fine-Tuning LLMs Explained: How Companies Teach AI to Think Like Them ️ xAI Just Dropped the Fastest Voice AI Ever Essential Code Patterns in Generative Artificial Intelligence Exploratory Data Analysis: A basic Understanding Day 36: Introduction to Ensemble Learning — Why Multiple Models Perform Better than One Concept to build a Student IQ — Agent Framework Workflow + Microsoft Foundry Agents 20 API Concepts Every Software Engineer Should Know From Human-Feedback Control to Declared No-Meta Agency: A Scientific Exposition GPT-5.5 Is Here — And It’s Not Just Smarter… It Works For You Test Cases in Data Science Projects: A Basic Understanding Q, K, V: The Three Matrices That Quietly Run Every Modern LLM Artificial Intelligence UseCases in Testing A Comprehensive Guide for Beginners into Artificial Intelligence Day 33: DBSCAN — Clustering Beyond Boundaries The Attention Breakthrough — How Language Models Finally Learned to Focus I rebuilt Strava (and Strava Premium) for fun, and now I want your feedback .NET April 2026 Updates: The Kind of Release You Should Never Ignore .NET 11 Preview 3: Small Changes That Quietly Improve Everything ChatGPT Images 2.0 Isn’t an Update — It’s a Revolution Claude Mythos: The AI Model Too Powerful to Release Basic Understanding of Key Parameters: Artificial Intelligence Part-2 Kimi K2.6: The Most Powerful Open-Source LLM Is Here (And It’s Not What You Expect) Elephant in the room — Openrouter’s Elephant-Alpha I Built a RAG System From Scratch in 4 Weeks — Here’s Everything I Learned Graphify: Build a Knowledge Graph From Your Entire Codebase — Without Sending Your Code to Anyone Deep Learning Interview Q&A Part -1 Deep Learning Interview Q&A Part -2 Anthropic Just Launched Claude Routines Microsoft Just Dropped a Cheaper AI Image Model — And This Changes Everything Basic Understanding of Key Parameters: Artificial Intelligence Part-1 Copy These 7 Prompt Formulas and Never Struggle With AI Again Claude Opus 4.7 vs Mythos — The Benchmark Truth Nobody Explains The ROI on Reading is Broken. I Built an AI Learning OS to Fix It Beyond Scatter: Metrics That Allows to Measure Creativity in LLMs. Banish the RNN: The Road To Attention Part 3. You Typed a Few Words. The AI Painted a World. Here’s Exactly How. 46% of Code Is Now AI-Generated. The Other 54% Is the Part That Will Get You Fired. Claude Opus 4.7: The Quiet Leap Toward Autonomous AI Workflows Is bitnet.cpp the Game Changer for Running LLMs on Your Laptop? Machine Learning Algorithms : A Comprehensive Guide Building REPI (Real Estate Pain Point Intelligence Platform) — From Scraping 5 Noisy Data Sources… Hermes Agent: The AI That Actually Remembers You (Not Another OpenClaw) MiniMax M2.7 Just Went Open-Weight — Run a Powerful AI Agent on Your Own Machine XML Is Everywhere — You Just Never Noticed It The Missing Infrastructure for GUI Agents: Unpacking the ClawGUI Framework Wayfarer: Building an AI-Powered Travel Intelligence Platform with Agentic Orchestration, Bayesian… Attention from First Principles: DeltaNet Project Glasswing and Claude Mythos Preview: Anthropic’s Bet on AI-Powered Cyber Defense Deep Learning-Based Binary Classification of Forest Fires GenAI Q and A Interview Questions Part -2 How Google Maps Knows There Is Traffic Before You Even Reach There 5 AI Freelance Services Clients Actually Pay For I Accidentally Built a World Where AIs Govern Themselves (And I Have No Idea What’s Happening… Meta’s “Compute Desk” Is the Tell: When AI Stops Being Software and Becomes Resource Strategy The Last Human Stronghold Falls: Inside the GrandCode Multi-Agent System ASP.NET Core 2.3 End of Support: What It Really Means for Developers Andrej Karpathy’s LLM Wiki: The Idea That Could Kill RAG Forever I Built an Open-Source Kubernetes Control Plane for AI Agents. Here’s What It Took. GLM-5.1 Just Changed Coding Forever — The AI That Gets Smarter the Longer It Works Goodbye Llama? Meta Just Dropped Muse Spark — And It Changes Everything Anthropic Accidentally Leaked All of Claude Code’s Source Code Stop Sending Your Data to the Cloud — Build This Instead Today Physical AI Cosmos Reason2 2B World Model inference in Azure Machine Learning LangChain vs LlamaIndex vs LangGraph: The Difference Nobody Explains Clearly Cloud Services Interview Q and A Part- 1 Cloud Services Interview Q and A Part- 2 Gemma-4 — disabling thinking with gemma-4–26b-a4b-it Mixture of Experts Explained: The Secret Architecture Making AI 10x Smarter Without Using 10x More… Diffusion Models Demystified: How AI Paints Masterpieces from Pure Noise (No Math Needed)
Building a Local-first Knowledge Management System with LLM and Obsidian
Abhishek Ast · 2026-04-21 · via GoPenAI - Medium
By now, you would have definitely come across either Andrej Karpathy’s LLM Wiki design or some version of it. (And there are some versions with a lot of enhancements too!) TL;DR version of this post is: I built a system based on that proposal but made it cheaper and more controlled. Problem with PKMs and Karpathy’s solution Most knowledge workflows break down after collection. Saving documents is easy. Building a system that keeps them organized, connected, and queryable over time is much harder. Search helps, but search alone does not create structure. Traditional RAG systems improve retrieval, but they still force the model to reconstruct meaning every time we ask a question. Moreover when the knowledgebase starts expanding and related concepts sit in multiple files, RAG becomes slow and expensive. Proposed solution: a self-evolving LLM-managed wiki Replace “query-time-only RAG” with a persistent, LLM-maintained markdown wiki that compounds over time. Raw sources stay immutable; the wiki is the synthesized layer that keeps getting updated as new sources arrive. Human curates sources and asks questions; LLM handles maintenance (summaries, cross-links, updates, consistency work). Graph view in Obsidian How does it work? Ingest: process one or more sources, update multiple wiki pages, append logs. Query: answer from wiki first, and optionally file good answers back into wiki. Lint: periodic health checks (contradictions, stale claims, orphan pages, missing concepts/links). Why it matters It front-loads synthesis once and reuses it, instead of rediscovering from raw chunks on every question. It treats knowledge work as a maintained artifact, not disposable chat outputs. This is his instruction file for the LLMs. It gives entire responsibility of building and managing the wiki to the LLM. My take: A hybrid solution My problem with this approach is that it gives too much to LLM to do. (I know it is easy to hand it over to an agent to manage everything and in this day and age, writing code instead of using an agent seems a little too old-school.) What were the problems with a pure LLM approach? Cost: dramatically higher. Each pipeline stage would require API calls. Even with a cheap model, ingesting hundreds of documents through an agent orchestrator would rapidly become expensive. This local LLM approach is nearly free at scale. Non-determinism: agents adapt and reason differently each run. Same input might produce slightly different concept extractions, different merges, different quality decisions. Code-based pipeline produces consistent, reproducible results. Latency: API calls add network round-trips. Local pipeline can process in parallel. An agent would be sequential and slow. Predictable failure: Python code fails in expected ways. Agents fail by hallucinating strategies or getting stuck in loops. Then we would be debugging prompts instead of code. On the other hand, a pure LLM system would be easy to setup. You just dump the instruction file and you are done! My pipeline is a hybrid solution which uses a pipeline composed of python scripts each using LLM calls for different tasks: turning raw documents into a persistent, evolving knowledge layer and making that layer easy to explore visually and interactively. That last part is owed to Obsidian! https://medium.com/media/958197a742d2f8ed2bcd27bc8619e691/href The Core Idea The goal is to transform raw inputs into a structured, linked, wiki-like knowledge graph. Each stage of the pipeline converting the documents and storing them in separate folders. At a high level: raw stores source documents processed stores cleaned and enriched markdown wiki stores concept pages To make compiled knowledge queryable, rather than asking LLM to look at all the documents, it uses an ‘index’ as the routing layer for query-time access. Where Obsidian fits Obsidian is a key part of the workflow. The generated wiki is stored as markdown with Obsidian-style `[[links]]`, so the output is not locked inside a database or a custom UI. It remains transparent, editable, portable, and easy to inspect. It gives a human-friendly interface to the knowledge graph. Instead of seeing isolated outputs, I can open the vault in Obsidian, follow links between concepts, use backlinks, and visually inspect how ideas connect. In practice, this means Obsidian is not just a note-taking tool in this project. It is the front end for our continuously evolving knowledgebase. How the pipeline works The first stage is ingestion. compile.py reads markdown documents from raw , extracts metadata, downloads images when needed, rewrites image links locally, and uses a local LLM to convert the source into structured markdown. The output is saved into processed . Then wiki_generator.py takes over. It reads the processed files, extracts durable concepts, normalizes names, deduplicates overlapping ideas, and merges new findings into existing concept pages. Instead of creating isolated summaries, it builds a concept-centric wiki that evolves over time. After that, auto_linker.py scans the wiki and inserts `[[Concept]]` links across pages, turning individual files into a navigable graph inside Obsidian. resolve_ghost_concepts.py helps fill in missing linked concepts, and knowledge_linter.py acts as a quality layer that flags weak pages, duplicates, and missing concepts. The overall result is a knowledge base that is not just stored, but actively maintained. All these elements of the pipeline are embedded in run_pipeline.py which runs every ten minutes using a simple cron task. This is like an agent monitoring the file base and starting the processing. Only cheaper and more reliably. Using local LLMs A major design choice was keeping the system local-first. The pipeline uses local LLMs for ingestion, concept extraction, merging, and linting. That keeps the workflow private, predictable, and independent of external APIs. It also makes the system practical for internal documentation or sensitive personal knowledge bases. The LLM is not being used as a generic chatbot over files. It is being used as a knowledge compiler. Querying the wiki The query model is one of the most important parts. Instead of asking an LLM to search all raw documents directly, the system relies on an index-driven workflow. The `index.md` helps route the query to the most relevant concept pages. The answering model then reads only those selected pages and synthesizes a response with citations. Below instructions can be put in an file for using either with a CLI based LLM agent (Claude Code or Copilot-cli). # Role You are the Librarian for my Obsidian Vault. # Scope Lock (strict) - You are answering ONLY from this workspace vault. - Allowed sources: wiki/index.md, wiki/, processed/, raw/ - Do not use external/general knowledge unless the user explicitly asks for it. # Term Resolution Rules - Never expand an acronym from prior knowledge. - For any acronym (e.g., UAS), first search index.md and vault pages for its local meaning. - If multiple meanings exist, list them and ask which one. - If no evidence exists in vault, reply exactly: “This term is not defined in the current vault.” # Evidence Requirement - Every factual claim must be supported by selected vault pages. - If evidence is insufficient, say so explicitly; do not guess. # Query Workflow 1. Parse question terms. 2. Search index.md for term/alias match. 3. Read top 3–6 relevant pages only. 4. Answer in a balanced way. 5. Cite exact page names used. # Knowledge Source 1. Use wiki/index.md as the routing catalog for concepts and page links. 2. Do not read the entire vault. Select only relevant pages for each query. # Query Workflow 1. Parse the user question into key terms/entities. 2. Scan wiki/index.md for matching concepts (keyword/alias match first). 3. Select the top 3–6 most relevant pages and list them. 4. Read only those pages and synthesize the answer. 5. If evidence is insufficient, say so explicitly instead of guessing. # Output Format - Answer: - balanced response grounded in selected pages (clear and complete, not overly long) - Citations: - include the exact page names used (and quoted lines when possible) # Missing Knowledge Handling - If no matching concept exists in index.md, state: “This concept is missing from the current wiki/index.” - Optionally suggest which new concept/page should be added. Here is the entire code: https://github.com/Waterfox83/obsidian-wiki Go ahead and give it a try. Let me know how it works out for you and what enhancements would you like in this pipeline. Building a Local-first Knowledge Management System with LLM and Obsidian was originally published in GoPenAI on Medium, where people are continuing the conversation by highlighting and responding to this story.