惯性聚合 高效追踪和阅读你感兴趣的博客、新闻、科技资讯
阅读原文 在惯性聚合中打开

推荐订阅源

月光博客
月光博客
D
Docker
腾讯CDC
J
Java Code Geeks
大猫的无限游戏
大猫的无限游戏
The Cloudflare Blog
Martin Fowler
Martin Fowler
MongoDB | Blog
MongoDB | Blog
博客园 - Franky
博客园 - 三生石上(FineUI控件)
Recent Announcements
Recent Announcements
F
Fortinet All Blogs
IT之家
IT之家
WordPress大学
WordPress大学
M
MIT News - Artificial intelligence
爱范儿
爱范儿
Microsoft Azure Blog
Microsoft Azure Blog
Vercel News
Vercel News
钛媒体:引领未来商业与生活新知
钛媒体:引领未来商业与生活新知
小众软件
小众软件
N
Netflix TechBlog - Medium
T
Tailwind CSS Blog
Engineering at Meta
Engineering at Meta
博客园 - 【当耐特】

DEV Community

Authentication Security Deep Dive: From Brute Force to Salted Hashing (With Java Examples) Why AI Systems Don’t Fail — They Drift Spilling beans for how i learn for exam😁"Reinforcement Learning Cheat Sheet" I Replaced Chrome with Safari for AI Browser Automation. Here's What Broke (and What Finally Worked) How Python Borrows Other People's Work The $40 Architecture: Processing 1 Billion API Requests with 99.99% Uptime Vibe Coding: A Workflow Guide (From Zero to SaaS) Most webhook security guides protect the wrong side. The scary part is delivery. Headless CMS for TanStack Start: Build a Blog with Cosmic EU Age Verification App "Hacked in 2 Minutes" — What Actually Happened Comfy Cloud’s delete function does not actually remove files Running AI Models on GPU Cloud Servers: A Beginner Guide Event-driven media intelligence with AWS Step Functions and Bedrock I scored 500 AI prompts across 8 quality dimensions — here's what broke How to Call Google Gemini API from Next.js (Free Tier, No Backend Needed) The Portal Protocol: Reclaiming Human Connection in the Age of AI How to Fix Your Team's Scattered Knowledge Problem With a Self-Hosted Forum Intro to tc Cloud Functors: A Graph-First Mental Model for the Modern Cloud Designing Multi-Tenant Backends With Both Ownership and Team Access I Built a Neumorphic CSS Library with 77+ Components — Here's What I Learned PostgreSQL Performance Optimization: Why Connection Pooling Is Critical at Scale Cómo construí un SaaS multi-rubro para gestionar expensas en Argentina con FastAPI + Vue 3 🚀 I Built an Ethical Hacking Scanner Tool – Open Source Project I Replaced /usage and /context in Claude Code With a Single Statusline A Pythonic Way to Handle Emails (IMAP/SMTP) with Auto-Discovery and AI-Ready Design I Collected 8.9 Million Polymarket Price Points — Here's What I Found About How Markets Really Move EcoTrack AI — Carbon Footprint Tracker & Dashboard Everyone's Using AI. No One Agrees How. 5 self-hosted ebook managers worth trying in 2026 Building Your First AI Agent with LangChain: From Chatbot to Autonomous Assistant
I Built a Private AI Brain on My Laptop for $0
TheonaiaO · 2026-06-14 · via DEV Community

TheonaiaO

Last week I couldn't shake an idea: what if I had an AI that knew everything I know? Not ChatGPT — something on my hardware, holding my knowledge, answering to no one's API bill.

Yesterday I built it. Here's the honest breakdown.

What it does

NEXUS runs on a regular Windows laptop — aging i7, 16GB RAM, no GPU. It:

  • Remembers everything. Drop any file in a folder; 60 seconds later it's searchable memory.
  • Answers from MY knowledge. "Which of my projects were formally closed and why?" — it answers from my actual records.
  • Watches the live web. Every 2 hours it pulls Hacker News and news feeds, learns what's trending, pings my Telegram.
  • Reports to my phone. 7 AM daily briefing: what it learned, what's running, what needs me.

The stack — all free, all open source

Ollama runs the models (Llama 3.2, Mistral 7B). Open WebUI is my private ChatGPT. Qdrant stores memory. n8n automates. SearXNG searches privately. PostgreSQL, Redis, and MinIO handle data.

Commercial equivalent: $300–500/month. My cost: electricity.

The memory trick nobody explains simply

  1. Parse — extract text from any file
  2. Chunk — split into ~300-word pieces
  3. Embed — each chunk becomes 768 numbers representing its meaning
  4. Store — a database that searches by similarity

Your question becomes 768 numbers too, and the database finds memories with similar meaning — not matching keywords. I asked "how do I get clients cheaper" and it found my notes on "reducing customer acquisition cost." Different words. Same meaning. That's the magic.

What surprised me

  • A 2GB model is genuinely useful. Llama 3.2 3B answers from my knowledge in seconds, on CPU.
  • The automation matters more than the AI. The watched folder + Telegram bot turned a cool demo into a system I actually use.
  • Windows is fine. Docker Desktop + WSL2 ran all nine services without drama.

The bill, honestly

  • Hardware: $0 (laptop I own)
  • Software: $0 (open source)
  • APIs: $0 (all local)
  • Time: one focused day

The only future cost is a cloud GPU server (~$65/mo) when I outgrow the laptop — and the plan is for the system to pay for that itself.

It already acts

By evening, NEXUS ran its first autonomous research mission: it searched the web, read five industry reports, cross-referenced its own memory, and delivered a cited market analysis to my phone — while I made coffee.

Next: a Writing Agent that drafts in my voice, and a Monitor Agent that hunts opportunities in the feeds it's already collecting.

An intelligence that doesn't just remember — it acts.


I'm documenting the whole build in public — every command, every dollar, every failure.

Ask me anything about the setup in the comments.

Stack: Ollama · Open WebUI · Qdrant · PostgreSQL · Neo4j · n8n · SearXNG · Redis · MinIO · Docker — all open source.