惯性聚合 高效追踪和阅读你感兴趣的博客、新闻、科技资讯
阅读原文 在惯性聚合中打开

推荐订阅源

Apple Machine Learning Research
Apple Machine Learning Research
OSCHINA 社区最新新闻
OSCHINA 社区最新新闻
G
Google Developers Blog
博客园 - 司徒正美
J
Java Code Geeks
aimingoo的专栏
aimingoo的专栏
A
About on SuperTechFans
博客园 - 三生石上(FineUI控件)
WordPress大学
WordPress大学
T
The Blog of Author Tim Ferriss
D
Docker
大猫的无限游戏
大猫的无限游戏
D
DataBreaches.Net
腾讯CDC
V
Visual Studio Blog
freeCodeCamp Programming Tutorials: Python, JavaScript, Git & More
C
Check Point Blog
M
MIT News - Artificial intelligence
Jina AI
Jina AI
I
InfoQ
雷峰网
雷峰网
The Cloudflare Blog
美团技术团队
Engineering at Meta
Engineering at Meta

DEV Community

Authentication Security Deep Dive: From Brute Force to Salted Hashing (With Java Examples) Why AI Systems Don’t Fail — They Drift Spilling beans for how i learn for exam😁"Reinforcement Learning Cheat Sheet" I Replaced Chrome with Safari for AI Browser Automation. Here's What Broke (and What Finally Worked) How Python Borrows Other People's Work The $40 Architecture: Processing 1 Billion API Requests with 99.99% Uptime Vibe Coding: A Workflow Guide (From Zero to SaaS) Most webhook security guides protect the wrong side. The scary part is delivery. Headless CMS for TanStack Start: Build a Blog with Cosmic EU Age Verification App "Hacked in 2 Minutes" — What Actually Happened Comfy Cloud’s delete function does not actually remove files Running AI Models on GPU Cloud Servers: A Beginner Guide Event-driven media intelligence with AWS Step Functions and Bedrock I scored 500 AI prompts across 8 quality dimensions — here's what broke How to Call Google Gemini API from Next.js (Free Tier, No Backend Needed) The Portal Protocol: Reclaiming Human Connection in the Age of AI How to Fix Your Team's Scattered Knowledge Problem With a Self-Hosted Forum Intro to tc Cloud Functors: A Graph-First Mental Model for the Modern Cloud Designing Multi-Tenant Backends With Both Ownership and Team Access I Built a Neumorphic CSS Library with 77+ Components — Here's What I Learned PostgreSQL Performance Optimization: Why Connection Pooling Is Critical at Scale Cómo construí un SaaS multi-rubro para gestionar expensas en Argentina con FastAPI + Vue 3 🚀 I Built an Ethical Hacking Scanner Tool – Open Source Project I Replaced /usage and /context in Claude Code With a Single Statusline A Pythonic Way to Handle Emails (IMAP/SMTP) with Auto-Discovery and AI-Ready Design I Collected 8.9 Million Polymarket Price Points — Here's What I Found About How Markets Really Move EcoTrack AI — Carbon Footprint Tracker & Dashboard Everyone's Using AI. No One Agrees How. 5 self-hosted ebook managers worth trying in 2026 Building Your First AI Agent with LangChain: From Chatbot to Autonomous Assistant
From Prompt Engineering to MCP Skills: What Rebuilding My...
neither galax · 2026-06-05 · via DEV Community

A recent comment on one of my dev.to posts asked a simple but insightful question:

What specifically was breaking before MCP: context loss between agents, or tool-call inconsistency?

At first, I thought the answer would be straightforward.

But after reflecting on my Tokyo Transit project, I realized the real issue wasn't either of those things.

The project has gone through three major iterations over the past year, and each version reflects a different stage in my understanding of AI agents, orchestration, and system architecture.

Looking back, the project became a timeline of how the AI ecosystem itself has evolved.

Version 1: SmolAgents and the "Big System Prompt" Era

The first version was built with SmolAgents.

Like many developers experimenting with agents at the time, I believed that if I wrote a sufficiently detailed system prompt, the agent would behave consistently.

As the project grew, the prompt grew with it.

More instructions.

More formatting rules.

More edge cases.

More exceptions.

Eventually, the system prompt became the primary mechanism for controlling behavior.

The result was predictable: the agent worked sometimes, but not reliably.

The biggest problem wasn't context loss between agents.

It wasn't tool-call failures either.

The real problem was prompt alignment.

I was trying to manage architecture through instructions.

Version 2: Google ADK and Higher-Level Abstractions

My next experiment used Google ADK.

Compared to the first version, ADK provided a high-level, cleaner, and more structured framework for agent development.

Many orchestration concerns were abstracted away.

This made development faster and reduced some of the complexity I had been manually managing.

But it also taught me something important:

Frameworks can simplify development, but they don't automatically solve architectural problems.

I still needed a clear way to define responsibilities, manage behavior, and structure agent workflows.

Version 3: MCP and Skill-Based Design

The current version uses the MCP Python SDK together with Skill-based design.

This was the point where the project finally started to feel maintainable.

Instead of pushing more logic into a growing system prompt, I could separate capabilities into tools and skills with clearly defined responsibilities.

MCP wasn't magical.

It didn't suddenly make the agent smarter.

What it provided was structure.

Skills gave me a dedicated place for behavioral instructions.

MCP provided a consistent interface for tools.

Together, they made the system easier to reason about, test, and improve.

So What Was Actually Breaking Before MCP?

Looking back, the biggest issue wasn't context loss or tool-call inconsistency.

It was architectural drift.

Whenever the agent behaved incorrectly, my solution was usually to add another instruction to the system prompt.

Over time, the prompt became harder to maintain and reason about.

The more complexity I added, the less predictable the behavior became.

In hindsight, I was trying to solve an architecture problem with prompt engineering.

Final Thoughts

The Tokyo Transit project is still under active development.

There are bugs to fix, improvements to make, and plenty left to learn.

But the most valuable outcome wasn't the transit tool itself.

It was seeing how my approach changed over time:

  • From large system prompts
  • To higher-level agent frameworks
  • To MCP and Skill-based architecture

The project became a record of my own learning journey as AI agents evolved from experiments into systems that can be structured, maintained, and improved over time.

If you've been building with agents, MCP, or AI frameworks, I'm curious: what was the biggest lesson that changed the way you design your systems? Feel free to share your experience in the comments below.