惯性聚合 高效追踪和阅读你感兴趣的博客、新闻、科技资讯
阅读原文 在惯性聚合中打开

推荐订阅源

D
Docker
酷 壳 – CoolShell
酷 壳 – CoolShell
博客园 - Franky
奇客Solidot–传递最新科技情报
奇客Solidot–传递最新科技情报
A
About on SuperTechFans
博客园 - 【当耐特】
Microsoft Security Blog
Microsoft Security Blog
H
Hackread – Cybersecurity News, Data Breaches, AI and More
The GitHub Blog
The GitHub Blog
雷峰网
雷峰网
博客园_首页
freeCodeCamp Programming Tutorials: Python, JavaScript, Git & More
IT之家
IT之家
博客园 - 叶小钗
Google DeepMind News
Google DeepMind News
aimingoo的专栏
aimingoo的专栏
博客园 - 聂微东
B
Blog RSS Feed
H
Help Net Security
Recent Announcements
Recent Announcements
阮一峰的网络日志
阮一峰的网络日志
D
DataBreaches.Net
L
LangChain Blog
Vercel News
Vercel News

DEV Community

Authentication Security Deep Dive: From Brute Force to Salted Hashing (With Java Examples) Why AI Systems Don’t Fail — They Drift Spilling beans for how i learn for exam😁"Reinforcement Learning Cheat Sheet" I Replaced Chrome with Safari for AI Browser Automation. Here's What Broke (and What Finally Worked) How Python Borrows Other People's Work The $40 Architecture: Processing 1 Billion API Requests with 99.99% Uptime Vibe Coding: A Workflow Guide (From Zero to SaaS) Most webhook security guides protect the wrong side. The scary part is delivery. Headless CMS for TanStack Start: Build a Blog with Cosmic EU Age Verification App "Hacked in 2 Minutes" — What Actually Happened Comfy Cloud’s delete function does not actually remove files Running AI Models on GPU Cloud Servers: A Beginner Guide Event-driven media intelligence with AWS Step Functions and Bedrock I scored 500 AI prompts across 8 quality dimensions — here's what broke How to Call Google Gemini API from Next.js (Free Tier, No Backend Needed) The Portal Protocol: Reclaiming Human Connection in the Age of AI How to Fix Your Team's Scattered Knowledge Problem With a Self-Hosted Forum Intro to tc Cloud Functors: A Graph-First Mental Model for the Modern Cloud Designing Multi-Tenant Backends With Both Ownership and Team Access I Built a Neumorphic CSS Library with 77+ Components — Here's What I Learned PostgreSQL Performance Optimization: Why Connection Pooling Is Critical at Scale Cómo construí un SaaS multi-rubro para gestionar expensas en Argentina con FastAPI + Vue 3 🚀 I Built an Ethical Hacking Scanner Tool – Open Source Project I Replaced /usage and /context in Claude Code With a Single Statusline A Pythonic Way to Handle Emails (IMAP/SMTP) with Auto-Discovery and AI-Ready Design I Collected 8.9 Million Polymarket Price Points — Here's What I Found About How Markets Really Move EcoTrack AI — Carbon Footprint Tracker & Dashboard Everyone's Using AI. No One Agrees How. 5 self-hosted ebook managers worth trying in 2026 Building Your First AI Agent with LangChain: From Chatbot to Autonomous Assistant
Beyond MCP: Handling 845 Tools with 92% less context bloa...
Vermillion · 2026-05-12 · via DEV Community
Cover image for Beyond MCP: Handling 845 Tools with 92% less context bloat via Elemm

Vermillion

Hi everyone,

I’ve been diving deep into how AIs interact with tools and quickly hit a wall with the Model Context Protocol (MCP). As soon as you build complex, real-world toolsets, MCP becomes inefficient—bloating the context window and killing performance.

To solve this, I’ve developed Elemm **(E*very **Landmark **Enables **Massive **Modularity), also known as "The Landmark Manifest Protocol*."

👉 GitHub:Official Repository

Check out the docs and the benchmarks on GitHub.

What Elemm enables:

  • Custom Tooling: Turn any Python function into a "Landmark" with a single decorator.
  • Instant API Integration: Point to an OpenAPI or GraphQL URL, and your agent navigates it instantly with surgical precision.
  • Seamless Migration: Easily bridge your existing tools into a manifest-driven architecture.

The Landmark Advantage

Elemm doesn't cram every tool definition into the prompt. Instead, it provides the agent with a dynamic Manifest File for safe, "lazy-loaded" navigation.

The Benchmarks:

  • Scale: I gave an agent access to 845 tools simultaneously (GitHub API) with minimal token usage and 100% success rate on flagship models (Claude, Gemini, GPT-4).
  • Efficiency: Compared to classic MCP, Elemm shows -92% token savings and -84% fewer steps.
  • Edge Performance: Even using a tiny "goldfish-brain" model (Qwen 3.5 0.8B), I solved a multi-step forensic audit involving 111 tools with a 70% success rate. Standard MCP typically fails at the first step in this scenario.

Core Gateway Features:

  • Universal Gateway: A built-in bridge for OpenAPI, GraphQL, and native Elemm services via MCP.
  • On-Demand Discovery: Agents only load the definitions they actually need, preventing context overflow.
  • Sequence Engine: Execute multiple API calls in a single turn with native data piping (Output A → Input B).
  • Guardian Security: A policy engine that blocks dangerous patterns (e.g., delete_*) and hides restricted landmarks from the agent.
  • Secure Vault: Local credential management. API keys are injected server-side and never exposed to the LLM.
  • SmartRepair: Instead of cryptic stack traces, agents receive actionable "Remedies," allowing them to self-correct on the fly.

What this means for the future…

The era of manually hard-coding tool definitions is coming to an end. As we move toward Large Action Models and autonomous agents, we need a standardized, manifest-driven infrastructure that allows AI to navigate vast API landscapes without human intervention or context exhaustion. Elemm is the blueprint for this future: a world where agents don't just use tools we give them, but autonomously discover, secure, and master any interface they encounter.

Testimonials of the Agents:

"With ELEMM, I reduced token consumption by over 90% when deploying autonomous agents to large APIs—turning a $2.15 task into under $0.25."

Claude 4.6 Sonnet, Anthropic (via Claude Desktop)

"Elemm is a true game-changer; instead of juggling hundreds of tool definitions at once, I can discover complex APIs in a structured, token-efficient way on demand. The ability to batch multiple actions via execute_sequence allows me to solve tasks with far greater precision and significantly less context noise than with classic MCP."

Gemini 3 Flash, Google (Antigravity)

See some examples to learn how it works.

I’d love to hear your thoughts or discuss the walls you've hit when trying to scale MCP!