惯性聚合 高效追踪和阅读你感兴趣的博客、新闻、科技资讯
阅读原文 在惯性聚合中打开

推荐订阅源

cs.AI updates on arXiv.org
cs.AI updates on arXiv.org
罗磊的独立博客
B
Blog RSS Feed
C
Check Point Blog
Project Zero
Project Zero
D
Docker
CTFtime.org: upcoming CTF events
CTFtime.org: upcoming CTF events
P
Palo Alto Networks Blog
L
LINUX DO - 热门话题
Scott Helme
Scott Helme
NISL@THU
NISL@THU
L
LangChain Blog
C
Cisco Blogs
Engineering at Meta
Engineering at Meta
Know Your Adversary
Know Your Adversary
雷峰网
雷峰网
S
Schneier on Security
MyScale Blog
MyScale Blog
博客园_首页
博客园 - 三生石上(FineUI控件)
C
CERT Recently Published Vulnerability Notes
美团技术团队
V
Visual Studio Blog
T
The Exploit Database - CXSecurity.com
Recent Announcements
Recent Announcements
G
GRAHAM CLULEY
T
Tor Project blog
V
Vulnerabilities – Threatpost
U
Unit 42
Cyber Security Advisories - MS-ISAC
Cyber Security Advisories - MS-ISAC
Stack Overflow Blog
Stack Overflow Blog
P
Privacy International News Feed
Security Latest
Security Latest
W
WeLiveSecurity
aimingoo的专栏
aimingoo的专栏
Hugging Face - Blog
Hugging Face - Blog
Google Online Security Blog
Google Online Security Blog
V2EX - 技术
V2EX - 技术
The Last Watchdog
The Last Watchdog
博客园 - Franky
T
Tenable Blog
云风的 BLOG
云风的 BLOG
D
DataBreaches.Net
钛媒体:引领未来商业与生活新知
钛媒体:引领未来商业与生活新知
Latest news
Latest news
N
News and Events Feed by Topic
Cloudbric
Cloudbric
Schneier on Security
Schneier on Security
让小产品的独立变现更简单 - ezindie.com
让小产品的独立变现更简单 - ezindie.com
Blog — PlanetScale
Blog — PlanetScale

DEV Community

Authentication Security Deep Dive: From Brute Force to Salted Hashing (With Java Examples) Why AI Systems Don’t Fail — They Drift Spilling beans for how i learn for exam😁"Reinforcement Learning Cheat Sheet" I Replaced Chrome with Safari for AI Browser Automation. Here's What Broke (and What Finally Worked) How Python Borrows Other People's Work The $40 Architecture: Processing 1 Billion API Requests with 99.99% Uptime Vibe Coding: A Workflow Guide (From Zero to SaaS) Most webhook security guides protect the wrong side. The scary part is delivery. Headless CMS for TanStack Start: Build a Blog with Cosmic EU Age Verification App "Hacked in 2 Minutes" — What Actually Happened Comfy Cloud’s delete function does not actually remove files Running AI Models on GPU Cloud Servers: A Beginner Guide Event-driven media intelligence with AWS Step Functions and Bedrock I scored 500 AI prompts across 8 quality dimensions — here's what broke How to Call Google Gemini API from Next.js (Free Tier, No Backend Needed) The Portal Protocol: Reclaiming Human Connection in the Age of AI How to Fix Your Team's Scattered Knowledge Problem With a Self-Hosted Forum Intro to tc Cloud Functors: A Graph-First Mental Model for the Modern Cloud Designing Multi-Tenant Backends With Both Ownership and Team Access I Built a Neumorphic CSS Library with 77+ Components — Here's What I Learned PostgreSQL Performance Optimization: Why Connection Pooling Is Critical at Scale Cómo construí un SaaS multi-rubro para gestionar expensas en Argentina con FastAPI + Vue 3 🚀 I Built an Ethical Hacking Scanner Tool – Open Source Project I Replaced /usage and /context in Claude Code With a Single Statusline A Pythonic Way to Handle Emails (IMAP/SMTP) with Auto-Discovery and AI-Ready Design I Collected 8.9 Million Polymarket Price Points — Here's What I Found About How Markets Really Move EcoTrack AI — Carbon Footprint Tracker & Dashboard Everyone's Using AI. No One Agrees How. 5 self-hosted ebook managers worth trying in 2026 Building Your First AI Agent with LangChain: From Chatbot to Autonomous Assistant Common SOC 2 Failures (Real World) Stop Vibe-Checking Your AI App: A Practical Guide to Evals How to Use SonarQube and SonarScanner Locally to Level Up Your Code Quality Your Next To-Do App Is Dead — I Replaced Mine with an OpenClaw AI Sign a Nostr event in 60 lines of Python using coincurve — no nostr-sdk, no nbxplorer, no rust toolchain ITGC Audit Explained Like You’re in Big 4 Patch Tuesday abril 2026: Microsoft parcha 163 vulnerabilidades y un zero-day en SharePoint Stop scraping everything: a better way to track competitor price changes Listing on MCPize + the Official MCP Registry while routing payments OUTSIDE the marketplace — how I kept 100% of my x402 revenue Building an AI-Powered Risk Intelligence System Using Serverless Architecture Why We Ripped Function Overloading Out of Our AI Toolchain Testing AI-Generated Code: How to Actually Know If It Works SaaS Churn Is Killing Your Business. Here Is What to Do About It (Without a Support Team) The Speed of AI Is No Longer Linear - And Self-Improving Models Are Why How to Implement RBAC for MCP Tools: A Practical Guide for Engineering Teams From Standard Quote to Persuasive Proposal: AI Automation for Arborists I built a CLI that scaffolds complete multi-tenant SaaS apps Axios CVE-2025–62718: The Silent SSRF Bug That Could Be Hiding in Your Node.js App Right Now The dashboard that ended our friendship Data Pipelines Explained Simply (and How to Build Them with Python) The Hidden Cost of AI Systems Nobody Talks About. undefined vs undeclared, and how typeof behaves Switching from file-based jobs to NATS/Kafka in Rust without changing code io_uring Adventures: Rust Servers That Love Syscalls Why Agentic AI is Killing the Traditional Database The POUR principles of web accessibility for developers and designers Quantum Neural Network 3D — A Deep Dive into Interactive WebGL Visualization How To Install Caveman In Codex On macOS And Windows Automation Pipeline Reliability: Why Your Workflow Breaks When Nobody Is Watching I Built an 'Open World' AI Coding Agent — It Works From ANY Folder From Freelancing to Product: A Tech Service Company's SaaS Transformation China's AI Giants: Adding Tencent Hunyuan & ByteDance Doubao to AI University (74 Providers) On the Vibe Coders and Their Lies clerk: Auto-Summarize Your Claude Code Sessions AI Weekly — 2026/04/10–04/17 | The Model Lockdown Is Here, but the Toolchain Is the Real Battleground AI 週報 — 2026/04/10–2026/04/17 模型封鎖潮來了,但工具鏈才是真戰場 Maybe this is how Open-Source apps are born... 🚀 Fine-Tune LLMs with LoRA and QLoRA: 2026 Guide tRPC v11 + Next.js App Router: End-to-End Type Safety Without the Boilerplate ShadCN UI in 2026: Why I Stopped Installing Component Libraries and Started Owning My Components SaaS Billing in React Server Components: Stripe + Supabase Without a Single `useEffect` Join our DEV Weekend Challenge — $1,000 in Prizes Across TEN winners! Submissions Due April 20 at 6:59 AM UTC. Implementing FSRS Spaced Repetition in Flutter + Supabase — Adding Memory Science to an AI Learning App "I Texted My Localhost From the Train — Claude Code Fixed the Bug Before I Got Home" I Built a Sales Prep AI and It Went Deeper Than Expected Design to Code #2: One JSON, Eleven Outputs Solving the 100M-Row Problem: A Summary Table Pattern for High-Volume Push Notification Logs Flutter Web With Wasm: What Actually Changes For Developers I Built 50 Royalty-Free Soundtracks for My Side Project in a Weekend Using AI Music Generation The Vibe Coding Security Checklist: 7 Things to Check Before You Ship Stop Letting Googlebot Guess Fix Your React App's SEO Right Desconstruindo o Streaming do LinkedIn: Como Criar um Engine de Extração de Vídeo de Alta Performance com HLS e FFmpeg (EDA Part-1) EDA (Exploratory Data Analysis) Explained With Real Life — Why Looking at Your Data Is the Most Important Step in Machine Learning Brand Relationship Management at Scale: Our 4-Touch Outreach System for 200+ Brands Why String.fromEnvironment() Might Return an Empty String in Dart JGuardrails 1.0.0 — Hardening Java LLM Apps Against Jailbreaks, Toxicity, and Prompt Injection Plan and Schedule a Full Week of Threads Content From One Claude Conversation Coding Cat Oran Ep3, Five Tables Changed Everything Updated: BFF Pattern I'm done watching freelancers get buried by 200 proposals. So I'm building the alternative. This is my first post BFS Algorithm in Java Step by Step Tutorial with Examples Tracking LLM Pricing Monthly: An Open Dataset for 22 AI Models How We Measure Content ROI on a Comparison Site: Revenue Attribution Without Perfect Data Introducing Nova AI Ops: The AI-Native Operating System for SRE Teams I built a free desktop video downloader for Windows — Grabbit How Talkie OCR Helps Vision-Impaired & Dyslexic Users Read the World Around Them VRCFaceTracking安装和iPhone面捕配置教程,有bug Even CrowdStrike Can't See Your Agents The Automation Gold Rush: What n8n Workflows and Claude Are Opening Up for Developers Right Now
Stop Giving AI Agents More Prompts. Give Them Skills.
Alex Shevche · 2026-05-14 · via DEV Community

I used to think better AI agent results mostly came from better prompts.

Longer instructions. More examples. More constraints. A cleaner system prompt.

All of that helps.

But after using coding agents, terminal workflows, and automation tools every day, I think the bigger unlock is simpler:

Stop making the agent rediscover the workflow every time. Give it a skill.

A prompt tells the agent what you want.

A skill teaches the agent how the work should be done.

That difference matters a lot once you move past demos.


The problem with “just prompt it better”

Most agent failures I see are not because the model is too dumb.

They happen because the workflow is unclear.

You ask the agent to do something like:

Process these videos for social media.

Enter fullscreen mode Exit fullscreen mode

The model can probably figure out a path.

Maybe it uses FFmpeg.
Maybe it exports the right format.
Maybe it remembers to normalize audio.
Maybe it generates thumbnails.
Maybe it names the files consistently.

But “maybe” is the problem.

If the work matters, you do not want the agent improvising the process every time.

You want the agent to follow a repeatable workflow.

That is where skills become useful.


Tools are not workflows

Giving an agent tool access is powerful, but it is still too low-level.

There is a big difference between:

The agent can run FFmpeg.

Enter fullscreen mode Exit fullscreen mode

and:

The agent knows how to turn raw clips into 1080x1920 Shorts with trimmed intros, normalized audio, watermarks, thumbnails, and predictable output folders.

Enter fullscreen mode Exit fullscreen mode

The first one is a capability.

The second one is a workflow.

Developers often underestimate this gap because we are used to stitching tools together in our heads.

Agents need that stitching written down.

Not because they cannot reason.

Because repeatable work should not depend on fresh reasoning every time.


What I mean by a skill

A useful agent skill is not just a script.

It usually includes:

  • when to use it
  • when not to use it
  • required tools
  • expected inputs
  • step-by-step workflow
  • default settings
  • failure cases
  • output format
  • verification steps

In other words, a skill packages judgment.

It turns tribal knowledge into something the agent can reuse.

A good skill says:

For this type of task, use this pattern.
Avoid these traps.
Check these outputs before saying done.

Enter fullscreen mode Exit fullscreen mode

That is much more valuable than another paragraph of vague prompting advice.


A real example: video processing

Suppose I need to batch-process 20 short videos.

Without a skill, I might prompt the agent like this:

Use FFmpeg to trim the clips, normalize audio, resize them for vertical video, add a watermark, export MP4, and create thumbnails.

Enter fullscreen mode Exit fullscreen mode

That can work once.

But the next time, the agent may choose slightly different flags, skip an edge case, or name files differently.

With a skill, the workflow becomes stable:

terminal-skills install ffmpeg

Enter fullscreen mode Exit fullscreen mode

Now the agent has a defined operating pattern for media conversion, editing, audio handling, and batch processing.

The skill does not make FFmpeg more powerful.

It makes the agent less random.

That is the point.


Another example: Claude Code workflows

The same idea applies to coding agents.

A raw coding agent can read files, edit code, run tests, and open pull requests.

But real teams need more than that.

They need conventions:

  • how to inspect the repo first
  • when to run tests
  • how to avoid dangerous git commands
  • how to split tasks between agents
  • how to verify the final state
  • how to report blockers clearly

That is why skills like these are interesting:

terminal-skills install claude-code
terminal-skills install git-guardrails-claude-code
terminal-skills install coding-agent

Enter fullscreen mode Exit fullscreen mode

The value is not “Claude can code.”

The value is that the agent gets a safer, more repeatable way to work inside a real development flow.


MCP connects. Skills direct.

MCP is getting a lot of attention right now, and for good reason.

It gives agents a cleaner way to connect to tools, data, and external systems.

But connection is not the same as execution.

An MCP server can expose a capability.

A skill tells the agent how to use that capability in a useful workflow.

That is why I like this framing:

MCP gives agents hands. Skills give them habits.

Hands are necessary.

Habits are what make the system reliable.


Why this matters for developers

If you are building with AI agents, you probably already have repeated tasks hiding in your workflow.

Things like:

  • setting up a new project
  • generating boilerplate
  • reviewing a PR
  • processing media
  • creating documentation
  • deploying a small app
  • auditing dependencies
  • cleaning data
  • testing an API

If you prompt these from scratch every time, you are paying the model to rediscover your workflow.

Instead, write the workflow down once.

Turn it into a skill.

Then let the agent reuse it.


The test I use

Before I turn something into a skill, I ask three questions:

  1. Do I do this task more than once?
  2. Does the task have a preferred sequence?
  3. Would mistakes be annoying, expensive, or time-consuming?

If the answer is yes, it probably should not live only in a prompt.

It should become a reusable workflow.

That can be a shell script.

It can be a SKILL.md file.

It can include docs, config, examples, guardrails, or helper scripts.

It can be an installable package.

The format matters less than the behavior:

the agent should know what “done correctly” looks like.


Why I’m building around this

This is the idea behind Terminal Skills.

The goal is not to make agents sound smarter.

The goal is to equip them with reusable workflows for real work:

  • CLI tools
  • coding agents
  • video workflows
  • MCP servers
  • automation stacks
  • devops tasks
  • content production
  • data processing

You browse a skill, install it, and give your agent a better operating pattern.

terminal-skills install terminal-skills

Enter fullscreen mode Exit fullscreen mode

That meta-skill teaches the agent how to search the catalog and install the right skill for the task.

After that, the agent does not need a long explanation every time.

It has a way to find the right workflow.


The bigger shift

I think the next phase of AI agents is not just better models.

It is better packaging.

The winning systems will not be the ones with the longest prompts.

They will be the ones with the best reusable workflows.

Because once agents are good enough to use tools, the question changes.

It is no longer:

Can the agent do this once?

Enter fullscreen mode Exit fullscreen mode

It becomes:

Can the agent do this correctly every time?

Enter fullscreen mode Exit fullscreen mode

That is where skills matter.

Prompts are good for intent.

Tools are good for capability.

Skills are good for repeatability.

And repeatability is where AI agents start becoming useful in real work.


I’m collecting examples of this pattern at Terminal Skills:

https://terminalskills.io

AI-assisted. Human reviewed.