惯性聚合 高效追踪和阅读你感兴趣的博客、新闻、科技资讯
阅读原文 在惯性聚合中打开

推荐订阅源

Engineering at Meta
Engineering at Meta
奇客Solidot–传递最新科技情报
奇客Solidot–传递最新科技情报
小众软件
小众软件
博客园_首页
T
Tailwind CSS Blog
美团技术团队
博客园 - 叶小钗
Microsoft Security Blog
Microsoft Security Blog
有赞技术团队
有赞技术团队
Apple Machine Learning Research
Apple Machine Learning Research
大猫的无限游戏
大猫的无限游戏
Microsoft Azure Blog
Microsoft Azure Blog
H
Hackread – Cybersecurity News, Data Breaches, AI and More
I
InfoQ
MongoDB | Blog
MongoDB | Blog
The Cloudflare Blog
J
Java Code Geeks
OSCHINA 社区最新新闻
OSCHINA 社区最新新闻
博客园 - 聂微东
酷 壳 – CoolShell
酷 壳 – CoolShell
Blog — PlanetScale
Blog — PlanetScale
IT之家
IT之家
Cyber Security Advisories - MS-ISAC
Cyber Security Advisories - MS-ISAC
Y
Y Combinator Blog

DEV Community

Authentication Security Deep Dive: From Brute Force to Salted Hashing (With Java Examples) Why AI Systems Don’t Fail — They Drift Spilling beans for how i learn for exam😁"Reinforcement Learning Cheat Sheet" I Replaced Chrome with Safari for AI Browser Automation. Here's What Broke (and What Finally Worked) How Python Borrows Other People's Work The $40 Architecture: Processing 1 Billion API Requests with 99.99% Uptime Vibe Coding: A Workflow Guide (From Zero to SaaS) Most webhook security guides protect the wrong side. The scary part is delivery. Headless CMS for TanStack Start: Build a Blog with Cosmic EU Age Verification App "Hacked in 2 Minutes" — What Actually Happened Comfy Cloud’s delete function does not actually remove files Running AI Models on GPU Cloud Servers: A Beginner Guide Event-driven media intelligence with AWS Step Functions and Bedrock I scored 500 AI prompts across 8 quality dimensions — here's what broke How to Call Google Gemini API from Next.js (Free Tier, No Backend Needed) The Portal Protocol: Reclaiming Human Connection in the Age of AI How to Fix Your Team's Scattered Knowledge Problem With a Self-Hosted Forum Intro to tc Cloud Functors: A Graph-First Mental Model for the Modern Cloud Designing Multi-Tenant Backends With Both Ownership and Team Access I Built a Neumorphic CSS Library with 77+ Components — Here's What I Learned PostgreSQL Performance Optimization: Why Connection Pooling Is Critical at Scale Cómo construí un SaaS multi-rubro para gestionar expensas en Argentina con FastAPI + Vue 3 🚀 I Built an Ethical Hacking Scanner Tool – Open Source Project I Replaced /usage and /context in Claude Code With a Single Statusline A Pythonic Way to Handle Emails (IMAP/SMTP) with Auto-Discovery and AI-Ready Design I Collected 8.9 Million Polymarket Price Points — Here's What I Found About How Markets Really Move EcoTrack AI — Carbon Footprint Tracker & Dashboard Everyone's Using AI. No One Agrees How. 5 self-hosted ebook managers worth trying in 2026 Building Your First AI Agent with LangChain: From Chatbot to Autonomous Assistant
Building a Developer Tool on Gemini's Free Tier — What's ...
hiyoyo · 2026-05-03 · via DEV Community

All tests run on an 8-year-old MacBook Air.

Most AI integration tutorials assume you're paying for API access.

HiyokoLogcat is built entirely on Gemini's free tier — and designed so users bring their own free API key. Here's what's possible, what the limits are, and how to design around them.


The free tier numbers (as of 2026)

Gemini 2.5 Flash Preview:

  • 15 requests per minute (RPM)
  • 1,000,000 tokens per day
  • 250 requests per day

For a developer tool used intermittently — click diagnose, read, fix, move on — these limits are generous. A typical diagnosis uses 3,000–6,000 tokens. At that rate, you'd need to run 166+ diagnoses to hit the daily token limit.


Design for the free tier from the start

Keep requests small. Don't send 500 lines when 100 will do. Every token saved is headroom for more requests.

Don't auto-trigger. Never send to the API automatically — only on explicit user action. Auto-triggering on every error would burn through the RPM limit instantly.

Cache results. If the user closes the diagnosis overlay and reopens it, serve the cached result. Don't make a new API call for the same log line.

use std::collections::HashMap;

pub struct DiagnosisCache {
    cache: HashMap,  // log hash → diagnosis
}

impl DiagnosisCache {
    pub fn get(&self, log_hash: &str) -> Option<&String> {
        self.cache.get(log_hash)
    }

    pub fn insert(&mut self, log_hash: String, diagnosis: String) {
        // Keep cache bounded
        if self.cache.len() > 50 {
            self.cache.clear();
        }
        self.cache.insert(log_hash, diagnosis);
    }
}

Enter fullscreen mode Exit fullscreen mode

50-entry cache. Old entries cleared when full. Simple and effective.


The "bring your own key" model

Asking users to get their own free API key has an unexpected benefit: they feel ownership over the AI feature.

"I set up my own Gemini key" is a different mental model than "the app has AI built in." Users who set up their own key understand the rate limits, understand that it's a free service, and are more forgiving when it's occasionally slow.

The setup friction (2 minutes to get a key from Google AI Studio) filters out users who aren't genuinely interested in the feature.


What free tier can't do

Batch processing. If a user wants to run AI analysis on 100 error lines at once, the free tier can't handle it without significant delay. Design your UI to discourage bulk AI use.

Real-time analysis. Don't try to send every log line to the API as it streams in. That's a 15 RPM limit burning in seconds.

High-volume production use. If your tool gets popular and thousands of users are hitting Gemini simultaneously, they share the free tier limits per API key. This is actually fine — each user has their own key.


The honest verdict

For a developer tool with intermittent AI use: the free tier is completely sufficient. I've never hit the daily limit in normal use.

Build for the free tier from day one. It keeps your costs at zero, your users' costs at zero, and forces you to design efficient AI interactions rather than spamming the API.


HiyokoLogcat is free and open source → github.com/hiyoyok/HiyokoLogcat
X → @hiyoyok