惯性聚合 高效追踪和阅读你感兴趣的博客、新闻、科技资讯
阅读原文 在惯性聚合中打开

推荐订阅源

Engineering at Meta
Engineering at Meta
博客园_首页
J
Java Code Geeks
Jina AI
Jina AI
B
Blog RSS Feed
量子位
有赞技术团队
有赞技术团队
M
MIT News - Artificial intelligence
L
LangChain Blog
Microsoft Security Blog
Microsoft Security Blog
小众软件
小众软件
博客园 - 聂微东
月光博客
月光博客
奇客Solidot–传递最新科技情报
奇客Solidot–传递最新科技情报
博客园 - 三生石上(FineUI控件)
Last Week in AI
Last Week in AI
MongoDB | Blog
MongoDB | Blog
I
InfoQ
罗磊的独立博客
H
Hackread – Cybersecurity News, Data Breaches, AI and More
爱范儿
爱范儿
Y
Y Combinator Blog
Vercel News
Vercel News
雷峰网
雷峰网

DEV Community

Authentication Security Deep Dive: From Brute Force to Salted Hashing (With Java Examples) Why AI Systems Don’t Fail — They Drift Spilling beans for how i learn for exam😁"Reinforcement Learning Cheat Sheet" I Replaced Chrome with Safari for AI Browser Automation. Here's What Broke (and What Finally Worked) How Python Borrows Other People's Work The $40 Architecture: Processing 1 Billion API Requests with 99.99% Uptime Vibe Coding: A Workflow Guide (From Zero to SaaS) Most webhook security guides protect the wrong side. The scary part is delivery. Headless CMS for TanStack Start: Build a Blog with Cosmic EU Age Verification App "Hacked in 2 Minutes" — What Actually Happened Comfy Cloud’s delete function does not actually remove files Running AI Models on GPU Cloud Servers: A Beginner Guide Event-driven media intelligence with AWS Step Functions and Bedrock I scored 500 AI prompts across 8 quality dimensions — here's what broke How to Call Google Gemini API from Next.js (Free Tier, No Backend Needed) The Portal Protocol: Reclaiming Human Connection in the Age of AI How to Fix Your Team's Scattered Knowledge Problem With a Self-Hosted Forum Intro to tc Cloud Functors: A Graph-First Mental Model for the Modern Cloud Designing Multi-Tenant Backends With Both Ownership and Team Access I Built a Neumorphic CSS Library with 77+ Components — Here's What I Learned PostgreSQL Performance Optimization: Why Connection Pooling Is Critical at Scale Cómo construí un SaaS multi-rubro para gestionar expensas en Argentina con FastAPI + Vue 3 🚀 I Built an Ethical Hacking Scanner Tool – Open Source Project I Replaced /usage and /context in Claude Code With a Single Statusline A Pythonic Way to Handle Emails (IMAP/SMTP) with Auto-Discovery and AI-Ready Design I Collected 8.9 Million Polymarket Price Points — Here's What I Found About How Markets Really Move EcoTrack AI — Carbon Footprint Tracker & Dashboard Everyone's Using AI. No One Agrees How. 5 self-hosted ebook managers worth trying in 2026 Building Your First AI Agent with LangChain: From Chatbot to Autonomous Assistant
Shipping FreeSay to 2GB-RAM Phones: What We Cut, What We ...
FreeSay · 2026-04-25 · via DEV Community

FreeSay

I built FreeSay — an AI-powered speaking tutor — and a surprising amount of engineering time went into making it run well on cheap phones. This post is about what got cut and what survived, targeting devices with 2GB of RAM and flaky connections.

Why this matters

Most of FreeSay's audience lives in markets where the median phone is not flagship-class. If the app chokes on a Samsung A03 or a Redmi 9A, we lose the exact users we were built for. So "runs on 2GB" was a hard constraint from day one, not a post-launch optimization.

What we kept

  • Real-time LLM-backed conversation in 15 target languages. Every turn goes to the cloud for quality; we gave up on fully offline speech-to-text because the accuracy gap was too large for beginners.
  • Cloud TTS for the tutor voice. We tested Piper on-device for Android — the voice-quality drop was too jarring relative to the APK bloat.
  • Aggressive per-turn caching on the server. Common corrections, translations, and vocabulary lookups are memoized so repeat learners pay close-to-zero latency on overlap.
  • A bare-metal server in Korea instead of serverless. Regional subscription pricing cannot survive Lambda bills once conversation volume grows.

What we cut

  • Heavy onboarding animations. Replaced with a single static screen and a play button. Every frame skipped was a frame the GPU did not have to allocate.
  • Rich-text chat bubbles. We tried markdown rendering in the chat log, then fell back to plain text with a handful of explicit highlight types — correction, new word, translation.
  • Pre-downloaded lesson content. Everything is fetched on demand; the APK ships small, and the first launch only downloads what the user actually opens.
  • Optional in-app video demos. For slow connections, a 30-second video was the difference between "intrigued" and "gave up." We replaced them with text + a single still frame.

The stack, briefly

React Native for a single codebase across iOS and Android. On-demand correction via LLM calls. Cloud TTS. A Puppeteer pipeline for localized Play Store / App Store screenshots.

What I would do differently

If I were starting again, I would benchmark the APK size and cold-start time on a 2GB device before writing a single feature. Almost every hard call we made later — TTS source, animation budget, bundle splitting — came back to that one number: how long until the learner can speak their first sentence?

Try it

Landing page: https://fasterwork.net/freesay/

Feedback from anyone who has shipped consumer apps to low-end Android welcome — especially on the APK-size vs feature-richness tradeoff.