惯性聚合 高效追踪和阅读你感兴趣的博客、新闻、科技资讯
阅读原文 在惯性聚合中打开

推荐订阅源

J
Java Code Geeks
Google DeepMind News
Google DeepMind News
奇客Solidot–传递最新科技情报
奇客Solidot–传递最新科技情报
小众软件
小众软件
Blog — PlanetScale
Blog — PlanetScale
腾讯CDC
A
About on SuperTechFans
Vercel News
Vercel News
I
InfoQ
阮一峰的网络日志
阮一峰的网络日志
月光博客
月光博客
钛媒体:引领未来商业与生活新知
钛媒体:引领未来商业与生活新知
人人都是产品经理
人人都是产品经理
S
SegmentFault 最新的问题
V
Visual Studio Blog
T
Tailwind CSS Blog
大猫的无限游戏
大猫的无限游戏
M
MIT News - Artificial intelligence
博客园 - 【当耐特】
freeCodeCamp Programming Tutorials: Python, JavaScript, Git & More
Microsoft Azure Blog
Microsoft Azure Blog
Apple Machine Learning Research
Apple Machine Learning Research
GbyAI
GbyAI
美团技术团队

Hacker News: Show HN

PurrrrrFocus: Pomodoro Timer App - App Store Workflow Engine — Multi-Step Orchestration for Bun RapidPhoto: Pro Photo Editor App - App Store GitHub - DheerG/swarms: Achieve extraordinary results with claude code across a variety of tasks SPICE simulation → oscilloscope → verification with Claude Code — Lucas Gerads Show HN: VCoding – A 5 MB native Windows IDE with no dynamic dependencies Show HN: LLMs don't hallucinate because they're bad at math, it's the format GitHub - Agent-FM/agentfm-core: AgentFM is a peer-to-peer network that turns everyday computers into a decentralized AI supercomputer. AgentFM lets you run massive AI workloads directly across a global mesh of idle CPUs and GPUs. Show HN: Tracking Top US Science Olympiad Alumni over Last 25 Years GitHub - Potarix/agent-hub: One place to talk to all your agents Show HN: Runtime security for AI agents(injection,tool abuse, data exfiltration) GitHub - dubeyKartikay/lazyspotify: Terminal Spotify client for macOS and Linux GitHub - the-banana-tool/king-louie: Easy to use GUI Personal AI Assistant. Win/Linux/Mac. Show HN I made my vacation rental bookable by AI agents–no Airbnb, 0% commission GitHub - basteez/jsf-autoreload: maven plugin to enable hot reload on jsf projects uvm32/hosts/host-gdbstub at main · ringtailsoftware/uvm32 GitHub - labsai/EDDI: Config-driven engine that turns JSON into production-grade AI agents. Multi-agent orchestration, 12+ LLM providers, MCP/A2A protocols, RAG, persistent memory, and enterprise compliance (EU AI Act, GDPR, HIPAA). Built on Quarkus. GitHub - glitchnsec/fortyone-oss: AI Executive Assistant Platform Quickstart | Alien GitHub - muxshed/shed: One stream in, or many. Every destination, simultaneously. No cloud middleman, no per-channel fees, no limits. GitHub - ocrbase-hq/ocrbase: 📄 PDF/IMG ->.MD/JSON Document OCR API for PaddleOCR and GLMOCR. Self-hostable. GitHub - impactjo/home-memory: MCP server that lets your AI assistant remember everything about your home. GitHub - Sets88/dbcls: DbCls is a powerful terminal database client that supports various databases GitHub - neptun2000/heor-agent-mcp GitHub - SeanFDZ/macmind: Single-layer transformer in HyperTalk for the classic Macintosh RollQuation: Math Puzzles - Apps on Google Play GitHub - dropbox/witchcraft Show HN: Agent-cache – Multi-tier LLM/tool/session caching for Valkey and Redis GitHub - opentalon/opentalon: OpenTalon is an open-source platform built from the ground up in Go as a robust alternative to OpenClaw LinkedIn™ 职位抓取工具 - Chrome 应用商店
Sipp - AI inference, freshly squeezed.
jjhartmann · 2026-06-24 · via Hacker News: Show HN

Blazing-fast WebGPU runtime · Open Source

The fastest runtime for the web. Run models right in the browser, zero install, and zero dependencies. Build games, agents, vision, and chat. Add a secure cloud gateway when you need it, all from one client.

Open sourceZero installLocal + gateway

SippPink Lemonade Ed.

Nutrition Facts

Serving size: 1 Sipp Client

Amount per pour

  • Dependencies0
  • Cold start< 1000ms
  • EngineRust · C++ · GGML
  • BackendBrowser-nativeWebGPU

% Dev Value

  • Open source100%
  • Type-safe100%
  • Framework sludge0%

$ npm install @sipphq/sipp

Ingredients

WebGPU · WASM · Rust · C++ · GGUF · TypeScript · 100% real tokens. Contains no frameworks, no concentrate, no added sugar.

Taste test

Easy Setup.
One Simple API.

Manage and query multiple inference endpoints through a single, unified API. Switch or split traffic between local browser execution and cloud gateways without rewriting your code.

  • Identical Code Paths
    Execute queries symmetrically across edge and cloud endpoints.
  • Multi-Endpoint Control
    Register local and remote models under one unified client.
  • Native Performance
    Tap local WebGPU execution or cloud gateways with equal ease.

Benchmark · WebGPU showdown

Same model.
Faster in the browser.

Sipp's WebGPU backend runs the same weights up to 5× faster than other browser runtimes. No native install. Pick a model and watch the multipliers stack up.

Measured on Qwen 2.5 0.5B · Q4_K_M. LILO · 1024 in / 512 out · NVIDIA 3080 · Chrome (N=3, 9 runs, 1 warmup). Multipliers show how many times faster Sipp runs vs each browser runtime.

Live demo · Fresh squeeze

Pick a model.
Sip the tokens.

A bare-bones chat running 100% in your browser. Pick a model, start the tap, and then chat. No account, no server.

Try the full demo

Built with Sipp · 100% in-browser

Pour it into
anything.

Real apps running real models with Sipp. No servers, no install, no waiting. Every one runs the model right in your browser.

Mobile support is currently being worked on. Try demos on desktop.

🪄Desktop only

PromptCast

A wizard duel where every spell is generated on the fly by a local LLM. No two casts the same.

Desktop only

🪄Live demo

PromptCast

A wizard duel where every spell is generated on the fly by a local LLM. No two casts the same.

Play in browser ›

🍌Desktop only

Banana Brawl

A swarm of little agents reason in-browser, each running a local model to pick its next move, all fighting for one banana.

Desktop only

🍌Live demo

Banana Brawl

A swarm of little agents reason in-browser, each running a local model to pick its next move, all fighting for one banana.

Play in browser ›

🎨Desktop only

Sketch Critic

Draw something and a local vision model snapshots the canvas, reads it, and gives you live feedback.

Desktop only

🎨Live demo

Sketch Critic

Draw something and a local vision model snapshots the canvas, reads it, and gives you live feedback.

Play in browser ›

💬Desktop only

Aria

Chat with a VRM character whose emotes, actions, and replies are all chosen live by a local model.

Desktop only

💬Live demo

Aria

Chat with a VRM character whose emotes, actions, and replies are all chosen live by a local model.

Play in browser ›

Fresh batch ready

Pour your first inference.

Install Sipp, run a model in your browser on WebGPU, then scale to Node, Rust, Python, or your own gateway.