惯性聚合 高效追踪和阅读你感兴趣的博客、新闻、科技资讯
阅读原文 在惯性聚合中打开

推荐订阅源

D
DataBreaches.Net
IT之家
IT之家
The Cloudflare Blog
Apple Machine Learning Research
Apple Machine Learning Research
WordPress大学
WordPress大学
N
Netflix TechBlog - Medium
阮一峰的网络日志
阮一峰的网络日志
P
Proofpoint News Feed
L
LangChain Blog
博客园 - Franky
美团技术团队
J
Java Code Geeks
Microsoft Security Blog
Microsoft Security Blog
博客园 - 叶小钗
小众软件
小众软件
Y
Y Combinator Blog
B
Blog RSS Feed
Cyber Security Advisories - MS-ISAC
Cyber Security Advisories - MS-ISAC
D
Docker
Hugging Face - Blog
Hugging Face - Blog
Jina AI
Jina AI
罗磊的独立博客
大猫的无限游戏
大猫的无限游戏
Vercel News
Vercel News

Hacker News: Show HN

PurrrrrFocus: Pomodoro Timer App - App Store Workflow Engine — Multi-Step Orchestration for Bun RapidPhoto: Pro Photo Editor App - App Store GitHub - DheerG/swarms: Achieve extraordinary results with claude code across a variety of tasks SPICE simulation → oscilloscope → verification with Claude Code — Lucas Gerads Show HN: VCoding – A 5 MB native Windows IDE with no dynamic dependencies Show HN: LLMs don't hallucinate because they're bad at math, it's the format GitHub - Agent-FM/agentfm-core: AgentFM is a peer-to-peer network that turns everyday computers into a decentralized AI supercomputer. AgentFM lets you run massive AI workloads directly across a global mesh of idle CPUs and GPUs. Show HN: Tracking Top US Science Olympiad Alumni over Last 25 Years GitHub - Potarix/agent-hub: One place to talk to all your agents Show HN: Runtime security for AI agents(injection,tool abuse, data exfiltration) GitHub - dubeyKartikay/lazyspotify: Terminal Spotify client for macOS and Linux GitHub - the-banana-tool/king-louie: Easy to use GUI Personal AI Assistant. Win/Linux/Mac. Show HN I made my vacation rental bookable by AI agents–no Airbnb, 0% commission GitHub - basteez/jsf-autoreload: maven plugin to enable hot reload on jsf projects uvm32/hosts/host-gdbstub at main · ringtailsoftware/uvm32 GitHub - labsai/EDDI: Config-driven engine that turns JSON into production-grade AI agents. Multi-agent orchestration, 12+ LLM providers, MCP/A2A protocols, RAG, persistent memory, and enterprise compliance (EU AI Act, GDPR, HIPAA). Built on Quarkus. GitHub - glitchnsec/fortyone-oss: AI Executive Assistant Platform Quickstart | Alien GitHub - muxshed/shed: One stream in, or many. Every destination, simultaneously. No cloud middleman, no per-channel fees, no limits. GitHub - ocrbase-hq/ocrbase: 📄 PDF/IMG ->.MD/JSON Document OCR API for PaddleOCR and GLMOCR. Self-hostable. GitHub - impactjo/home-memory: MCP server that lets your AI assistant remember everything about your home. GitHub - Sets88/dbcls: DbCls is a powerful terminal database client that supports various databases GitHub - neptun2000/heor-agent-mcp GitHub - SeanFDZ/macmind: Single-layer transformer in HyperTalk for the classic Macintosh RollQuation: Math Puzzles - Apps on Google Play GitHub - dropbox/witchcraft Show HN: Agent-cache – Multi-tier LLM/tool/session caching for Valkey and Redis GitHub - opentalon/opentalon: OpenTalon is an open-source platform built from the ground up in Go as a robust alternative to OpenClaw LinkedIn™ 职位抓取工具 - Chrome 应用商店
Classer - High-performance AI classification
wdrezek · 2026-06-11 · via Hacker News: Show HN

High-performance
AI classification

Beats GPT-5.4 accuracy · up to 100x cheaper · real-time latency

No prompt engineeringNo training required

Text

Labels

billing×technical_support×sales×spam×

Built for our own apps. Now open to everyone.

The Problem

You're overpaying for most AI tasks

You need to sort a support ticket. Detect spam. Route a phone call. Tag a product image. So you:

1

Iterate Prompts

Rewrite "be accurate" 47 ways until the model stops making things up.

2

Debug Schema Violations

Catch the 15% of responses that ignore your output format.

3

Face the Bill

Watch costs explode when you scale past demo.

It works. But it's slow, expensive, and embarrassingly over-engineered for a task that should take milliseconds.

The Solution

A dedicated engine for classification at scale

We stripped away the "chat" and let the intelligence focus on: turning messy data into accurate labels.

1

No more prompt engineering.

Provide your labels and let the engine auto-calibrate. No "Act as" fluff, no manual tweaking.

2

Zero schema violations.

Pure classification means zero hallucinations. Get the right format, every single time.

3

10x lower overhead.

Scale without the "LLM tax." Built for high-volume apps where speed and margins matter.

It's precise. It's predictable. It's the specialized infrastructure for the 90% of AI tasks that don't need a chat interface.

Comparison

How we compare to General-Purpose LLMs

Benchmarks

Tested on 33 public datasets

Classer beats GPT-5.4-mini on the top classification benchmarks — with zero training data.

See all 33 benchmarks

The Journey

Start in 60 seconds. Improve without ML engineers.

1

Zero-shot

Just pass your labels. It works out of the box.

2

Monitor

See every prediction in your console. Inspect confidence scores. Spot edge cases.

3

Correct

Add class descriptions. Label a few examples, or let a high-reasoning LLM do it automatically.

4

Auto-improve

Enable auto-calibration. The system distills your data into a custom model that lives in your account.

You stay focused on your product. The model gets smarter in the background.

PRICING

Predictable, low-cost AI pricing

Pay for input tokens only. Choose the priority tier and save up to 8x vs public LLMs.

Free tier: 10M tokens/mo free - no credit card required

Enterprise: Need volume pricing or dedicated infra? Contact us

Fast Batch Processing

Millions of results in 15 minutes

No concurrency scriptsNo retry loopsNo rate limit workaroundsNo 24-hour waits

Just upload a file — up to 50 million rows per job — and get labeled results back at whatever scale your pipeline needs.

Lowest cost

If you don't need instant answers, Batch is the cheapest way to run Classer

See batch docs

Generative LLMs

OpenAI-compatible inference, your prices

Point any OpenAI SDK at api.classer.ai/v1 and pick the model that fits.

deepseek-ai/DeepSeek-V4-Flash

DeepSeek V4 Flash

1M context · text

ToolsJSON mode

Input$0.10Output$0.28

per 1M tokens

Open-weights workhorse for high-volume, cost-sensitive workloads.

Qwen/Qwen3.6-35B-A3B

Qwen 3.6 35B-A3B

256K context · text + vision

ToolsJSONStructuredVision

Input$0.15Cache read$0.05Output$1.00

per 1M tokens

Multimodal MoE with structured outputs and per-token cache discount.

moonshotai/Kimi-K2.6

Kimi K2.6

256K context · text + vision · reasoning

ReasoningToolsJSONStructuredVision

Input$0.70Output$4.00

per 1M tokens

Reasoning model with explicit thinking traces and image input.

View generative LLM details

FAQ

Frequently Asked Questions

Stop burning money

Get your API key in 30 seconds. First 10M tokens free.

Start Saving