惯性聚合 高效追踪和阅读你感兴趣的博客、新闻、科技资讯
阅读原文 在惯性聚合中打开

推荐订阅源

雷峰网
雷峰网
爱范儿
爱范儿
宝玉的分享
宝玉的分享
Apple Machine Learning Research
Apple Machine Learning Research
博客园 - Franky
freeCodeCamp Programming Tutorials: Python, JavaScript, Git & More
OSCHINA 社区最新新闻
OSCHINA 社区最新新闻
博客园 - 三生石上(FineUI控件)
人人都是产品经理
人人都是产品经理
阮一峰的网络日志
阮一峰的网络日志
钛媒体:引领未来商业与生活新知
钛媒体:引领未来商业与生活新知
Last Week in AI
Last Week in AI
博客园 - 聂微东
大猫的无限游戏
大猫的无限游戏
让小产品的独立变现更简单 - ezindie.com
让小产品的独立变现更简单 - ezindie.com
罗磊的独立博客
博客园 - 叶小钗
WordPress大学
WordPress大学
奇客Solidot–传递最新科技情报
奇客Solidot–传递最新科技情报
酷 壳 – CoolShell
酷 壳 – CoolShell
小众软件
小众软件
博客园 - 司徒正美
博客园 - 【当耐特】
IT之家
IT之家

Show HN

GitHub - astefanutti/shaderbang: Shebang for Shaders Show HN: Generate Claude Code Workflows using Spec Driven Development approach Show HN: AI agents for UK GDAD PCF roles and their skills The Two Pillars: Mixer Mode and Meta-Software in the Reorganization of Software Work After AI GitHub - JaiCode08/teleport-env What 1,000+ Harness Experiments Taught Me About Self-Improving Agents Show HN: Liiists, a Markdown-first, iOS and CLI list app SwiperTab – Get this Extension for 🦊 Firefox (en-US) GitHub - kouhxp/fftext: Summarize, explain, fact-check, or translate any text, URL, or file. No GPU. No cloud. One command GitHub - sweetpad-dev/sweetpad: Develop Swift/iOS projects using VSCode GitHub - dogmaticdev/IRON: IRON a.k.a. Intermediate Representation Object Notation is a Interpreter/Database that is used to create Programming Languages. GitHub - sjhalani7/vaen: Package your AI coding harness into a portable .agent file, and share it across repos, teams, & the community without ever having to copy-paste instructions, skills, MCP config, or secrets. Show HN: Gandalf the Grader Show HN: Citadeld – replay any CI failure locally from a single file GitHub - tdortman/cuSBF: High-Performance GPU Super Bloom Filter coral-ai/claude-code-token-xray at main · Coral-Bricks-AI/coral-ai GitHub - ulyssestenn/funes: Funes is a Git-based framework for LLM-managed knowledge work: an AI Librarian ingests raw sources, builds an interlinked Markdown knowledge base, and uses it to produce cited reports, analyses, and other outputs. GitHub - ThatXliner/gah: Git Add Hunk, built for agents to use GitHub - harmont-dev/harmont-cli: Command-line client for the Harmont CI platform GitHub - brooksmcmillin/mcp-authflow: OAuth 2.0 Authorization Server framework for MCP servers GitHub - javaid-codes/audit-supply-chain-agents GitHub - amorey/gochan: A small library of common channel architectures for Go, inspired by Rust GitHub - arifozgun/OpenGem: Free, Open-Source AI API Gateway with Gemini, OpenAI & Anthropic Compatibility in 1 file GitHub - Pranesh950/BioPetals: 🌸 Run BIOxAI models at home, BitTorrent-style. Fine-tuning and inference up to 10x faster than offloading GitHub - cnguyen14/bounty-doctor: Diagnose a GitHub bounty issue before you waste hours: detects honeypot scam repos, AI-bot attempt swarms, and stale contests. Show HN: CoreMCP – MCP Server for On-Prem DBs Show HN: KittyHTML – Render HTML/CSS as an inline image in your terminal GitHub - bingud/filemat: Web-based file manager Show HN: TruthLens – Free multi-signal deepfake image detector GitHub - apexlocal-jz/claude-usage-tray: Windows system-tray app showing your Claude Code rate-limit usage at a glance. Zero deps, ~300 lines of PowerShell. Cross-IDE (works regardless of VS Code, Cursor, plain terminal).
Classer - High-performance AI classification
wdrezek · 2026-06-11 · via Show HN

High-performance
AI classification

Beats GPT-5.4 accuracy · up to 100x cheaper · real-time latency

No prompt engineeringNo training required

Text

Labels

billing×technical_support×sales×spam×

Built for our own apps. Now open to everyone.

The Problem

You're overpaying for most AI tasks

You need to sort a support ticket. Detect spam. Route a phone call. Tag a product image. So you:

1

Iterate Prompts

Rewrite "be accurate" 47 ways until the model stops making things up.

2

Debug Schema Violations

Catch the 15% of responses that ignore your output format.

3

Face the Bill

Watch costs explode when you scale past demo.

It works. But it's slow, expensive, and embarrassingly over-engineered for a task that should take milliseconds.

The Solution

A dedicated engine for classification at scale

We stripped away the "chat" and let the intelligence focus on: turning messy data into accurate labels.

1

No more prompt engineering.

Provide your labels and let the engine auto-calibrate. No "Act as" fluff, no manual tweaking.

2

Zero schema violations.

Pure classification means zero hallucinations. Get the right format, every single time.

3

10x lower overhead.

Scale without the "LLM tax." Built for high-volume apps where speed and margins matter.

It's precise. It's predictable. It's the specialized infrastructure for the 90% of AI tasks that don't need a chat interface.

Comparison

How we compare to General-Purpose LLMs

Benchmarks

Tested on 33 public datasets

Classer beats GPT-5.4-mini on the top classification benchmarks — with zero training data.

See all 33 benchmarks

The Journey

Start in 60 seconds. Improve without ML engineers.

1

Zero-shot

Just pass your labels. It works out of the box.

2

Monitor

See every prediction in your console. Inspect confidence scores. Spot edge cases.

3

Correct

Add class descriptions. Label a few examples, or let a high-reasoning LLM do it automatically.

4

Auto-improve

Enable auto-calibration. The system distills your data into a custom model that lives in your account.

You stay focused on your product. The model gets smarter in the background.

PRICING

Predictable, low-cost AI pricing

Pay for input tokens only. Choose the priority tier and save up to 8x vs public LLMs.

Free tier: 10M tokens/mo free - no credit card required

Enterprise: Need volume pricing or dedicated infra? Contact us

Fast Batch Processing

Millions of results in 15 minutes

No concurrency scriptsNo retry loopsNo rate limit workaroundsNo 24-hour waits

Just upload a file — up to 50 million rows per job — and get labeled results back at whatever scale your pipeline needs.

Lowest cost

If you don't need instant answers, Batch is the cheapest way to run Classer

See batch docs

Generative LLMs

OpenAI-compatible inference, your prices

Point any OpenAI SDK at api.classer.ai/v1 and pick the model that fits.

deepseek-ai/DeepSeek-V4-Flash

DeepSeek V4 Flash

1M context · text

ToolsJSON mode

Input$0.10Output$0.28

per 1M tokens

Open-weights workhorse for high-volume, cost-sensitive workloads.

Qwen/Qwen3.6-35B-A3B

Qwen 3.6 35B-A3B

256K context · text + vision

ToolsJSONStructuredVision

Input$0.15Cache read$0.05Output$1.00

per 1M tokens

Multimodal MoE with structured outputs and per-token cache discount.

moonshotai/Kimi-K2.6

Kimi K2.6

256K context · text + vision · reasoning

ReasoningToolsJSONStructuredVision

Input$0.70Output$4.00

per 1M tokens

Reasoning model with explicit thinking traces and image input.

View generative LLM details

FAQ

Frequently Asked Questions

Stop burning money

Get your API key in 30 seconds. First 10M tokens free.

Start Saving