惯性聚合 高效追踪和阅读你感兴趣的博客、新闻、科技资讯
阅读原文 在惯性聚合中打开

推荐订阅源

Microsoft Security Blog
Microsoft Security Blog
J
Java Code Geeks
GbyAI
GbyAI
aimingoo的专栏
aimingoo的专栏
L
LangChain Blog
I
InfoQ
D
Docker
F
Fortinet All Blogs
Y
Y Combinator Blog
Martin Fowler
Martin Fowler
月光博客
月光博客
B
Blog
Engineering at Meta
Engineering at Meta
T
Tailwind CSS Blog
罗磊的独立博客
博客园_首页
G
Google Developers Blog
Stack Overflow Blog
Stack Overflow Blog
Recent Announcements
Recent Announcements
D
DataBreaches.Net
freeCodeCamp Programming Tutorials: Python, JavaScript, Git & More
B
Blog RSS Feed
IT之家
IT之家
V
V2EX

Show HN

GitHub - astefanutti/shaderbang: Shebang for Shaders Show HN: Generate Claude Code Workflows using Spec Driven Development approach Show HN: AI agents for UK GDAD PCF roles and their skills The Two Pillars: Mixer Mode and Meta-Software in the Reorganization of Software Work After AI GitHub - JaiCode08/teleport-env What 1,000+ Harness Experiments Taught Me About Self-Improving Agents Show HN: Liiists, a Markdown-first, iOS and CLI list app SwiperTab – Get this Extension for 🦊 Firefox (en-US) GitHub - kouhxp/fftext: Summarize, explain, fact-check, or translate any text, URL, or file. No GPU. No cloud. One command GitHub - sweetpad-dev/sweetpad: Develop Swift/iOS projects using VSCode GitHub - dogmaticdev/IRON: IRON a.k.a. Intermediate Representation Object Notation is a Interpreter/Database that is used to create Programming Languages. GitHub - sjhalani7/vaen: Package your AI coding harness into a portable .agent file, and share it across repos, teams, & the community without ever having to copy-paste instructions, skills, MCP config, or secrets. Show HN: Gandalf the Grader Show HN: Citadeld – replay any CI failure locally from a single file GitHub - tdortman/cuSBF: High-Performance GPU Super Bloom Filter coral-ai/claude-code-token-xray at main · Coral-Bricks-AI/coral-ai GitHub - ulyssestenn/funes: Funes is a Git-based framework for LLM-managed knowledge work: an AI Librarian ingests raw sources, builds an interlinked Markdown knowledge base, and uses it to produce cited reports, analyses, and other outputs. GitHub - ThatXliner/gah: Git Add Hunk, built for agents to use GitHub - harmont-dev/harmont-cli: Command-line client for the Harmont CI platform GitHub - brooksmcmillin/mcp-authflow: OAuth 2.0 Authorization Server framework for MCP servers GitHub - javaid-codes/audit-supply-chain-agents GitHub - amorey/gochan: A small library of common channel architectures for Go, inspired by Rust GitHub - arifozgun/OpenGem: Free, Open-Source AI API Gateway with Gemini, OpenAI & Anthropic Compatibility in 1 file GitHub - Pranesh950/BioPetals: 🌸 Run BIOxAI models at home, BitTorrent-style. Fine-tuning and inference up to 10x faster than offloading GitHub - cnguyen14/bounty-doctor: Diagnose a GitHub bounty issue before you waste hours: detects honeypot scam repos, AI-bot attempt swarms, and stale contests. Show HN: CoreMCP – MCP Server for On-Prem DBs Show HN: KittyHTML – Render HTML/CSS as an inline image in your terminal GitHub - bingud/filemat: Web-based file manager Show HN: TruthLens – Free multi-signal deepfake image detector GitHub - apexlocal-jz/claude-usage-tray: Windows system-tray app showing your Claude Code rate-limit usage at a glance. Zero deps, ~300 lines of PowerShell. Cross-IDE (works regardless of VS Code, Cursor, plain terminal).
Classer - High-performance AI classification
wdrezek · 2026-06-11 · via Show HN

High-performance
AI classification

Beats GPT-5.4 accuracy · up to 100x cheaper · real-time latency

No prompt engineeringNo training required

Text

Labels

billing×technical_support×sales×spam×

Built for our own apps. Now open to everyone.

The Problem

You're overpaying for most AI tasks

You need to sort a support ticket. Detect spam. Route a phone call. Tag a product image. So you:

1

Iterate Prompts

Rewrite "be accurate" 47 ways until the model stops making things up.

2

Debug Schema Violations

Catch the 15% of responses that ignore your output format.

3

Face the Bill

Watch costs explode when you scale past demo.

It works. But it's slow, expensive, and embarrassingly over-engineered for a task that should take milliseconds.

The Solution

A dedicated engine for classification at scale

We stripped away the "chat" and let the intelligence focus on: turning messy data into accurate labels.

1

No more prompt engineering.

Provide your labels and let the engine auto-calibrate. No "Act as" fluff, no manual tweaking.

2

Zero schema violations.

Pure classification means zero hallucinations. Get the right format, every single time.

3

10x lower overhead.

Scale without the "LLM tax." Built for high-volume apps where speed and margins matter.

It's precise. It's predictable. It's the specialized infrastructure for the 90% of AI tasks that don't need a chat interface.

Comparison

How we compare to General-Purpose LLMs

Benchmarks

Tested on 33 public datasets

Classer beats GPT-5.4-mini on the top classification benchmarks — with zero training data.

See all 33 benchmarks

The Journey

Start in 60 seconds. Improve without ML engineers.

1

Zero-shot

Just pass your labels. It works out of the box.

2

Monitor

See every prediction in your console. Inspect confidence scores. Spot edge cases.

3

Correct

Add class descriptions. Label a few examples, or let a high-reasoning LLM do it automatically.

4

Auto-improve

Enable auto-calibration. The system distills your data into a custom model that lives in your account.

You stay focused on your product. The model gets smarter in the background.

PRICING

Predictable, low-cost AI pricing

Pay for input tokens only. Choose the priority tier and save up to 8x vs public LLMs.

Free tier: 10M tokens/mo free - no credit card required

Enterprise: Need volume pricing or dedicated infra? Contact us

Fast Batch Processing

Millions of results in 15 minutes

No concurrency scriptsNo retry loopsNo rate limit workaroundsNo 24-hour waits

Just upload a file — up to 50 million rows per job — and get labeled results back at whatever scale your pipeline needs.

Lowest cost

If you don't need instant answers, Batch is the cheapest way to run Classer

See batch docs

Generative LLMs

OpenAI-compatible inference, your prices

Point any OpenAI SDK at api.classer.ai/v1 and pick the model that fits.

deepseek-ai/DeepSeek-V4-Flash

DeepSeek V4 Flash

1M context · text

ToolsJSON mode

Input$0.10Output$0.28

per 1M tokens

Open-weights workhorse for high-volume, cost-sensitive workloads.

Qwen/Qwen3.6-35B-A3B

Qwen 3.6 35B-A3B

256K context · text + vision

ToolsJSONStructuredVision

Input$0.15Cache read$0.05Output$1.00

per 1M tokens

Multimodal MoE with structured outputs and per-token cache discount.

moonshotai/Kimi-K2.6

Kimi K2.6

256K context · text + vision · reasoning

ReasoningToolsJSONStructuredVision

Input$0.70Output$4.00

per 1M tokens

Reasoning model with explicit thinking traces and image input.

View generative LLM details

FAQ

Frequently Asked Questions

Stop burning money

Get your API key in 30 seconds. First 10M tokens free.

Start Saving