惯性聚合 高效追踪和阅读你感兴趣的博客、新闻、科技资讯
阅读原文 在惯性聚合中打开

推荐订阅源

GbyAI
GbyAI
B
Blog
Stack Overflow Blog
Stack Overflow Blog
量子位
freeCodeCamp Programming Tutorials: Python, JavaScript, Git & More
T
Tailwind CSS Blog
MongoDB | Blog
MongoDB | Blog
小众软件
小众软件
博客园 - 三生石上(FineUI控件)
Recent Announcements
Recent Announcements
U
Unit 42
奇客Solidot–传递最新科技情报
奇客Solidot–传递最新科技情报
腾讯CDC
D
DataBreaches.Net
Microsoft Azure Blog
Microsoft Azure Blog
G
Google Developers Blog
M
MIT News - Artificial intelligence
P
Proofpoint News Feed
罗磊的独立博客
L
LangChain Blog
V
Visual Studio Blog
雷峰网
雷峰网
aimingoo的专栏
aimingoo的专栏
宝玉的分享
宝玉的分享

DEV Community

Authentication Security Deep Dive: From Brute Force to Salted Hashing (With Java Examples) Why AI Systems Don’t Fail — They Drift Spilling beans for how i learn for exam😁"Reinforcement Learning Cheat Sheet" I Replaced Chrome with Safari for AI Browser Automation. Here's What Broke (and What Finally Worked) How Python Borrows Other People's Work The $40 Architecture: Processing 1 Billion API Requests with 99.99% Uptime Vibe Coding: A Workflow Guide (From Zero to SaaS) Most webhook security guides protect the wrong side. The scary part is delivery. Headless CMS for TanStack Start: Build a Blog with Cosmic EU Age Verification App "Hacked in 2 Minutes" — What Actually Happened Comfy Cloud’s delete function does not actually remove files Running AI Models on GPU Cloud Servers: A Beginner Guide Event-driven media intelligence with AWS Step Functions and Bedrock I scored 500 AI prompts across 8 quality dimensions — here's what broke How to Call Google Gemini API from Next.js (Free Tier, No Backend Needed) The Portal Protocol: Reclaiming Human Connection in the Age of AI How to Fix Your Team's Scattered Knowledge Problem With a Self-Hosted Forum Intro to tc Cloud Functors: A Graph-First Mental Model for the Modern Cloud Designing Multi-Tenant Backends With Both Ownership and Team Access I Built a Neumorphic CSS Library with 77+ Components — Here's What I Learned PostgreSQL Performance Optimization: Why Connection Pooling Is Critical at Scale Cómo construí un SaaS multi-rubro para gestionar expensas en Argentina con FastAPI + Vue 3 🚀 I Built an Ethical Hacking Scanner Tool – Open Source Project I Replaced /usage and /context in Claude Code With a Single Statusline A Pythonic Way to Handle Emails (IMAP/SMTP) with Auto-Discovery and AI-Ready Design I Collected 8.9 Million Polymarket Price Points — Here's What I Found About How Markets Really Move EcoTrack AI — Carbon Footprint Tracker & Dashboard Everyone's Using AI. No One Agrees How. 5 self-hosted ebook managers worth trying in 2026 Building Your First AI Agent with LangChain: From Chatbot to Autonomous Assistant
Why our B2B matching engine is a pure function, not an ML...
Anatolii Nikolskii · 2026-06-20 · via DEV Community

Anatolii Nikolskii

When you build a marketplace that connects two sides, manufacturers on one side and distributors on the other, the first instinct is often to reach for machine learning to rank matches. We went the other way. The core of Hell of a Partner is a deterministic scoring function, and that choice has paid off in ways I did not expect.

The shape

The whole matcher is one pure function:

scoreMatch(offer, profile) => {
  score: number,      // 0 to 100
  tier: "strong" | "fair" | "weak",
  rationale: Reason[] // one line per dimension
}

It scores a pair across seven weighted dimensions (category fit, country and trade bloc, certifications, capacity, and so on). Each dimension is a small scorer that returns a value from 0 to 1, then we take a weighted sum.

const MATCH_WEIGHTS = {
  category: 0.30,
  geography: 0.20,
  tradeBloc: 0.15,
  certifications: 0.15,
  // ...
}

Why deterministic beats a model here

It is explainable. Every score comes with a rationale, for example "category match, both in food and beverage" or "no shared trade bloc". A buyer sees why a supplier was suggested. With a black box ranker you get a number and a shrug.

It is testable. Our tests are plain scripts that assert invariants: same input, same output, every time. No flaky thresholds, no drift, no retraining. A pull request that changes a weight shows up as a clean diff in expected scores.

It is cheap and instant. No inference cost, no extra API latency, no cold start. The function runs inline with the request.

It works on day one. A learned ranker needs interaction data you do not have at launch. A scoring function encodes domain knowledge directly, so it is useful before you have a single click to learn from.

The tradeoff

The cost is that the domain knowledge lives in the weights, and you tune them by hand. That is fine while the rules stay legible and the team can reason about them. The day there is enough real interaction data, the deterministic scorer becomes a strong baseline and a source of features, not something to throw away.

Takeaway

Reach for the pure function first. Make the thing explainable, testable, and cheap, then add learning later when you have data and a baseline to beat.

If you want to see it in action, the matcher powers the supplier and distributor suggestions on Hell of a Partner.