惯性聚合 高效追踪和阅读你感兴趣的博客、新闻、科技资讯
阅读原文 在惯性聚合中打开

推荐订阅源

MyScale Blog
MyScale Blog
F
Full Disclosure
Microsoft Azure Blog
Microsoft Azure Blog
Jina AI
Jina AI
Recent Announcements
Recent Announcements
美团技术团队
L
LangChain Blog
P
Privacy & Cybersecurity Law Blog
M
MIT News - Artificial intelligence
www.infosecurity-magazine.com
www.infosecurity-magazine.com
W
WeLiveSecurity
Engineering at Meta
Engineering at Meta
S
Security @ Cisco Blogs
H
Hackread – Cybersecurity News, Data Breaches, AI and More
Exploit-DB.com RSS Feed
Exploit-DB.com RSS Feed
Help Net Security
Help Net Security
L
LINUX DO - 最新话题
S
Secure Thoughts
O
OpenAI News
Hacker News - Newest:
Hacker News - Newest: "LLM"
A
About on SuperTechFans
NISL@THU
NISL@THU
C
Cyber Attacks, Cyber Crime and Cyber Security
Cyberwarzone
Cyberwarzone
T
Tailwind CSS Blog
T
The Blog of Author Tim Ferriss
Recent Commits to openclaw:main
Recent Commits to openclaw:main
F
Fortinet All Blogs
B
Blog
IT之家
IT之家
T
Tor Project blog
L
Lohrmann on Cybersecurity
Webroot Blog
Webroot Blog
博客园 - 叶小钗
Simon Willison's Weblog
Simon Willison's Weblog
Application and Cybersecurity Blog
Application and Cybersecurity Blog
量子位
Security Latest
Security Latest
TaoSecurity Blog
TaoSecurity Blog
S
Schneier on Security
C
Cisco Blogs
博客园 - 三生石上(FineUI控件)
J
Java Code Geeks
Last Week in AI
Last Week in AI
The Last Watchdog
The Last Watchdog
博客园 - 聂微东
Cisco Talos Blog
Cisco Talos Blog
Security Archives - TechRepublic
Security Archives - TechRepublic
Google DeepMind News
Google DeepMind News
SecWiki News
SecWiki News

Databricks

How lakebase architecture delivers 5x faster Postgres writes Why Talent Transformation Is the Missing Focus of Enterprise AI Public Health Intelligence Shouldn't Require a Data Scientist Mean Time to Detect Is a Data Access Problem First-party audience data is the ad sales relationship now Rethinking Distributed Systems for Serverless Performance and Reliability The AI Scaling Gap Hiding in Digital Native Companies 10 trillion samples a day: Scaling beyond traditional monitoring infra at Databricks AI success starts with clean data, not just better models How nOps Rebuilt Their Cloud Optimization Platform on Databricks Lakebase, and Why Other ISVs Should Too Peril Predicts: Precision Payouts for a Volatile World The foundation of AI scalability: one team, one platform, one operating model The Federal Data Paradox: Rich in Data, Poor in Access Driving Budapest Forward: How BKK Uses Databricks to Transform City Mobility LLM Vs AI: A Practical Guide to Differences, Use Cases, and Tools Model Risk Governance Is Not the Same as Risk Intelligence Generative AI for Business: A Complete Strategy and Implementation Guide Data Science vs Data Engineering: Choosing Analysis or Infrastructure AI Applications: Tools, Use Cases, and Platforms MLOps vs DevOps: A Practical Guide for Data Scientists and IT Teams Top Data Warehouse Tools For Modern Data Analytics Unlocking SAP Business Context in Databricks with Semantic Metadata Delta Sharing The marketing activation gap has a fix: Databricks and Stitch partner to turn data infrastructure into marketing performance Alert Fatigue Is a Business Risk Backstage with Lakebase Shipping Faster isn’t Learning Faster Why Your OEE Dashboard Is Lying to You The Turbine That Tried to Tell You It Was Failing Predicting Readmissions Isn't Enough. Acting in Time Is. Clinical Trials Run Longer Than They Have To. That's a Patient Problem Network Quality Is a Revenue Problem, Not a Technical One Shelf Availability Starts with Better Demand Visibility When Predicting the Next Hit Requires More Than Intuition Approximate Answers, Exact Decisions: New Sketch Functions for Analytics Companies Winning with AI Built the Data Layer First Rethinking SQL ETL for modern data platforms Stripe data now available on Databricks via Databricks Marketplace Databricks and Stripe Projects: Infrastructure Built for Agents Agents are ready but your architecture probably isn't Interoperability Between Unity Catalog and Google BigQuery via Catalog Federation Built In, Not Bolted On: What AI-Native Actually Means in Cybersecurity Operationalizing AI for public sector fraud prevention From months to minutes: Building real-time clinical data pipelines with natural language Agentic Data Engineering with Genie Code and Lakeflow Securely send first-party conversion signals with Snapchat Conversions API on Databricks Marketplace How leading tech companies are killing the builder’s tax with Lakebase Inside one of the first production deployments of Lakebase: LangGuard's agentic workflow governance engine The next generation of Databricks Genie Model Risk Management in 2026: A Banker’s Guide to the Revised Interagency Guidance OpenAI GPT-5.5 now available on Databricks, fully-governed through Unity AI Gateway Operational databases: How they work and when to use them Databricks partners with OpenAI on GPT-5.5 Announcing the Public Preview of Lakeflow Designer Are LLM agents good at join order optimization? How conversational analytics removes the BI bottleneck How to transform document activation workflows with Genie and Agent Bricks Beyond the spreadsheet: how Databricks is delivering the modern CFO in Financial Services AI App Development: Guide To Building AI-Powered Apps IoT in Manufacturing: Strategy, Components, Use Cases, and Challenges Stop Hand-Coding Change Data Capture Pipelines Multimodal Data Integration: Production Architectures for Healthcare AI Personalization Strategies for Media Companies A Modern AI Risk Management Framework Introducing the Databricks Excel Add-in for Business Users Real-Time Decisioning for AI Agents: Why you Need a Customer Context Layer First A Practical Guide to LLM Fine Tuning AI Data Transformation Guide for Data Engineers and Data Scientists Concurrency Control in DBMS: How Locking, MVCC and Optimistic Strategies Keep Data Consistent Bridging data science and marketing: Databricks unveils Delta Sharing integration for Adobe Experience Platform and agentic marketing workflows Take Control: Customer-Managed Keys for Lakebase Postgres Get hands on with agents, vibe coding and more at Data+ AI Summit Mercedes-Benz Builds a Cross-Cloud Data Mesh with Delta Sharing and Intelligent Replication, Cutting Costs by 66% What Is a Transactional Database? Introducing Genie Agent Mode Governing coding agent sprawl with Unity AI Gateway Governing Coding Agent Sprawl with Unity AI Gateway What is pgvector? Banks Don’t Have an AI Problem – They Have a Data Platform Problem Open Platform, Unified Pipelines: Why dbt on Databricks is Accelerating Why Your Agents Can’t Read Enterprise Documents — and How to Fix It Building with Databricks Document Intelligence and Lakeflow Databricks on Google Cloud: Innovate Faster. Smarter. Together. Introducing the Databricks Connector for Google Sheets: Real-Time, Governed Lakehouse Data in the Sheets Users Love Unity AI Gateway: How to connect agents to external MCPs securely Expanding agent governance with Unity AI Gateway Agentic reasoning in practice: Making sense of structured and unstructured data Agent Bricks: The Governed Enterprise Agent Platform 8 AI and data trends shaping financial services in 2026 Lovable + Databricks: Build Data-Driven Apps at the Speed of Thought Memory scaling for AI agents Powering clinical research innovation: How TriNetX uses Databricks to accelerate drug development Database Branching in Postgres: Git-Style Workflows with Databricks Lakebase How Zalando built a unified data foundation for AI and analytics on Databricks The next era of the open lakehouse: Apache Iceberg™ v3 in Public Preview on Databricks How FSIs eliminate silos between clients, operations, and finance How MakeMyTrip achieved millisecond personalization at scale with Databricks A multi-agent approach to audience intelligence AiChemy: Next-generation agent with MCP, skills and custom data for drug discovery Accelerate business insights with Lakeflow Connect, now with a Free Tier Unlocking Next-Gen Customer Experiences with Data Intelligence for Marketing
Building real-time product search on Databricks
Jiayi Wu, Luke Lefebure, Adam Gurary · 2026-04-14 · via Databricks

Imagine you're designing a search system for an online marketplace selling cars. In milliseconds, users expect results that fit their budget, match their preferences, are available near them, and feel relevant.

That's what modern web product search looks like. It's not just a lookup tool, but a real-time decision engine that must retrieve, filter, rank, and respond almost instantly — all while balancing business and technical metrics like revenue, click-through rate, latency, and relevance.

Databricks provides the end-to-end platform for building these systems — from scalable data ingestion (Lakeflow) to vector-powered retrieval (AI Search) to real-time operational data (Lakebase) to agent-powered search experiences (Agent Bricks). This blog walks through how these pieces come together to power real-time product search.

Components for Product Search

Product search isn't simply about answering a question or surfacing information via a chatbot. It's a discovery and decision process — dynamic, personalized, and deeply tied to revenue. Buyers expect to browse, compare, and explore. The goal isn't to generate a single answer, but to present a ranked set of choices that feel relevant, trustworthy, and worth considering.

A real-time product search system generally has 3 segments (Figure 1).

  • Ingestion prepares product data for search. Product titles, descriptions, and attributes are processed, converted into embeddings, enriched with metadata, and indexed for fast retrieval.
  • Retrieval finds what could be relevant by generating a candidate set using full-text, semantic, or hybrid search combined with structured filtering.
  • Refinement determines how results should be interpreted and ordered by applying query understanding, ranking logic, personalization, and business rules.

Functional Breakdown of a Modern Search Pipeline

Figure 1: Functional Breakdown of a Modern Search Pipeline

Behind the Search Bar

None of that experience exists without strong infrastructure and meaningful metrics.

  • Infrastructure makes speed and relevance possible.
  • Metrics prove that your system is actually fast and relevant — not just on paper.

Modern product search requires both: the engineering foundation to deliver results, and the metrics discipline to continuously validate that those results are good enough.

Architecture Walkthrough

First, let's look at the architecture. Figure 2 shows a detailed example of a real-time product search architecture.

Reference Architecture for Real-Time Product Search on Databricks

Figure 2: Reference Architecture for Real-Time Product Search on Databricks

At the center of this design is Databricks AI Search, which handles ingestion, retrieval, and refinement in a single platform — eliminating the need to stitch together multiple external systems.

  • Ingestion prepares product data so it can be searched efficiently. Unstructured sources such as product listings and images are processed through scalable pipelines using Databricks Auto LoaderLakeflow Spark Declarative Pipeline and AI Functions (e.g., ai_parse_document). The data can then be chunked and converted into embeddings with metadata (e.g., car model, color, or price) in Databricks AI Search.
  • Retrieval handles real-time queries. User input is transformed into embeddings and structured filters, and Databricks AI Search retrieves the top candidates using semantic search, full-text search, or hybrid search with metadata filtering.
  • Refinement enhances retrieved candidates into final results. While retrieval provides a strong baseline, this layer refines outcomes by interpreting intent, applying ranking logic, and incorporating personalization and business rules when needed. Real-time operational context, such as session state, inventory, pricing, and user preferences, can be served via Lakebase, enabling sub-10ms low-latency signals to influence final ordering.

A few practical guidelines when building search systems on Databricks:

  • Experiment with models easily. Swap embedding models with minimal friction and leverage native reranking capabilities. Future updates will enable one-click fine-tuning of reranking models directly within the platform, simplifying relevance optimization.
  • Serve application state at the speed of search. Use Lakebase to store real-time application state — session context, inventory, pricing, user preferences — with sub-10ms latency. Managed CDC syncs Lakebase to Delta automatically, so ranking models and analytics always reflect current operational data without custom pipelines.
  • Test for scale before production. Validate latency and throughput under realistic traffic, including high-QPS scenarios. You can simulate production workloads today using search load testing notebook, with native one-click load testing support coming in a future release. For sustained traffic, leverage high QPS endpoints to handle concurrency at scale, and monitor performance through endpoint observability to track latency, throughput, and system health.
  • Build agent-ready search from day one. Every AI Search index with managed embeddings automatically gets a managed MCP server. Use it for zero-config agent integration, the VectorSearchRetrieverTool for code-first control, or point a Knowledge Assistant at your index for instant Q&A with citations — powered by Instructed Retriever, which delivers 70% better accuracy than standard RAG systems.

Metrics that Matter

A search system isn't successful because it looks elegant on a diagram. It's successful because it delivers fast, relevant results that drive business outcomes.

As shown in Figure 3, three categories of metrics help teams evaluate a search pipeline — each tied to a different layer of the system.

  • Operational metrics ensure the system is fast and reliable enough to serve users at scale. These are critical across ingestion, retrieval and refinement steps.
  • Retrieval quality metrics measure whether the system is actually retrieving and ranking relevant candidates, and are most closely tied to the retrieval and refinement stages where ranking and reranking occur.
  • User engagement metrics capture real-world behavior — whether users click, refine, or ultimately convert - providing feedback that informs improvements in retrieval and refinements over time.

Metrics Framework for Real-Time Product Search

Figure 3: Metrics Framework for Real-Time Product Search

A few practical guidelines when evaluating search systems on Databricks:

  • Balance metrics, not just optimize one. Effective search systems must balance multiple metrics — you rarely win on all metrics at once. For example, aggressively optimizing precision may increase latency or hide relevant results, ultimately leading to frustrated buyers.
  • Monitor real-time latency carefully. Break down latency by pipeline stages and track tail latency such as p95/p99 to quickly identify bottlenecks. Techniques such as caching may help meet strict latency SLA.
  • Track metrics systematically. Use MLflow to log and evaluate retrieval and engagement metrics across experiments. Native retrieval quality evaluation is coming soon to Databricks AI Search, making this even easier.

Search in Production at Scale — FOX Sports

FOX Sports built their AI-powered search bar on Databricks AI Search, handling thousands of QPS with a 2x improvement in query success rate. Launched for Super Bowl LIX, their architecture demonstrates several patterns covered in this blog:

  • Real-time ingestion. Spark Structured Streaming continuously ingests content into Delta Sync Indexes as it's published
  • Two-phase retrieval. Exact entity matching for players and teams, plus time-weighted semantic search for articles and videos, orchestrated by Databricks Model Serving
  • Production optimization. A caching layer and trending searches feature — driving over 25% of all search requests — handle high-traffic spikes during live events

From Search to Intelligent Applications

Product search doesn't exist in isolation — it's one layer in a broader application stack. Here's how the rest of the Databricks platform extends what you can build on top of AI Search.

Real-Time Applications with Lakebase

For customer-facing search applications — marketplaces, product catalogs, media platforms — the search index is only part of the story. Applications also need a transactional database for operational state: inventory levels, pricing, user sessions, personalization preferences. Lakebase provides this as a fully managed, PostgreSQL-compatible database natively integrated with the Databricks platform. Managed bidirectional sync with Delta Lake means ranking models train on the freshest operational data, and analytical insights flow back to the application layer — all governed by Unity Catalog.

Agent-Powered Search with Agent Bricks

Databricks automatically provides a managed MCP server for every AI Search index, unlocking multiple integration patterns:

  • Knowledge Assistant. A question-and-answer chatbot over your documents. Point it at a AI Search index and get production-ready document search with citations. Uses Instructed Retriever under the hood — 70% better accuracy than vanilla RAG and 30% better than agentic RAG.
  • Custom agents. Use the VectorSearchRetrieverTool or MCP with any framework (OpenAI Agents SDK, LangGraph, LlamaIndex). Full control over retrieval parameters, embeddings, and filters. Deploy as Databricks Apps with MLflow tracing.
  • Supervisor Agent. Orchestrate multiple subagents: a Knowledge Assistant for document Q&A, a Genie space for structured data queries, and UC Functions for custom business logic — all coordinated by a single supervisor.

Conclusion

Building a modern product search system requires more than a search index. It requires infrastructure designed to handle real-world scale, performance, and observability:

  • Low-latency execution. Query understanding, retrieval, filtering, and reranking must be completed within strict p95/p99 latency budgets.
  • Hybrid retrieval capability. Combine semantic similarity (embeddings) with structured filtering such as price, category, or availability.
  • Scalability under load. Sustain high QPS and concurrency during peak traffic without degrading performance.
  • Observability. Maintain clear visibility into latency breakdowns, ranking performance, and overall system health.
  • Agent-ready by default. Every AI Search index is an MCP tool, immediately usable by Knowledge Assistant, custom agents, and Supervisor Agents.
  • Full-stack operational support. Lakebase provides the transactional database for real-time application state, synced to Delta without ETL.

Ready to build? Follow the retrieval quality guide to benchmark and optimize your search pipeline, see how FOX Sports built AI-powered search at scale, and dive into the AI Search documentation to get started.