惯性聚合 高效追踪和阅读你感兴趣的博客、新闻、科技资讯
阅读原文 在惯性聚合中打开

推荐订阅源

M
MIT News - Artificial intelligence
罗磊的独立博客
Hugging Face - Blog
Hugging Face - Blog
J
Java Code Geeks
G
Google Developers Blog
美团技术团队
让小产品的独立变现更简单 - ezindie.com
让小产品的独立变现更简单 - ezindie.com
腾讯CDC
奇客Solidot–传递最新科技情报
奇客Solidot–传递最新科技情报
T
The Blog of Author Tim Ferriss
月光博客
月光博客
B
Blog
WordPress大学
WordPress大学
云风的 BLOG
云风的 BLOG
博客园_首页
人人都是产品经理
人人都是产品经理
aimingoo的专栏
aimingoo的专栏
Y
Y Combinator Blog
Jina AI
Jina AI
S
SegmentFault 最新的问题
H
Help Net Security
博客园 - 聂微东
Microsoft Azure Blog
Microsoft Azure Blog
Google DeepMind News
Google DeepMind News

Databricks

Why Talent Transformation Is the Missing Focus of Enterprise AI Public Health Intelligence Shouldn't Require a Data Scientist Mean Time to Detect Is a Data Access Problem First-party audience data is the ad sales relationship now Rethinking Distributed Systems for Serverless Performance and Reliability The AI Scaling Gap Hiding in Digital Native Companies 10 trillion samples a day: Scaling beyond traditional monitoring infra at Databricks AI success starts with clean data, not just better models How nOps Rebuilt Their Cloud Optimization Platform on Databricks Lakebase, and Why Other ISVs Should Too Peril Predicts: Precision Payouts for a Volatile World The foundation of AI scalability: one team, one platform, one operating model The Federal Data Paradox: Rich in Data, Poor in Access Driving Budapest Forward: How BKK Uses Databricks to Transform City Mobility LLM Vs AI: A Practical Guide to Differences, Use Cases, and Tools Model Risk Governance Is Not the Same as Risk Intelligence Generative AI for Business: A Complete Strategy and Implementation Guide Data Science vs Data Engineering: Choosing Analysis or Infrastructure AI Applications: Tools, Use Cases, and Platforms MLOps vs DevOps: A Practical Guide for Data Scientists and IT Teams Top Data Warehouse Tools For Modern Data Analytics Unlocking SAP Business Context in Databricks with Semantic Metadata Delta Sharing The marketing activation gap has a fix: Databricks and Stitch partner to turn data infrastructure into marketing performance Alert Fatigue Is a Business Risk Backstage with Lakebase Shipping Faster isn’t Learning Faster Why Your OEE Dashboard Is Lying to You The Turbine That Tried to Tell You It Was Failing Predicting Readmissions Isn't Enough. Acting in Time Is. Clinical Trials Run Longer Than They Have To. That's a Patient Problem Network Quality Is a Revenue Problem, Not a Technical One
Introducing AI Spend Controls with Unity AI Gateway
2026-05-19 · via Databricks

Today, we're announcing AI Spend Controls in Unity AI Gateway. This release extends Unity AI Gateway's existing cost visibility with proactive budget alerts to give you full control over your organization's AI spend - from the coding agents your developers use every day, to the production agents serving your customers, to the batch jobs running overnight:

AI workloads deliver disproportionate value - but their cost profile is fundamentally more challenging to manage than your traditional cloud spend:

  • Your nightly batch job translating call transcripts may run perfectly for a month, then start failing halfway through and trigger retry logic that multiplies its cost 10x overnight.
  • Your engineering org's coding agents save thousands of developer hours a week - but the same agents make it easy for one engineer to kick off an accidental multi-agent experiment Friday night that burns through the team's monthly budget by Sunday.

Employees across engineering, support, sales, and ops are onboarding to AI faster than any technology in the last decade, unlocking net-new use cases week over week. But that adoption brings a management challenge: foundation model usage now spans dozens of teams, hundreds of users, and thousands of agents with a shifting mix of providers and model tiers. Spend controls need to apply uniformly across all AI workloads, so your organization can confidently lean into AI without worrying about surprises on the bill.

Configure Budget Alerts at Every Granularity 

While spend controls need to apply uniformly, different parts of your organization need different cost controls. A platform team cares about workspace-wide totals. A FinOps lead cares about the org-level monthly burn. An engineering manager cares about per-developer experimentation budgets. AI Spend Controls let you set them all from one place and is deeply integrated with Databricks’ existing budgets:

  • Per user: Set budgets for individual experimentation — for example, $2000 per user per month for the engineering org. Catch the developer whose agent is stuck in a loop before it shows up on the P&L.
  • Per use case: Get alerted if your organizations’ spend on coding agents like codex or claude code exceeds $1000 per user per month
  • Per workspace: Hold each unit to its own budget. Production gets $50,000/month; sandbox gets $5,000.
  • Per account: Set a top-line ceiling — say, $200,000/month across every model, every provider, every workspace — and get alerted long before you approach it.

Get Started with Unity AI Gateway Budgets Today

To track your organization’s AI spend, follow these steps:

Create your Unity AI Gateway Budget

  • Open your account settings, navigate to Usage in the sidebar and open the Budgets tab
  • Create a Budget and select “Unity AI Gateway” as the Resource type
  • Optionally apply the budget only to a subset of workspaces 
  • Optionally apply “Resource tags” to configure budgets for a subset of your AI Gateway LLMs. Only AI Gateway LLMs whose tags match your budget tags will count towards the budget. This is useful to configure use-case specific budgets.
  • Configure a “Shared threshold” that sets the monthly spend limit globally across all resources in your selected workspace(s) that match the resource tags
  • Configure a “Per-user threshold” that sets a monthly spend limit per user in your account 
  • Configure email addresses that receive alerts when the thresholds are exceeded

Once created, look out for budget alerts

When one of your budgets is exceeded you will receive a notification email:

Analyze your active budgets

The Cost section of your account console lets you respond to budget alert emails or proactively monitor the status of your live budgets. On the Budgets page, you see at a glance how your budgets are trending:

Open up any budget to see how your AI spend is trending:

If you configured per-user level budget thresholds, the Budget detail page will show you how your organization’s users' individual AI spend is trending. When users exceed their individual threshold, their status and spend are clearly surfaced so you can act quickly:

To increase a budget’s threshold, you can simply edit the Budget and modify its spend limits.

Analyze your organization’s AI Spend in detail

Unity AI Gateway Budgets give you a high-level overview of per-user and per-budget spend. To further analyze which users, models or use cases are driving your spend, you can use Unity AI Gateway’s existing cost tracking capabilities. Every request gets logged to Unity Catalog system tables with DBU costs and not just token counts. Provisioned throughput, uptime, pay-per-token usage, and even the token costs of external model providers are all automatically calculated. You can slice the data however your organization tracks spend:

  • Identity: Aggregate by user or service principal — map spend to the people and systems driving it.
  • Workspace, endpoint and tags: Group by team, environment, or cost center.
  • Model and provider: See which models (Opus vs. Sonnet) and providers (Anthropic vs. OpenAI vs. open source) are driving costs.
  • Request tags: Dynamic attribution for SaaS platforms proxying to end customers.

Access the Cost Analytics dashboard by navigating to the Unity AI Gateway page in your Databricks workspace and click on “View Dashboard”:

This opens up a usage & cost analytics dashboard that you can fully customize:

One platform to govern data and AI

AI Spend Controls are a natural extension of the governance capabilities you already use in Databricks:

  • Unity AI Gateway is your organization’s central AI Gateway to manage and access LLMs and MCPs.
  • Unity Catalog is your central catalog to register and discover your organization’s data and AI assets. Access permissions, audit logs and usage data all live in Unity Catalog. 
  • Databricks budgets provide the foundation for cost monitoring and alerting. With this release, Databricks budgets now allow you to configure AI-tailored budgets for your organization’s AI workloads.

Databricks provides you with a single, consistent system for governing what your agents can do, who they can do it for, and how much they can spend doing it. Get started today!