惯性聚合 高效追踪和阅读你感兴趣的博客、新闻、科技资讯
阅读原文 在惯性聚合中打开

推荐订阅源

WordPress大学
WordPress大学
小众软件
小众软件
MongoDB | Blog
MongoDB | Blog
Hugging Face - Blog
Hugging Face - Blog
Jina AI
Jina AI
Cyber Security Advisories - MS-ISAC
Cyber Security Advisories - MS-ISAC
Stack Overflow Blog
Stack Overflow Blog
L
LangChain Blog
大猫的无限游戏
大猫的无限游戏
量子位
A
About on SuperTechFans
G
Google Developers Blog
雷峰网
雷峰网
奇客Solidot–传递最新科技情报
奇客Solidot–传递最新科技情报
IT之家
IT之家
钛媒体:引领未来商业与生活新知
钛媒体:引领未来商业与生活新知
博客园_首页
H
Hackread – Cybersecurity News, Data Breaches, AI and More
Vercel News
Vercel News
V
Visual Studio Blog
OSCHINA 社区最新新闻
OSCHINA 社区最新新闻
博客园 - 聂微东
U
Unit 42
Apple Machine Learning Research
Apple Machine Learning Research

Databricks

Using AI_Functions in Your Data Warehouse: Top Use Cases How Scottish Water Made Its Capital Investment Data Conversational With Databricks Genie What are AI Hallucinations? Smart Routing in Unity AI Gateway: Match frontier quality with 30%+ lower cost per task Databricks Network Configuration delivery to Tens of Millions of Serverless VMs How Amtrak is building the data backbone for its largest transformation in over 50 years How a major freight railroad scaled pipeline creation with Genie Code The Future of Data Analytics: Why AI is rewriting the Analyst’s Job Description Taking AUTO CDC to the next level: Solving the hardest real-world use cases Open-sourcing Metals v2: Databricks’ Java and Scala language server for multi‑million line codebases Modern Risk Demands a Real-Time Foundation: The CRO’s Mandate Electric joins Databricks to bring WASM Postgres to AI agent sandboxes How to ground Genie Agents in both structured data and documents without losing governance Innocent until combined: Blocking the lethal trifecta with Omnigent Contextual Policies Introducing FILE type: a native column type for multimodal data Managing AI Coding Costs at Scale What is an AI Assistant? What are Agentic Workflows? What is Tool Calling? Kimi K3 from Moonshot AI is now available on Databricks through Unity AI Gateway Introducing OfficeQA Pro V2: A New Benchmark for Enterprise Grounded-Reasoning BigQuery to Databricks: A Strategic Framework for Modern Migration Unity AI Gateway is Generally Available Granular Usage Attribution for dbt Pipelines with Query Tags - Cloned Databricks joins the Open Secure AI Alliance to advance AI safety and security The New Monday Morning Report: How Generative AI can deliver the insights your executives need. Ingest semi-structured data faster and more efficiently with Variant - Now Generally Available Databricks Completes Acquisition of Panther: Accelerating the Security Lakehouse Era Backstage with Lakebase, part 3 Foundations for an AI-forward healthcare organization
Introducing AI Spend Controls with Unity AI Gateway
2026-05-19 · via Databricks

Today, we're announcing AI Spend Controls in Unity AI Gateway. This release extends Unity AI Gateway's existing cost visibility with proactive budget alerts to give you full control over your organization's AI spend - from the coding agents your developers use every day, to the production agents serving your customers, to the batch jobs running overnight:

AI workloads deliver disproportionate value - but their cost profile is fundamentally more challenging to manage than your traditional cloud spend:

  • Your nightly batch job translating call transcripts may run perfectly for a month, then start failing halfway through and trigger retry logic that multiplies its cost 10x overnight.
  • Your engineering org's coding agents save thousands of developer hours a week - but the same agents make it easy for one engineer to kick off an accidental multi-agent experiment Friday night that burns through the team's monthly budget by Sunday.

Employees across engineering, support, sales, and ops are onboarding to AI faster than any technology in the last decade, unlocking net-new use cases week over week. But that adoption brings a management challenge: foundation model usage now spans dozens of teams, hundreds of users, and thousands of agents with a shifting mix of providers and model tiers. Spend controls need to apply uniformly across all AI workloads, so your organization can confidently lean into AI without worrying about surprises on the bill.

Configure Budget Alerts at Every Granularity 

While spend controls need to apply uniformly, different parts of your organization need different cost controls. A platform team cares about workspace-wide totals. A FinOps lead cares about the org-level monthly burn. An engineering manager cares about per-developer experimentation budgets. AI Spend Controls let you set them all from one place and is deeply integrated with Databricks’ existing budgets:

  • Per user: Set budgets for individual experimentation — for example, $2000 per user per month for the engineering org. Catch the developer whose agent is stuck in a loop before it shows up on the P&L.
  • Per use case: Get alerted if your organizations’ spend on coding agents like codex or claude code exceeds $1000 per user per month
  • Per workspace: Hold each unit to its own budget. Production gets $50,000/month; sandbox gets $5,000.
  • Per account: Set a top-line ceiling — say, $200,000/month across every model, every provider, every workspace — and get alerted long before you approach it.

Get Started with Unity AI Gateway Budgets Today

To track your organization’s AI spend, follow these steps:

Create your Unity AI Gateway Budget

  • Open your account settings, navigate to Usage in the sidebar and open the Budgets tab
  • Create a Budget and select “Unity AI Gateway” as the Resource type
  • Optionally apply the budget only to a subset of workspaces 
  • Optionally apply “Resource tags” to configure budgets for a subset of your AI Gateway LLMs. Only AI Gateway LLMs whose tags match your budget tags will count towards the budget. This is useful to configure use-case specific budgets.
  • Configure a “Shared threshold” that sets the monthly spend limit globally across all resources in your selected workspace(s) that match the resource tags
  • Configure a “Per-user threshold” that sets a monthly spend limit per user in your account 
  • Configure email addresses that receive alerts when the thresholds are exceeded

Once created, look out for budget alerts

When one of your budgets is exceeded you will receive a notification email:

Analyze your active budgets

The Cost section of your account console lets you respond to budget alert emails or proactively monitor the status of your live budgets. On the Budgets page, you see at a glance how your budgets are trending:

Open up any budget to see how your AI spend is trending:

If you configured per-user level budget thresholds, the Budget detail page will show you how your organization’s users' individual AI spend is trending. When users exceed their individual threshold, their status and spend are clearly surfaced so you can act quickly:

To increase a budget’s threshold, you can simply edit the Budget and modify its spend limits.

Analyze your organization’s AI Spend in detail

Unity AI Gateway Budgets give you a high-level overview of per-user and per-budget spend. To further analyze which users, models or use cases are driving your spend, you can use Unity AI Gateway’s existing cost tracking capabilities. Every request gets logged to Unity Catalog system tables with DBU costs and not just token counts. Provisioned throughput, uptime, pay-per-token usage, and even the token costs of external model providers are all automatically calculated. You can slice the data however your organization tracks spend:

  • Identity: Aggregate by user or service principal — map spend to the people and systems driving it.
  • Workspace, endpoint and tags: Group by team, environment, or cost center.
  • Model and provider: See which models (Opus vs. Sonnet) and providers (Anthropic vs. OpenAI vs. open source) are driving costs.
  • Request tags: Dynamic attribution for SaaS platforms proxying to end customers.

Access the Cost Analytics dashboard by navigating to the Unity AI Gateway page in your Databricks workspace and click on “View Dashboard”:

This opens up a usage & cost analytics dashboard that you can fully customize:

One platform to govern data and AI

AI Spend Controls are a natural extension of the governance capabilities you already use in Databricks:

  • Unity AI Gateway is your organization’s central AI Gateway to manage and access LLMs and MCPs.
  • Unity Catalog is your central catalog to register and discover your organization’s data and AI assets. Access permissions, audit logs and usage data all live in Unity Catalog. 
  • Databricks budgets provide the foundation for cost monitoring and alerting. With this release, Databricks budgets now allow you to configure AI-tailored budgets for your organization’s AI workloads.

Databricks provides you with a single, consistent system for governing what your agents can do, who they can do it for, and how much they can spend doing it. Get started today!