惯性聚合 高效追踪和阅读你感兴趣的博客、新闻、科技资讯
阅读原文 在惯性聚合中打开

推荐订阅源

AWS News Blog
AWS News Blog
Engineering at Meta
Engineering at Meta
T
Tailwind CSS Blog
博客园 - 【当耐特】
F
Fortinet All Blogs
博客园 - 司徒正美
Stack Overflow Blog
Stack Overflow Blog
罗磊的独立博客
Google DeepMind News
Google DeepMind News
博客园_首页
A
About on SuperTechFans
博客园 - 聂微东
N
Netflix TechBlog - Medium
MongoDB | Blog
MongoDB | Blog
H
Help Net Security
GbyAI
GbyAI
O
OpenAI News
CTFtime.org: upcoming CTF events
CTFtime.org: upcoming CTF events
C
Check Point Blog
S
Schneier on Security
D
Darknet – Hacking Tools, Hacker News & Cyber Security
Apple Machine Learning Research
Apple Machine Learning Research
让小产品的独立变现更简单 - ezindie.com
让小产品的独立变现更简单 - ezindie.com
Martin Fowler
Martin Fowler
J
Java Code Geeks
T
Tor Project blog
T
Threatpost
G
GRAHAM CLULEY
P
Privacy & Cybersecurity Law Blog
Cyber Security Advisories - MS-ISAC
Cyber Security Advisories - MS-ISAC
Y
Y Combinator Blog
The Register - Security
The Register - Security
阮一峰的网络日志
阮一峰的网络日志
D
Docker
Vercel News
Vercel News
I
Intezer
Microsoft Security Blog
Microsoft Security Blog
爱范儿
爱范儿
cs.AI updates on arXiv.org
cs.AI updates on arXiv.org
H
Hacker News: Front Page
P
Privacy International News Feed
Cyberwarzone
Cyberwarzone
Spread Privacy
Spread Privacy
N
News and Events Feed by Topic
N
News | PayPal Newsroom
The GitHub Blog
The GitHub Blog
U
Unit 42
WordPress大学
WordPress大学
Threat Intelligence Blog | Flashpoint
Threat Intelligence Blog | Flashpoint
Simon Willison's Weblog
Simon Willison's Weblog

Databricks

How lakebase architecture delivers 5x faster Postgres writes Why Talent Transformation Is the Missing Focus of Enterprise AI Public Health Intelligence Shouldn't Require a Data Scientist Mean Time to Detect Is a Data Access Problem First-party audience data is the ad sales relationship now Rethinking Distributed Systems for Serverless Performance and Reliability The AI Scaling Gap Hiding in Digital Native Companies 10 trillion samples a day: Scaling beyond traditional monitoring infra at Databricks How nOps Rebuilt Their Cloud Optimization Platform on Databricks Lakebase, and Why Other ISVs Should Too Peril Predicts: Precision Payouts for a Volatile World The foundation of AI scalability: one team, one platform, one operating model The Federal Data Paradox: Rich in Data, Poor in Access Driving Budapest Forward: How BKK Uses Databricks to Transform City Mobility LLM Vs AI: A Practical Guide to Differences, Use Cases, and Tools Model Risk Governance Is Not the Same as Risk Intelligence Generative AI for Business: A Complete Strategy and Implementation Guide Data Science vs Data Engineering: Choosing Analysis or Infrastructure AI Applications: Tools, Use Cases, and Platforms MLOps vs DevOps: A Practical Guide for Data Scientists and IT Teams Top Data Warehouse Tools For Modern Data Analytics Unlocking SAP Business Context in Databricks with Semantic Metadata Delta Sharing The marketing activation gap has a fix: Databricks and Stitch partner to turn data infrastructure into marketing performance Alert Fatigue Is a Business Risk Backstage with Lakebase Shipping Faster isn’t Learning Faster Why Your OEE Dashboard Is Lying to You The Turbine That Tried to Tell You It Was Failing Predicting Readmissions Isn't Enough. Acting in Time Is. Clinical Trials Run Longer Than They Have To. That's a Patient Problem Network Quality Is a Revenue Problem, Not a Technical One Shelf Availability Starts with Better Demand Visibility When Predicting the Next Hit Requires More Than Intuition Approximate Answers, Exact Decisions: New Sketch Functions for Analytics Companies Winning with AI Built the Data Layer First Rethinking SQL ETL for modern data platforms Stripe data now available on Databricks via Databricks Marketplace Databricks and Stripe Projects: Infrastructure Built for Agents Agents are ready but your architecture probably isn't Interoperability Between Unity Catalog and Google BigQuery via Catalog Federation Built In, Not Bolted On: What AI-Native Actually Means in Cybersecurity Operationalizing AI for public sector fraud prevention From months to minutes: Building real-time clinical data pipelines with natural language Agentic Data Engineering with Genie Code and Lakeflow Securely send first-party conversion signals with Snapchat Conversions API on Databricks Marketplace How leading tech companies are killing the builder’s tax with Lakebase Inside one of the first production deployments of Lakebase: LangGuard's agentic workflow governance engine The next generation of Databricks Genie Model Risk Management in 2026: A Banker’s Guide to the Revised Interagency Guidance OpenAI GPT-5.5 now available on Databricks, fully-governed through Unity AI Gateway Operational databases: How they work and when to use them Databricks partners with OpenAI on GPT-5.5 Announcing the Public Preview of Lakeflow Designer Are LLM agents good at join order optimization? How conversational analytics removes the BI bottleneck How to transform document activation workflows with Genie and Agent Bricks Beyond the spreadsheet: how Databricks is delivering the modern CFO in Financial Services AI App Development: Guide To Building AI-Powered Apps IoT in Manufacturing: Strategy, Components, Use Cases, and Challenges Stop Hand-Coding Change Data Capture Pipelines Multimodal Data Integration: Production Architectures for Healthcare AI Personalization Strategies for Media Companies A Modern AI Risk Management Framework Introducing the Databricks Excel Add-in for Business Users Real-Time Decisioning for AI Agents: Why you Need a Customer Context Layer First A Practical Guide to LLM Fine Tuning AI Data Transformation Guide for Data Engineers and Data Scientists Concurrency Control in DBMS: How Locking, MVCC and Optimistic Strategies Keep Data Consistent Bridging data science and marketing: Databricks unveils Delta Sharing integration for Adobe Experience Platform and agentic marketing workflows Take Control: Customer-Managed Keys for Lakebase Postgres Get hands on with agents, vibe coding and more at Data+ AI Summit Mercedes-Benz Builds a Cross-Cloud Data Mesh with Delta Sharing and Intelligent Replication, Cutting Costs by 66% What Is a Transactional Database? Introducing Genie Agent Mode Governing coding agent sprawl with Unity AI Gateway Governing Coding Agent Sprawl with Unity AI Gateway What is pgvector? Banks Don’t Have an AI Problem – They Have a Data Platform Problem Open Platform, Unified Pipelines: Why dbt on Databricks is Accelerating Why Your Agents Can’t Read Enterprise Documents — and How to Fix It Building with Databricks Document Intelligence and Lakeflow Databricks on Google Cloud: Innovate Faster. Smarter. Together. Introducing the Databricks Connector for Google Sheets: Real-Time, Governed Lakehouse Data in the Sheets Users Love Unity AI Gateway: How to connect agents to external MCPs securely Expanding agent governance with Unity AI Gateway Agentic reasoning in practice: Making sense of structured and unstructured data Agent Bricks: The Governed Enterprise Agent Platform 8 AI and data trends shaping financial services in 2026 Building real-time product search on Databricks Lovable + Databricks: Build Data-Driven Apps at the Speed of Thought Memory scaling for AI agents Powering clinical research innovation: How TriNetX uses Databricks to accelerate drug development Database Branching in Postgres: Git-Style Workflows with Databricks Lakebase How Zalando built a unified data foundation for AI and analytics on Databricks The next era of the open lakehouse: Apache Iceberg™ v3 in Public Preview on Databricks How FSIs eliminate silos between clients, operations, and finance How MakeMyTrip achieved millisecond personalization at scale with Databricks A multi-agent approach to audience intelligence AiChemy: Next-generation agent with MCP, skills and custom data for drug discovery Accelerate business insights with Lakeflow Connect, now with a Free Tier Unlocking Next-Gen Customer Experiences with Data Intelligence for Marketing
AI success starts with clean data, not just better models
2026-05-06 · via Databricks

Kraken, the AI-powered operating system behind some of the world's largest utilities, manages over 90 million customer accounts across 27 countries for clients including EDF, E.ON, National Grid, and Tokyo Gas. Kraken uses Databricks as its internal data platform and partners with Databricks to help clients maximize the value of the data they receive via secure, scalable data distribution.

Kristy Mayer-Mejia is the Global Head of Data Transformation at Kraken, where her team helps utility clients understand, adopt, and extract value from the data Kraken provides. Her mandate is twofold: speed up the time it takes clients to use the data, and increase the value they get from it.

I sat down with Kristy to understand how data functions as a business asset and is the foundation of a successful AI strategy. A key point from our conversation is that becoming data-driven is as much about clean, unified data as it is about deep business context and ownership. Platforms like Kraken and Databricks solve what Kristy calls the foundational unification problem, the prerequisite that makes everything else viable. But once that foundation is in place, the part most leaders underestimate is the business context that makes unified data usable.

Why data unification is table stakes

Aly McGue: In your experience, why do siloed data and fragmented systems remain a big hurdle for organizations trying to extract value from their investments?

Kristy Mayer-Mejia: What we see repeatedly with our clients is that low-quality, siloed data is the single biggest blocker to getting value from any other investment. Until the data is in one place, nothing else works at scale, and solving that is exactly the problem Kraken’s platform is designed to address. And I've lived this as a data leader in all my prior roles, too. Your team spends 80% of their time cleaning data, and that's just not valuable work. It's not necessary.

The real unlock is self-service, but it’s only possible once the underlying data is clean, unified, and accessible. Especially in the age of AI, self-service is possible at scale. You're never going to move quickly as a business, innovate or make day-to-day data-driven decisions if every question has to be answered by the data team. But when the data is scattered across systems with no documentation and no clear way to join it, self-service is impossible. Unification is this foundational unlock that makes everything else viable: the analytics, the AI, the speed of decision-making. It's table stakes.

The number no one trusts

Aly: We’ve all been in meetings where leadership spends more time debating 'which number is right’ than actually making a decision. What is the hidden cost of that lack of trust in the data?

Kristy: I give this example all the time, and it's been true at every company I've ever worked at. Before you have unified data, the classic question is: how many customers do we have? And no one totally knows. You know the rough magnitude. But when I give that example, every time people laugh because they know it's true.

What it leads to is a lack of trust in the data. And one of the primary early values that unified data provides is the speed of decision-making, the ability to embed data-driven thinking in the company's DNA. You can't move quickly if every time you pull a number, you're thinking, ’ Am I sure this is right? ’ Let me check five other places. Let me ask someone. And then it's different. And then you have to run down why it's different. Suddenly, it's two weeks later or a month later, and you might as well have just picked a random direction and kept moving.

AI is the forcing function enterprise analytics needed

Aly: We often talk about data fueling AI, but you’ve suggested that AI might actually be a 'forcing function’ for better data. How is the push for AI changing the way organizations approach documentation and context?

Kristy: AI has actually been a forcing function. The inputs AI needs are the same inputs humans need: clear data, documentation, context on what columns mean, and how things join together. When data is hard to use, self-service analytics feels like a nice-to-have because the value is hard to pin down. It's a few hours saved here and there on individual decisions, which doesn't feel compelling in isolation. But accumulated across the organization, it's huge. It's just hard to see.

AI has made that value visible and has made clean data & documentation table stakes. It takes what everyone always knew was needed and makes it non-negotiable. And then on the other side, AI itself provides the tools to unlock analytics. Things like conversational interfaces that let people query data without writing SQL. So it's both the forcing function that drives unification and the payoff that comes out of it.

Metadata as the missing ingredient

Aly: You've talked about the need to unify and document data. But when it comes to AI specifically, is documentation in a knowledge base or a PDF enough?

Kristy: It used to be. We shared our data documentation the way most companies do: a PDF, or a page on a website that a data analyst could reference when they needed context. That works well enough for humans. It does not work for AI.

Every client I talk to now is asking the same question: can you share the metadata in context, alongside the data itself, so we can actually feed it into models and have them understand what they're working with? That shift, from documentation as a reference artifact to documentation as a live input, is one of the more underappreciated changes AI is forcing. With Unity Catalog and Delta Sharing, we can share that context with the data rather than separately from it. For our clients, that is often the difference between AI that can reason about the data and AI that cannot.

From monthly reports to hourly decisions

Aly: What does 'data unification’ look like in practice? How does near-real-time visibility change day-to-day operations?

Kristy: A few examples from our clients stand out. One is call center operations, which is a massive function for utilities. We had a client go from monthly reporting on call volume, which was so painful to put together, to dashboards that update every couple of hours, with a predictive model layered on top of what calls they're likely to see going forward. That ability to fine-tune operations in near real time, rather than looking backward once a month, is a completely different way of running the business.

Another area is product innovation. In the utility space, clients are determining which products and tariffs to offer to attract and retain customers. That's a decision that can be deeply optimized with data. Clean, clear data give clients easy insight, and rapid test-and-learn cycles to optimize their product offers – and then Kraken's platform lets them quickly launch those new tariffs.

Getting people into the data

Aly: The 'analyst bottleneck’ is a classic pain point for leadership. How do natural language interfaces, like Databricks Genie, shift the culture from waiting weeks for a report to getting answers in minutes?

Kristy: Most of our Genie clients are still in the early stages. But what we're seeing is that it's accelerating their time to get started by weeks or more. They don't need to deeply model the data the way you would to feed it into a traditional BI tool. They need clear documentation, they need the context, they need the data in one place, but they don't have to structure it so precisely that a user can explore it through a rigid interface.

But beyond the speed, there's a really clear cultural knock-on effect. One of the bigger barriers to data value is the cultural shift of making data part of your DNA. And I firmly believe one of the keys to that is making it incredibly easy and intuitive. When the barrier is low, and people can get in quickly, the culture, and the compounding value, follows.

The advice most C-suite get wrong

Aly: What is the biggest misconception C-level leaders have when they task their IT departments with 'getting the data ready’ for AI?

Kristy: Data is a business asset. And the biggest mistake I see leaders make is treating it like an IT platform. They disconnect it from the business and say, "Okay, IT, go prepare our data.” But the key to building a solid data foundation is the deep business context. How is the data generated? How is it used? How do people interpret it? What does this field actually mean? Once the technical foundation is in place, the hardest part becomes that deep business context. And the vast majority of that work sits with the business, not the data team.

So my advice is to embed data within the business. The roadmap to getting your data ready for AI is a shared roadmap. It's a business roadmap as much as it is a technical one.

What good looks like from here

Aly: Kraken sits across a large share of the utility industry's data. Where do you see AI and data taking your clients over the next three to five years?

Kristy: What I find most interesting is how quickly AI is raising the ceiling on what clients can do once they have a solid data foundation. For a long time, the question was: how do we get our data into a usable state? That work is still real, and it still takes time. But the question is shifting toward: now that the foundation is there, what becomes possible? And the answer to that keeps expanding. AI is changing where clients start from and what good looks like. Clients who would have considered a monthly report a success two years ago are now running hourly dashboards with predictive models layered on top and looking quickly toward broad use of agentic AI.

The ones who invested early in their data capabilities – and not just their tech but their skills and culture – are the ones moving fastest now, and the gap between them and everyone else is only going to widen.

Closing Thoughts

Kristy's perspective adds an often-missing layer to the data infrastructure conversation. The platform and the unification it enables are the foundational unlock. But where she sees most organizations stall is in the work that comes after: the business knowledge that makes data usable, the documentation that makes AI possible, and the cultural shift that makes self-service real.

As you develop your roadmap to embed AI across your organization and products, download the Databricks State of AI Agents to help you benchmark your investments.