LSEG unifies 30 data systems, reducing product development time with Fabric
LSEG, a trusted data partner to 44,000 customers in more than 170 countries, partnered with Microsoft to build a unified data platform on Microsoft Fabric, consolidating 30 systems, 1,200 datasets and 33 petabytes of data. Using Apache Spark on Fabric, Apache Airflow in Fabric, Microsoft Purview and OneLake, LSEG reduced product development timelines for new data products from years to months; onboarding data and creating a new product now happens in months, accelerating launches such as its ESG and Fundamentals data products. LSEG processes around 80,000 files daily on Spark on Fabric, consuming approximately 280,000 capacity units per day, with month-on-month consumption growing more than 50%. The company is also building AI-readiness by hosting Model Context Protocol (MCP)-powered data in OneLake and enabling financial professionals to build custom AI agents via Copilot Studio.
Overview
LSEG, a trusted data partner to 44,000 customers in more than 170 countries, partnered with Microsoft to build a unified data platform on Microsoft Fabric, consolidating 30 systems, 1,200 datasets and 33 petabytes of data. Using Apache Spark on Fabric, Apache Airflow in Fabric, Microsoft Purview and OneLake, LSEG reduced product development timelines for new data products from years to months; onboarding data and creating a new product now happens in months, accelerating launches such as its ESG and Fundamentals data products. LSEG processes around 80,000 files daily on Spark on Fabric, consuming approximately 280,000 capacity units per day, with month-on-month consumption growing more than 50%. The company is also building AI-readiness by hosting Model Context Protocol (MCP)-powered data in OneLake and enabling financial professionals to build custom AI agents via Copilot Studio.
This entry has 11 published fields tied to exact passages in an immutable source capture.
Inspect the highlighted sourceThe challenge
LSEG was dealing with petabytes of data across 30 systems and 1,200 datasets, spanning a variety of longstanding systems each with its own tech stack, distribution method, and data model, and wanted a more unified customer experience across its product offerings, since product instability would be an economy-wide issue given global reliance on LSEG's financial data.
The solution
LSEG partnered with Microsoft to co-engineer a unified data platform built on Microsoft Fabric, using Apache Spark on Fabric, Apache Airflow in Fabric, and Microsoft Purview to build scalable pipelines with embedded data quality checks, and OneLake for interoperability. The company is also building AI-readiness by hosting Model Context Protocol (MCP)-powered data in OneLake, enabling financial professionals to build custom AI agents via Copilot Studio.
Reported business value
LSEG consolidated 30 systems, 1,200 datasets and 33 petabytes of data, reducing product development timelines for new data products from years to months, and accelerating launches such as its ESG and Fundamentals data products. LSEG processes around 80,000 files daily on Spark on Fabric, consuming approximately 280,000 capacity units per day, with month-on-month consumption growing more than 50%.
Sources
Open any source and check the claim yourself — that is the point of the register.
This record was researched and written with AI assistance, and its claims were checked against the sources above. (EU AI Act art. 50 transparency notice.)
Other financial services entries in the register.
Navy Federal Transforms Service With AI
Navy Federal Credit Union is reshaping banking for military members by unifying data and leveraging generative and agentic AI on the Databricks Data + AI Platform. By embracing AI-augmented workflows and upskilling teams, Navy Federal delivers customized services while streamlining productivity through responsible change management and data readiness.
Banking Innovator bunq Supports Growth, Strengthens Security Using AWS
bunq, a Dutch neobank with over 11 million users across Europe, uses Amazon Bedrock for several generative AI use cases including summarizing new user data with large language models, removing the need for agents to process onboarding documents manually. Using Amazon Bedrock, bunq tripled user support process efficiency while maintaining over 90 percent accuracy. Sensitive data stays within bunq's AWS virtual private cloud, supporting GDPR and PCI DSS compliance alongside tools such as AWS CloudHSM, AWS Security Hub and AWS KMS.
TBC Bank Operationalizes Trusted Data with Lakebase
TBC Bank, the largest banking group in the Caucasus region, built a Lakehouse on Databricks and adopted Lakebase and Databricks Apps to move from on-premises SQL Server instances and month-long reporting cycles to self-service analytics and AI-driven applications, including a web-based AI chatbot and AutoML-based credit risk scoring. Credit risk model deployment fell from 14 weeks to two days, and more than 600 users regularly query governed data through Genie.
Worldline enables real-time insights for smarter merchant decisions with Databricks
European payment processor Worldline consolidated data from multiple acquisitions onto a Databricks medallion architecture with Delta Lake and Unity Catalog to unify over 50 billion annual transactions, reducing infrastructure costs by €200,000 per month, lifting team productivity 40%, and increasing scheme reporting speed 93%.
Was this helpful?
Your feedback helps us improve our use case database

