NBCUniversal’s Seamless Migration: Unlocking Scalable Analytics with Databricks
Summary
NBCUniversal cut costs by 30% by moving from a slot-based reservation model to Databricks' dedicated job compute, letting parallel pipelines scale independently and consistently hit SLA targets. Partner EXL executed the migration in phases with custom accelerators for code conversion, data migration, and automated validation, giving NBCUniversal a unified platform for ML development, real-time analytics, and collaborative data engineering.
Summary generated by brickster.ai. For the full article, follow the source link above.
More from Databricks Blog
Managed Postgres: What Lakebase Actually Takes Off Your Plate
Lakebase runs PostgreSQL on serverless infrastructure with automatic scaling, scale-to-zero, point-in-time recovery, branching, pgvector, and PostGIS, handling most managed Postgres operations—patching, scaling, failover, backups—within a region. Cross-region disaster recovery is the one piece it doesn't fully own, still requiring customer-managed recovery procedures.
Announcing On-Demand State Repartitioning for Apache Spark™ Structured Streaming on Databricks
Databricks now supports on-demand state repartitioning for stateful Structured Streaming queries: set spark.sql.streaming.stateStore.partitions and restart on DBR 18+ with the RocksDB state store provider to redistribute state to a new partition count without rebuilding the checkpoint. This lets you resize long-running streams to match workload demands and monitor each resize through query progress metrics.
Health Plans: Your BI Tells You MLR Moved. Can Your AI Tell You Why?
Databricks and Abacus Insights combine to let payer finance teams use conversational AI to decompose MLR variances across claims, utilization, cost, and population risk in minutes instead of waiting on analyst reports. The approach embeds payer-specific business logic—covering IBNR, rebates, risk adjustment, and provider settlements—into a governed data and AI foundation so AI explains MLR shifts accurately rather than compounding confusion from inconsistent definitions across systems.
Unify your marketing data with Lakeflow Connect
Lakeflow Connect now offers native, fully managed connectors for marketing and ad platforms—including Salesforce, HubSpot, Google Ads, Meta Ads, TikTok Ads, LinkedIn Ads, Marketo, and more—landing governed data directly into Unity Catalog without any infrastructure to manage. Paired with the Ad-Genie solution accelerator, teams can turn raw ad data into a governed customer 360 and a natural-language Genie agent in just three steps.
Unifying governance across engines and catalogs in the Open Lakehouse
Apache Iceberg now has two new specs, read restrictions and catalog labels, that standardize policy enforcement across engines and catalogs by letting trusted engines handle access decisions directly and making governance metadata portable across federated catalogs. Together with centralized enforcement for untrusted engines, they give the Open Lakehouse a clear governance model covering every access pattern.
Improving Lakebase Postgres compute cache
Lakebase Postgres now runs an autoscaling cache that works alongside shared buffers to keep more data resident on compute instead of falling back to the storage layer. In production this delivered 2x throughput, fewer storage-layer reads, and lower latency compared to standard Postgres caching.
