Lakehouse
PulseAugur coverage of Lakehouse — every cluster mentioning Lakehouse across labs, papers, and developer communities, ranked by signal.
4 day(s) with sentiment data
-
Databricks introduces native FILE type for multimodal data in lakehouse
Databricks has introduced a new beta feature called FILE type, designed to store unstructured data like documents, images, and videos as a native column within tables. This innovation aims to integrate multimodal data d…
-
Microsoft Fabric integrates Lakehouse and Warehouse options with OneLake
Microsoft Fabric offers both Lakehouse and Warehouse options within a unified platform, leveraging OneLake for data storage. The Lakehouse architecture, centered around Apache Spark and open Delta Parquet tables, is ide…
-
Databricks launches AI agent for SQL code migration
Databricks has launched an agentic code converter, powered by Genie Code, to simplify the migration of proprietary SQL dialects to open ANSI SQL. This beta feature supports conversions from T-SQL, Snowflake, Redshift, O…
-
Databricks MCP server grants AI agents direct lakehouse access
Databricks has released a new tool, the Databricks MCP server, which allows AI agents like Claude to directly access and interact with a user's data lakehouse. This integration enables conversational execution of notebo…
-
Hospitality Lakehouse Data Platform Built Using MLOps Principles
This article details the construction of a hospitality data platform utilizing a Lakehouse architecture. The platform is designed to manage various data streams including bookings, revenue, occupancy, payment reconcilia…
-
Databricks offers framework for ETL migration via SQL, SDP, or PySpark
Databricks has introduced a new framework to help organizations migrate their existing ETL (Extract, Transform, Load) pipelines. The framework outlines three primary migration paths: utilizing Databricks SQL for SQL-hea…
-
Databricks outlines best practices for modern data pipeline architecture and deployment
Databricks has published a comprehensive guide on data pipeline best practices, covering architecture, modern pipeline design, and deployment strategies. The guide emphasizes the importance of deliberate architectural c…
-
Databricks Lakebase adds Change Data Feed for direct operational data access
Databricks has introduced a new Change Data Feed (CDF) feature for its Lakebase product, now in public preview. This feature allows operational data to be directly accessed by various engines, models, and agents without…
-
AI-augmented lakehouse architecture proposed to improve data governance
A new paper proposes an AI-augmented hub-and-spoke model built on a lakehouse architecture to address the challenges of enterprise data platforms. This approach uses a central hub for AI-enabled governance, automating t…
-
Databricks guides analytics teams on choosing data warehouse tools for modern data needs
Databricks has published a guide to selecting data warehouse tools, emphasizing the importance of evaluating them across performance, scalability, integration, cost, and governance. The company advocates for the lakehou…
-
Databricks and Snapchat partner to send conversion data directly to advertisers
Databricks has integrated Snapchat's Conversions API into its Marketplace, allowing businesses to directly send first-party conversion data from their Lakehouse to Snapchat. This integration aims to improve advertising …
-
Databricks scales monitoring with Hydra; nOps rebuilds on Lakebase
Databricks has developed a new monitoring platform called Hydra, built on its Lakehouse architecture, to handle the massive scale of its operations, ingesting over 10 trillion samples daily and managing 5 billion active…