Microsoft Fabric Updates Blog

Announcing Delta Lake support in Real-Time Analytics KQL Database

As part of the One logical copy effort, we’re excited to announce that you can now enable availability of KQL Database in Delta Lake format.

Delta Lake  is the unified data lake table format chosen to achieve seamless data access across all compute engines in Microsoft Fabric.

The data streamed into KQL Database is stored in an optimized columnar storage format with full text indexing and supports complex analytical queries at low latency on structured, semi-structured, and free text data.

Enabling data availability of KQL Database in OneLake means that customers can enjoy the best of both worlds: they can query the data with high performance and low latency in their KQL database and query the same data in Delta Lake format via any other Fabric engines such as Power BI Direct Lake mode, Warehouse, Lakehouse, Notebooks, and more.

KQL Database offers a robust mechanism to batch the incoming streams of data into one or more Parquet files suitable for analysis. The Delta Lake representation is provided to keep the data open and reusable. This logical copy is managed once, is paid for once and users should consider it a single data set.

Users will only be charged once for the data storage after enabling the KQL Database availability in OneLake.

Enable OneLake availability

  1. To enable data availability in OneLake, browse to the details page of your KQL database or table.
  2. Next to OneLake folder in the Database details pane, select the Edit (pencil) icon.

Screenshot of the Database details pane in Real-Time Analytics showing an overview of the database with the edit OneLake folder option highlighted.

3. Enable the feature by toggling the button to Active, then select Done.

Screenshot of the OneLake folder details window in Real-Time Analytics in Microsoft Fabric. The option to expose data to OneLake is turned on.

You can enable data availability at a KQL database or table level.

Once you enable data availability, you can access all the new data added to your database at the given OneLake path in Delta parquet.

You can also choose to create a OneLake shortcut from Lakehouse, Data warehouse, or query the data directly via Power BI Direct Lake mode.

End-to-end streaming architecture in Fabric

Customers can now leverage data availability in OneLake to build more efficient and performant systems to handle high volume, and low latency streaming data in Microsoft Fabric.

  1. Eventstream can capture streaming data from multiple sources at scale.
  2. Eventstream allows pushing raw data into KQL Database seamlessly.
  3. KQL Database can be used to build a medallion architectural pattern with the help of in-built transformation functions such as update policies, and materialized views. The medallion structure is reflected as follows:–
    1. Bronze layer: raw data as received from Eventstream.
    2. Silver layer: deduplicated and enriched data.
    3. Gold layer: aggregated data suitable for reporting.
  4. Then you can either choose to make all three layers available in OneLake, or only make the aggregated data available for building your reports directly with Power BI.

This data design pattern allows you to scale efficiently for large incoming streams. KQL Database serves real-time analysis with low latency while making the data available in Delta Lake.

For more information on enabling data availability in OneLake, see One logical copy.

Related blog posts

Announcing Delta Lake support in Real-Time Analytics KQL Database

February 9, 2024 by Ruixin Xu

During January 2024, we announced the worldwide availability for public preview of Copilot in Microsoft Fabric. This preview includes Copilot for Power BI, Data Factory and Data Science & Data Engineering. With the Copilot in preview, Microsoft Fabric brings an improved way to transform, enrich and analyze data, and shortens the time to insights.  Today, we announce that Copilot … Continue reading “Announcing Fabric Copilot pricing “

February 1, 2024 by Kimberly Williams

We are excited to announce that Microsoft Fabric, our all-in-one analytics solution for enterprises, has achieved new certifications for HIPAA and ISO 27017, ISO 27018, ISO 27001, ISO 27701. These certifications demonstrate our commitment to providing the highest level of security and privacy for our customers’ data. What are these certifications and why do they matter? HIPAA (Health Insurance Portability … Continue reading “Microsoft Fabric is now HIPAA compliant”