What is Synapse Data Engineering

Definition

Synapse Data Engineering in Microsoft Fabric provides the Spark-based environment for building large-scale data engineering workloads, using notebooks and lakehouses to ingest, transform, and prepare big data at scale within the unified Fabric platform.
« Back to Glossary Index
  • Runs scalable Spark data engineering on lakehouse data
  • Uses notebooks for flexible transformation and preparation
  • Integrates with OneLake so outputs are shared across Fabric
  • Handles big-data workloads within the unified platform

Real World Example

A data engineering team uses Synapse Data Engineering notebooks in Fabric to transform terabytes of raw lakehouse data into clean tables with Spark, feeding the results to the warehouse and Power BI.

FAQs

What does Synapse Data Engineering provide?

A Spark-based environment with notebooks and lakehouses for building scalable big-data engineering workloads in Fabric.

What language and tools does it use?

It uses Spark with notebooks supporting languages like PySpark and Spark SQL over lakehouse data.

How does it fit into Fabric?

It is the data engineering workload that prepares and transforms data in OneLake for other Fabric experiences.

Hello popup window