Senior Data Engineer
Senior
Adls Gen2
Apache Airflow
Azure
Azure Arm Templates
Azure Data Factory
Azure Data Lake Storage
Azure Databricks
Azure Event Hubs
Azure Monitor
Big Data
Bigdata
Cloud
Cloud Platforms
Data
Data Architecture
Data Engineer
Data Governance
Data Integration
Data Lake
Data Pipeline
Data Platform
Data Processing
Data Warehouse
Database
Databricks
DevOps
ETL
Infrastructure As Code
Key Vault
Microsoft Azure
Serverless Compute
Spark
SQL
Job Description
Deloitte is seeking a Senior Data Engineer to shape and optimize comprehensive data platforms, building scalable ETL and ELT pipelines and partnering with engagement teams and client architects to deliver robust solutions. The role is anchored on hands-on development with Azure and Databricks, conducted onsite in Morristown, New Jersey, as part of a client-facing data engineering practice.
Responsibilities
- Communicate regularly with Engagement Managers (Directors), project team members, and representatives from multiple functional and technical teams, escalating issues that require leadership attention.
- Design, develop, and optimize ETL/ELT pipelines using Azure Data Factory and Databricks.
- Write and tune PySpark and Spark SQL notebooks for large-scale data transformations.
- Architect end-to-end data solutions across dev, UAT, and prod environments using Unity Catalog.
- Lead design discussions with client architects and counterparts to shape data solutions.
- Collaborate with teams on data contracts and schema agreements.
- Lead the design and optimization of high-volume data pipelines.
- Define and enforce data engineering standards including naming conventions, partitioning strategies, cluster configurations, and Spark tuning.
- Drive performance optimization through AQE tuning, liquid clustering, broadcast joins, and shuffle partition management.
- Design Databricks cluster policies, autoscaling configurations, and cost optimization strategies.
- Conduct root cause analysis on production incidents and implement permanent fixes.
- Mentor junior and mid-level engineers through code reviews and pair programming.
- Evaluate new technologies and recommend adoption (examples include DABs, DLT, Auto Loader, Serverless Compute, and Azure Event Hubs).
Requirements
- Python, PySpark, Spark SQL, SQL Server
- Azure (ADF, ADLS Gen2, Key Vault, Azure Monitor)
- Databricks (Delta Lake, Unity Catalog, Workflows)
- Apache Airflow
- Git / Azure DevOps
- Deep Spark internals (DAG optimization, spill analysis, skew handling)
- Delta Lake advanced features (time travel, deletion vectors, predictive I/O)
- Unity Catalog governance (row/column security, external locations, system tables)
- IaC - Terraform, Azure ARM templates
- Bachelor's degree, preferably in Computer Science, Information Technology, Computer Engineering, or related IT discipline; or equivalent experience
- Limited immigration sponsorship may be available
- Ability to travel 10% on average, depending on client needs
Technologies
- Python
- PySpark
- Spark SQL
- SQL Server
- Azure Data Factory (ADF)
- ADLS Gen2
- Key Vault
- Azure Monitor
- Databricks
- Delta Lake
- Unity Catalog
- Workflows
- Apache Airflow
- Git
- Azure DevOps
- DABs
- DLT
- Auto Loader
- Serverless Compute
- Azure Event Hubs
- Terraform
- Azure ARM templates
Location
Morristown, NJ (onsite)
Compensation
USD 95,000 - 150,000 per year
Education
Bachelor's degree, preferably in Computer Science, Information Technology, Computer Engineering, or related IT discipline; or equivalent experience
Accommodations
- Information for applicants with a need for accommodation: accommodation information page