Senior Data Engineer (Databricks & Cloud Analytics)
Senior
Azure Data Factory
Azure Data Lake
Big Data
Data Analytics
Data Architecture
Data Engineer
Data Factory Azure
Data Governance
Data Integration
Data Lake
Data Lakehouse
Data Pipeline
Data Platform
Data Processing
Data Security
Data Warehouse
Database
Databricks
Databricks Mlflow
Delta Lake
Delta Live Tables
ETL
Microsoft Azure
Spark
SQL
Job Description
CGI Group, Inc. is seeking a Senior Data Engineer focused on Databricks and cloud analytics to help design and deliver secure, scalable, high-performing enterprise data platforms. This onsite role is based in the Salt Lake City, UT area and centers on building modern cloud data pipelines and analytics capabilities.
In this position, you will apply strong engineering experience across Databricks, Spark, and Azure to develop end-to-end data solutions, including ingestion, transformation, governance, and production support.
Key Responsibilities
- Design, develop, test, deploy, and maintain enterprise data engineering solutions using modern cloud and big data technologies.
- Design and implement scalable ETL/ELT pipelines with Databricks, Apache Spark (PySpark), Delta Lake, and Azure Data Factory.
- Build high-performance ingestion, transformation, and integration frameworks that support enterprise analytics and reporting.
- Develop, optimize, and maintain complex SQL queries, stored procedures, and data transformation processes.
- Design and implement scalable data models for business intelligence, analytics, and AI initiatives.
- Build and integrate RESTful APIs, event-driven architectures, and legacy SOAP services for enterprise data exchange.
- Develop and maintain streaming and messaging solutions using Apache Kafka.
- Monitor, troubleshoot, and optimize production data pipelines, performing root cause analysis and implementing long-term fixes.
- Implement data quality, governance, security, and performance best practices across enterprise platforms.
- Manage source code using Git and follow CI/CD and enterprise DevOps practices.
- Collaborate with Solution Architects, Product Owners, Business Analysts, and cross-functional engineering teams to translate business requirements into technical solutions.
- Participate in Agile ceremonies including sprint planning, backlog refinement, architecture discussions, code reviews, and retrospectives.
- Create technical documentation, deployment artifacts, testing documentation, and operational runbooks.
- Mentor junior engineers and contribute to engineering standards, best practices, and continuous improvement.
Required Qualifications
- Bachelor's degree in Computer Science, Information Systems, Engineering, or a related technical discipline (or equivalent professional experience).
- 6+ years of experience designing, developing, and supporting enterprise data platforms, cloud analytics solutions, or large-scale data engineering initiatives.
- Strong hands-on experience with Databricks.
- Strong hands-on experience with Apache Spark (PySpark).
- Strong hands-on experience with Python.
- Strong hands-on experience with SQL.
- Strong hands-on experience with Scala.
- Experience with Databricks Workspaces.
- Experience with Delta Lake.
- Experience with Unity Catalog.
- Experience with Delta Live Tables (DLT).
- Experience with Databricks SQL.
- Experience with MLflow.
- Experience with Databricks Jobs.
- Strong SQL development experience, including query optimization and performance tuning.
- Experience designing and implementing ETL/ELT solutions using Azure Data Factory or comparable cloud integration platforms.
- Experience with Apache Kafka or other event streaming technologies.
- Experience integrating enterprise applications using REST APIs, JSON, XML, and SOAP web services.
- Experience with Azure Data Lake Storage (ADLS Gen2) or comparable cloud storage platforms.
- Experience using Git for source code management and collaborative software development.
- Experience working within Linux environments, including shell scripting and command-line utilities.
- Familiarity with Infrastructure as Code (IaC) tools such as Terraform is preferred.
- Experience implementing CI/CD pipelines using Azure DevOps, GitHub Actions, or similar DevOps platforms is preferred.
- Strong analytical, troubleshooting, and problem-solving skills.
- Experience working in Agile environments utilizing Scrum, Kanban, or SAFe methodologies.
- Demonstrated ability to manage multiple priorities while delivering high-quality solutions.
- Proven ability to work independently while mentoring teammates and contributing to technical leadership.
Technology Focus
- Databricks, Apache Spark, PySpark
- Azure Data Factory, Azure Data Lake, Azure Data Lake Storage (ADLS Gen2)
- Delta Lake, Unity Catalog, Delta Live Tables (DLT)
- Databricks SQL, Databricks Jobs, MLflow
- Kafka
- Python, SQL, Scala
- Git, RESTful APIs, SOAP, JSON, XML
- Linux, shell scripting
- Terraform, Azure DevOps, GitHub Actions, CI/CD
- Scrum, Kanban, SAFe
Benefits
- Competitive compensation
- Comprehensive insurance options
- Matching contributions through the 401(k) plan and the share purchase plan
- Paid time off for vacation, holidays, and sick time
- Paid parental leave
- Learning opportunities and tuition assistance
- Wellness and Well being programs
Additional Collaboration and Communication
- Clearly communicate complex technical concepts to technical and non-technical audiences.
- Collaborate with architects, developers, business analysts, product owners, and client stakeholders.
- Build trusted relationships across cross-functional teams while delivering high-quality solutions.
- Foster knowledge sharing, continuous learning, and engineering excellence.