EngineerJobs.io
← Back to all jobs

Job Description

Responsibilities

  • Lead the end-to-end design, development, and operation of scalable ETL and ELT pipelines to support enterprise and commercial data delivery within the Databricks ecosystem.
  • Develop and operate pipelines that ingest, transform, and publish data across multiple domains.
  • Build and optimize data transformations and validations using Python, PySpark, Scala, and SQL.
  • Implement configuration-driven pipeline frameworks to onboard and manage datasets efficiently.
  • Ensure data products are well-structured, performant, and optimized for downstream consumption.
  • Collaborate with stakeholders to define data contracts, schemas, SLAs, and quality standards.
  • Apply best practices for data reliability, observability, and cost optimization.
  • Contribute to CI/CD practices, including automated testing, deployment, and promotion.
  • Support data governance initiatives, including data quality, lineage, and access controls.

Requirements

  • Minimum five years of experience in data engineering, building and operating production-grade data pipelines.
  • Strong programming proficiency with Python, PySpark, Scala, and SQL.
  • Proven experience designing or working with configuration- or metadata-driven data pipelines.
  • Hands-on experience operating within a Databricks-based data platform.
  • Solid understanding of data modeling, schema evolution, and large-scale dataset management.
  • Experience deploying and operating data solutions in AWS and Azure cloud environments.
  • Working knowledge of CI/CD concepts and integrating data pipelines into automated workflows.
  • Demonstrated proficiency with AI tools to streamline workflows, with awareness of responsible and ethical AI use.

Technologies

  • Python
  • PySpark
  • Scala
  • SQL
  • Databricks
  • AWS
  • Azure

Benefits

  • Medical insurance
  • Dental insurance
  • Vision insurance
  • Parental leave
  • Paid time off
  • 401(k) plan with employee and company contribution opportunities
  • Life insurance
  • Disability insurance
  • Accident insurance
  • Discounted Employee Stock Purchase Plan
  • Tuition reimbursement

About the Team

This team sits at the heart of Moody's Enterprise Data Platform (CORE), supporting a complex and impactful portion of Moody's data estate. It is a modern, innovation-driven data engineering group that powers Moody's extensive financial and corporate databases, enabling data-driven insights and supporting Moody's progression in an increasingly AI-driven industry.

Similar Jobs