EngineerJobs.io
← Back to all jobs

Job Description

Vytwo is hiring a Senior Data Engineer to help design and maintain scalable data pipelines and keep Databricks and Spark workloads running efficiently for analytics-ready data. This role is based in Minneapolis, MN (onsite), with flexible work from home options available.

Responsibilities

  • Design, develop, and maintain scalable data pipelines using Apache Spark with PySpark and/or Scala
  • Build and optimize Databricks data workflows, including Delta Lake, notebooks, and scheduled jobs
  • Ingest, transform, and curate large-scale structured and semi-structured datasets
  • Perform performance tuning and cost optimization for Spark workloads and Databricks clusters
  • Implement data quality checks, monitoring, and error handling
  • Collaborate with analytics and business stakeholders to deliver well-modeled, analytics-ready data
  • Support batch processing and, where applicable, streaming data pipelines
  • Follow best practices for testing, documentation, security, and version control

Requirements

  • 5+ years of experience in Data Engineering or a related role
  • Strong hands-on experience with Apache Spark (with PySpark or Scala)
  • Proven experience working in Databricks environments
  • Strong SQL skills and experience with relational and analytical databases
  • Experience building and maintaining ETL/ELT pipelines at scale
  • Familiarity with modern data lake architectures (including Delta Lake preferred)
  • Experience with Git-based version control
  • Must be a U.S. Citizen
  • Must be currently located in Minnesota (MN), or willing to relocate

Technology Focus

Apache Spark, PySpark, Scala, Databricks, Delta Lake, Git, SQL

Similar Jobs