EngineerJobs.io
← Back to all jobs

Job Description

Sentara Hospitals is hiring a fully remote Data Engineer to help build and scale a modern data platform on Databricks. This role emphasizes data quality, governance, and reliable ingestion, so your work contributes directly to trusted data for stakeholders. Enjoy a schedule in first (days) shift and benefits designed to support both your life and growth.

What you’ll get

  • Medical, Dental, Vision plans
  • Adoption, Fertility and Surrogacy Reimbursement up to $10,000
  • Paid Time Off and Sick Leave
  • Paid Parental & Family Caregiver Leave
  • Emergency Backup Care
  • Long-Term, Short-Term Disability, and Critical Illness plans
  • Life Insurance
  • 401k/403B with Employer Match
  • Tuition Assistance of $5,250/year and discounted educational opportunities through Guild Education
  • Student Debt Pay Down of $10,000
  • Reimbursement for certifications and free access to complete CEUs and professional development
  • Pet Insurance
  • Legal Resources Plan
  • Colleagues may earn an annual discretionary bonus if established system and employee eligibility criteria are met

Responsibilities

  • Develop and maintain data pipelines using PySpark and Databricks
  • Use a metadata-driven ingestion framework to onboard new datasets
  • Embed data quality checks and validation rules within pipelines
  • Support ingestion from file-based sources and tools such as Fivetran
  • Manage schema changes, incremental loads, and file processing patterns
  • Contribute to data governance practices including tagging, metadata, and lineage
  • Troubleshoot and resolve pipeline failures and performance issues
  • Collaborate with architects and stakeholders on data onboarding and requirements
  • Follow and contribute to coding standards, reusable components, and best practices

Required qualifications

  • 3 to 5 years of relevant experience (required)
  • Experience in lieu of a Bachelor’s Degree
  • Hands-on experience with PySpark and Databricks
  • Strong SQL skills
  • Experience building ETL/ELT data pipelines
  • Understanding of Delta Lake concepts including merge, schema evolution, and partitions
  • Familiarity with cloud platforms, with Azure preferred
  • Basic experience with Git and version control
  • Exposure to data catalog or governance tools such as DataHub
  • Experience with Fivetran or similar ingestion tools
  • Understanding of data quality and validation concepts
  • Experience working with metadata-driven frameworks
  • Strong problem-solving and debugging skills
  • Ability to work in a structured, framework-driven environment
  • Focus on data quality, not only pipeline execution
  • Willingness to learn and adapt in a fast-evolving data ecosystem

Location: Virginia Beach, VA (fully remote). Remote opportunities available in: Alabama, Delaware, Florida, Georgia, Idaho, Indiana, Kansas, Louisiana, Maine, Maryland, Minnesota, Nebraska, Nevada, New Hampshire, North Carolina, North Dakota, Ohio, Oklahoma, Pennsylvania, South Carolina, South Dakota, Tennessee, Texas, Utah, Virginia, Washington (state), West Virginia, Wisconsin, Wyoming.

Salary range: USD 80,204 - 133,681 per year.

Technologies: PySpark, Databricks, Fivetran, SQL, Delta Lake, Azure, Git, DataHub.

Similar Jobs