Data Engineer
Job Description
Sentara Hospitals is hiring a fully remote Data Engineer to help build and scale a modern data platform on Databricks. This role emphasizes data quality, governance, and reliable ingestion, so your work contributes directly to trusted data for stakeholders. Enjoy a schedule in first (days) shift and benefits designed to support both your life and growth.
What you’ll get
- Medical, Dental, Vision plans
- Adoption, Fertility and Surrogacy Reimbursement up to $10,000
- Paid Time Off and Sick Leave
- Paid Parental & Family Caregiver Leave
- Emergency Backup Care
- Long-Term, Short-Term Disability, and Critical Illness plans
- Life Insurance
- 401k/403B with Employer Match
- Tuition Assistance of $5,250/year and discounted educational opportunities through Guild Education
- Student Debt Pay Down of $10,000
- Reimbursement for certifications and free access to complete CEUs and professional development
- Pet Insurance
- Legal Resources Plan
- Colleagues may earn an annual discretionary bonus if established system and employee eligibility criteria are met
Responsibilities
- Develop and maintain data pipelines using PySpark and Databricks
- Use a metadata-driven ingestion framework to onboard new datasets
- Embed data quality checks and validation rules within pipelines
- Support ingestion from file-based sources and tools such as Fivetran
- Manage schema changes, incremental loads, and file processing patterns
- Contribute to data governance practices including tagging, metadata, and lineage
- Troubleshoot and resolve pipeline failures and performance issues
- Collaborate with architects and stakeholders on data onboarding and requirements
- Follow and contribute to coding standards, reusable components, and best practices
Required qualifications
- 3 to 5 years of relevant experience (required)
- Experience in lieu of a Bachelor’s Degree
- Hands-on experience with PySpark and Databricks
- Strong SQL skills
- Experience building ETL/ELT data pipelines
- Understanding of Delta Lake concepts including merge, schema evolution, and partitions
- Familiarity with cloud platforms, with Azure preferred
- Basic experience with Git and version control
- Exposure to data catalog or governance tools such as DataHub
- Experience with Fivetran or similar ingestion tools
- Understanding of data quality and validation concepts
- Experience working with metadata-driven frameworks
- Strong problem-solving and debugging skills
- Ability to work in a structured, framework-driven environment
- Focus on data quality, not only pipeline execution
- Willingness to learn and adapt in a fast-evolving data ecosystem
Location: Virginia Beach, VA (fully remote). Remote opportunities available in: Alabama, Delaware, Florida, Georgia, Idaho, Indiana, Kansas, Louisiana, Maine, Maryland, Minnesota, Nebraska, Nevada, New Hampshire, North Carolina, North Dakota, Ohio, Oklahoma, Pennsylvania, South Carolina, South Dakota, Tennessee, Texas, Utah, Virginia, Washington (state), West Virginia, Wisconsin, Wyoming.
Salary range: USD 80,204 - 133,681 per year.
Technologies: PySpark, Databricks, Fivetran, SQL, Delta Lake, Azure, Git, DataHub.