EngineerJobs.io
← Back to all jobs

Job Description

MyHealthTeam is hiring a Senior Data Engineer (hybrid in Redwood City, CA) to take ownership of the data platform end-to-end. This role builds ingestion, transformation, storage, query, access, and the roadmap that enables other teams to move faster across product, analytics, and operations.

Core Responsibilities

  • Own the data platform end-to-end, including ingestion, transformation, storage, query, and access, and guide the platform roadmap as company data needs grow.
  • Design and evolve batch and streaming pipelines using PySpark/EMR, Kinesis, Lambda, and Step Functions; ingest data from Postgres, Salesforce, third-party vendors, and product event streams into an Iceberg-based lake.
  • Model the warehouse by designing SCD tables, event tables, and the data conventions used by engineers and analysts when introducing new data sources.
  • Collaborate with product, engineering, analytics, and operations stakeholders to scope data requests into reliable pipelines, and create documentation and internal tooling for self-serve enablement.
  • Own the security and compliance foundation for data systems, including audit logging, access control, and temporary-access workflows.
  • Optimize backend query performance at the intersection of data and product code, including read-replica routing, indexing, caching, and IO instrumentation in Java/Spring services.
  • Lead investigations and remediation when data infrastructure misbehaves, such as IOPS spikes, pipeline failures, schema drift, and late data, and implement durable fixes.
  • Use AI daily to accelerate pipeline scaffolding, schema work, and ad-hoc investigations, and ship internal AI tooling for other teams.
  • Mentor engineers and analysts on how to work effectively with the data platform.

Requirements

  • 5+ years of experience building production data pipelines and platforms.
  • Deep proficiency in Python (PySpark) and SQL, including tuning Spark jobs at scale.
  • Comfort and capability working on surrounding services in the data layer ecosystem (Java, Spring Boot).
  • Hands-on experience with distributed compute (Spark/EMR), streaming (Kinesis), and object storage (S3).
  • Strong Postgres fundamentals, including query optimization, indexing, replication, replica routing, and assessing when the database is the bottleneck.
  • Experience with Iceberg and Trino, or similar technologies.
  • Comfort with CI/CD and Terraform.
  • Evidence of using AI in practice, including frontier models, agentic coding tools, or rapid experiments.
  • Track record of working cross-functionally with teams that may not share the same technical focus (for example, product and operations).

Technologies

  • PySpark, EMR, Kinesis, Lambda, Step Functions
  • Postgres, Salesforce, Iceberg, Trino, S3
  • Java, Spring Boot, CI/CD, Terraform
  • SQL

Role Overview

As a Senior Data Engineer, you will own MyHealthTeam’s data platform, the layer other teams rely on to understand the business, ship product features, and meet compliance obligations. The platform you build is intended to create long-term leverage across product, analytics, operations, and data science. This position sits at the intersection of production engineering and data engineering, with success measured by how much faster other teams can deliver because of your work.

Additional Notes

  • The role is open to Senior or Lead candidates; focus is on what you can build rather than current title.

Location and Compensation

  • Location: Redwood City, CA (hybrid)
  • Salary Range: USD 205,000 - 240,000 per year

Similar Jobs