Data Engineer II
Job Description
Hybrid work, competitive compensation, and a mission-driven culture await. This Data Engineer II role at Inland Empire Health Plan blends hands-on data engineering with cross-functional collaboration to enhance data reliability, accessibility, and analytics across the enterprise. The position offers a salary range of $104,041.60 to $137,841.60 per year, a hybrid schedule in Rancho Cucamonga, CA, comprehensive medical, dental, and vision coverage, retirement benefits, and ample opportunities for professional development. You will partner with teams to design scalable data warehouse solutions that empower data-driven decision making for IEHP and its stakeholders.
- Competitive salary
- Hybrid schedule
- State-of-the-art fitness center on site
- Medical insurance with dental and vision
- Life, short-term, and long-term disability options
- Career advancement opportunities and professional development
- Wellness programs that promote a healthy work-life balance
- Flexible Spending Account for Health Care and Childcare
- CalPERS retirement
- 457(b) option with a contribution match
- Paid life insurance for employees
- Pet care insurance
Responsibilities
- Design and develop data warehouse ETL solutions using SQL Server Integration Services (SSIS), Azure Data Factory, Synapse Analytics, Azure Databricks, and PySpark ETL.
- Develop and implement data collection processes to support the data warehouse.
- Source data from legacy systems to feed a centralized data warehouse and reporting platform.
- Deliver technical solutions to meet Data Warehouse, BI and Analytics requirements.
- Collaborate with informaticists and analysts to translate analytic requirements into technical designs.
- Contribute to data integration strategies and visions.
- Create and maintain scalable data pipeline architectures based on microservices, aligned to platform and application needs.
- Partner with the Data Engineering, BI & Analytics teams, the Data Warehouse Architect, and the Data Systems Architect to craft data and analytics solutions that improve usability, completeness, and accuracy of enterprise data.
- Engage with stakeholders including Executives, Product, Data and Design teams to address data infrastructure needs and data-related technical issues.
- Analyze user requirements and translate them into database requirements, implementing them in database code.
- Provide detailed analysis of data issues, data mapping, and automation to enhance data quality.
- Assemble large, complex data sets that meet functional and non-functional business requirements.
- Identify, design, and implement process improvements such as automating manual pipelines, optimizing data ingestion and consumption, and redesigning infrastructure for greater scalability and microservices.
- Create, maintain, and optimize SQL queries and routines.
- Analyze potential data quality issues to determine root causes and implement effective solutions.
- Develop, adopt, and enforce Data Warehouse and ETL standards and architecture.
- Monitor and support ETL processes ensuring data integrity and seamless integration of all data sources.
- Facilitate problem management and clear communication among data architects, managers, informaticists and analysts.
- Perform source-to-target mapping validations.
- Identify, document, and execute unit test cases and scripts; participate in peer reviews and document results.
- Provide ongoing proactive technical support for ETL and the data warehouse to ensure business continuity.
- Perform other duties as required to support Health Plan operations and department needs.
Requirements
- Four years of relevant work experience.
- Experience and knowledge in logical, rational, dimensional, and physical data modeling.
- Background in database systems with strong SQL capabilities.
- Experience with orchestration tools, Azure DevOps, and CI/CD.
- Intermediate experience with the following tools and technologies:
- Azure Data Catalogue / Purview
- Azure Cloud
- Databricks
- Power BI Dataflows
- Power Query
- Azure Cosmos
- Azure Monitor
- PowerShell
- Python
- Development experience using PySpark, Spark, Hadoop, Kubernetes, and RDMIS is highly desired.
- Bachelor's degree from an accredited institution is required; Master's degree is preferred.
- Azure Data Engineering Certification is preferred.
Technologies
- SQL Server Integration Services (SSIS)
- Azure Data Factory
- Synapse Analytics
- Azure Databricks
- PySpark
- SQL, T-SQL
- Power BI Dataflows, Power Query
- Azure Cosmos, Azure Monitor
- PowerShell, Python
- Databricks, Spark, Hadoop, Kubernetes
- RDMIS, MS SQL Server
- Azure Data Catalogue, Purview
- Azure Cloud, Azure DevOps
Work Model and Location
This position follows a hybrid work model with remote work on Mondays and Fridays and on-site collaboration Tuesday through Thursday in Rancho Cucamonga, California.
Pay range: USD 104,041.60 - 137,841.60 per year.