Senior Data Engineer
Backend Developer
Senior
3d Design Tools
Amazon Web Services
Automation
AWS
Aws Cloudformation
Aws Data Lake
Aws Data Warehouse
Aws Fargate
Aws Glue
Big Data
Bigdata
Cloud
Cloud Computing
Cloud Data Warehouse
Cloud Infrastructure
Cloud Native
Cloud Platform
Cloud Platforms
Cloud Technology
Data
Data Analysis
Data Analytics
Data Architecture
Data Engineer
Data Engineering
Data Engineering Lead
Data Integration
Data Lake
Data Lakehouse
Data Pipeline
Data Platform
Data Processing
Data Warehouse
Database
Databases
Databricks
DevOps
DevSecOps
Engineering
ETL
Informatica
Information Technology (IT)
Infrastructure As Code
IT Services
Kubernetes
Platform Engineering
Programming
Programming Language
Programming Languages
Rendering Engines
Security Automation
Spark
SQL
Job Description
Build and evolve an AWS-based data platform in a senior engineering role focused on scalable pipelines and data products.
Responsibilities
- Analyze large, complex data sets
- Architect, design, and develop data pipelines
- Coordinate pipeline generation and delivery across the team
- Collaborate with engineers to plan and solve complex technical problems
- Design new pipelines or optimize existing ones for performance and cost
Requirements
- Expert Python and SQL development
- Strong AWS experience including:
- AWS Lambda
- AWS Glue
- AWS Step Functions (State Machines)
- AWS Fargate
- AWS CloudFormation
- Amazon RedShift
- Strong experience handling semi-structured data formats: XML, CSV, JSON, and Parquet
- Strong automation experience
- Experience with Databricks:
- Delta Lake
- Notebooks
- SQL analytics
- Experience with Kubernetes:
- Cluster orchestration
- Container networking
- Pod management
- Experience with DevOps and CI/CD, including deployment orchestration (Octopus Deploy)
- Experience with agentic AI tooling:
- LangGraph
- MCP Servers
- Agent frameworks
- Experience with data handling and distributed processing:
- Pandas / NumPy
- Apache Spark
- Ray
Technology Stack
- Python, SQL
- AWS Lambda, AWS Glue, AWS Step Functions, AWS Fargate, AWS CloudFormation, Amazon RedShift
- XML, CSV, JSON, Parquet
- Databricks, Delta Lake, Databricks Notebooks, SQL Analytics
- Kubernetes (cluster orchestration, container networking, pod management)
- Octopus Deploy
- LangGraph, MCP Servers, Agent frameworks
- Pandas, NumPy, Apache Spark, Ray
Location and Education
- Troy, MI (hybrid)
- Bachelor's Degree
Benefits
- Hybrid schedule: 3 days onsite in Troy, MI
- No visa sponsorship available at this time
Helpful Experience (Nice to Have)
- Experience with LLMs, machine learning and deep learning libraries, AI application frameworks, vector databases
- Experience with automotive data or the automotive industry
- Experience with Agile / Scrum
About MOTOR Information Systems (Hearst)
- MOTOR is the industry leader in automotive data and insights
- Powers the aftermarket ecosystem using comprehensive datasets
- Investing heavily in modernizing the data platform to support next-generation products and AI-driven capabilities