Amazon’s eCommerce Foundation Business Data Technologies (BDT) team builds scalable data platforms that help power the Amazon website and customer experience. In this Data Engineer II role, you’ll design, develop, and run large-scale data structures used for analytics and deep learning, working with data at petabyte scale from thousands of sources across Amazon and supported subsidiaries.
BDT collects and serves data from systems including the Amazon catalog, inventory, customer orders, page views, and Alexa, and also supports Amazon subsidiaries such as IMDb and Audible. Internal customers query this data hundreds of thousands of times per day through AWS Redshift, Hive, and Spark.
What you’ll do
- Design, develop, implement, test, and operate large-scale, high-volume, high-performance data structures for analytics and deep learning.
- Build real-time and batch data ingestion routines using AWS technologies and big data tools, applying best practices across data modeling and ETL/ELT.
- Collect business and functional requirements and translate them into robust, scalable, operable solutions aligned to the broader data architecture.
- Analyze upstream source data systems and contribute best practices to source teams.
- Support the end-to-end development lifecycle, from design and implementation through testing, documentation, delivery, support, and maintenance.
- Create dataset documentation and metadata that are usable for stakeholders and downstream teams.
- Review and make decisions on dataset implementations proposed by peer data engineers.
- Assess and decide on adoption of new or existing software products and tools.
- Mentor junior data engineers.
What we’re looking for
- 3+ years of data engineering experience
- Experience with data modeling, warehousing, and building ETL pipelines
Preferred: Experience with AWS technologies such as Redshift, S3, AWS Glue, EMR, Kinesis, FireHose, Lambda, and IAM roles and permissions; and experience with non-relational databases and data stores including object storage, document or key-value stores, graph databases, and column-family databases.
Tools and technologies
AWS, Redshift, Hive, Spark, EMR, RDBMS, S3, AWS Glue, Kinesis, FireHose, Lambda, and IAM roles and permissions, including real-time and batch ingestion with ETL/ELT, data modeling, non-relational data stores, object storage, document/key-value stores, graph databases, and column-family databases.
Compensation and benefits
Salary: USD 132,100 - 178,800 per year. Location: Seattle, WA (onsite).
- Health insurance (medical, dental, vision, prescription, Basic Life & AD&D insurance, option for Supplemental life plans, EAP, Mental Health Support, Medical Advice Line, Flexible Spending Accounts, Adoption and Surrogacy Reimbursement coverage)
- 401(k) matching
- Paid time off
- Parental leave
- Sign-on payments
- Restricted stock units (RSUs)
About the organization
Amazon’s eCommerce Foundation (eCF) organization delivers core components that support the Amazon website and customer experience. The Business Data Technologies (BDT) group is responsible for collecting petabytes of data from thousands of sources, and enabling fast access and querying for internal customers using AWS services such as Redshift, Hive, and Spark. Solutions are designed to scale as the business grows.