Build scalable data pipelines and automation for Google’s GCS Data Science team. In an agile environment, you will help turn fast-changing business needs into analytics and actionable insights by supporting feature engineering and ML/AI model development and scaling. This onsite role is based in New York, NY.
Compensation: USD 106,000 - 151,000 per year. Bonus and benefits: 15% bonus target, equity, benefits, health insurance, and a 401k match.
What you’ll do
- Develop data pipelines, reports, and reusable best practices/frameworks that enable analysts and other stakeholders across the organization.
- Support feature engineering by creating ETL pipeline solutions that meet data requirements for ML/AI model development and enable scaling.
- Apply pipeline and insights rigor across data integrity, test design, analysis, validation, and documentation.
- Create scalable, actionable solutions such as dashboards, automated collateral, and web applications that communicate insights clearly and help advertisers grow.
- Partner with stakeholders to identify feature/tooling gaps and generate improvements for customers.
What you’ll bring
- Bachelor’s degree in Computer Science or a related technical field (or equivalent practical experience).
- 1 year of experience designing data pipelines (ETL) and model data.
- Experience analyzing data and creating reports, plus database query (for example, SQL) and visualization tools (for example, Tableau, dashboards).
- Experience with one or more general purpose programming languages, such as Python, C/C++, or Java.
Preferred qualifications
- Experience in a data science environment supporting feature engineering and model automation needs.
- Experience with big data tools, distributed computing, and non-relational databases.
- Structured problem-solving skills, including the ability to break down ambiguous problems and propose impactful data modeling designs.
- Strong interest in analyzing large, complex datasets and turning them into insights that support business decisions.
About the team and role
The GCS Data Science team is tackling challenging problems for the GCS (Google Customer Solutions) division within the Global Business Organization (GBO). The team aims to build efficient, scalable ML models that help small and midsize businesses grow using Google solutions.
Within Data Science Engineering (DSet), the subteam focuses on the data engineering needs of GCS Data Science teams, including building new data pipelines, automation, observability, and reporting tools. In this role, you will approach big data work in an agile way, using analytical methods to develop a deep understanding of a fast-changing business.
Technologies: ETL, ML/AI, SQL, Tableau, dashboards, Python, C/C++, Java