Data Engineer 4
Job Description
Capital One is seeking a Data Engineer 4 to design, develop, implement, and support cloud-first data solutions and pipelines. In this onsite role based in McLean, VA, you will help build scalable and secure data engineering platforms using Python, Spark, and AWS-related and data warehousing technologies such as Databricks and Snowflake.
This position emphasizes collaboration across Agile teams, technical leadership across full-stack systems, and independent delivery of high-quality cloud data solutions. You will also play an ambassador role, translating data engineering outcomes and concepts clearly for internal and external stakeholders.
Key Responsibilities
- Collaborate with and across Agile teams to design, develop, test, implement, and support technical solutions using full-stack development tools and technologies
- Influence a team of developers, data analysts, and data scientists, leveraging experience in machine learning, distributed microservices, lakehouse architecture, and full-stack systems
- Use Python and Spark with open-source relational and NoSQL databases, along with cloud-based data warehousing platforms such as Databricks and Snowflake
- Stay current with data trends by experimenting with and learning new technologies, participating in internal and external technology communities, and mentoring others
- Work with product managers and software engineers to deliver robust cloud-first data solutions that support experiences for millions of Americans
- Independently design, build, and deliver cloud data solutions and applications with limited or no support from supervisors or managers
- Architect and enforce common data engineering design patterns to improve code quality, maintainability, and reusability
- Design and build data pipelines and platforms focused on scalability, resilience, and operational efficiency to maintain performance as data volume and business demands grow
- Implement data security standards, including encryption at rest and in transit and fine-grained access control to support compliance with data privacy regulations
- Communicate technical concepts and data outcomes clearly to drive alignment and shared understanding
Required Qualifications
- Bachelor’s Degree or higher in Computer Science or a related quantitative field (Statistics, Economics, Operations Research, Analytics, Mathematics, Engineering)
- At least 4 years of experience in application development (internship experience does not apply)
- At least 2 years of experience in distributed data
- At least 2 years of experience with SQL
- At least 2 years of experience with one programming language: Python, Java, or Scala
- At least 2 years of experience in data pipeline design and development
- At least 1 year of experience in data modeling and designing end-to-end data solutions using both relational and non-relational database systems
Technologies
- Python, Spark, Databricks, Snowflake, SQL
- EMR, Glue, Airflow, Dagster
- NoSQL, MongoDB, Cassandra, DynamoDB
- Redshift
- Monte Carlo, Splunk
- AWS, Microsoft Azure, Google Cloud
Benefits
- Comprehensive, competitive, and inclusive set of health, financial, and other benefits
Work Location, Salary, and Experience
- Location: McLean, VA (onsite)
- Salary: USD 197,300 - 225,100 per yearly
- Minimum experience: 4 years
Preferred Qualifications
- 7+ years of experience in application development with demonstrated proficiency in Python, SQL, Scala, or Java
- 4+ years of hands-on experience designing, deploying, and operating data workloads in at least one public cloud environment (AWS, Microsoft Azure, or Google Cloud)
- 4+ years of experience building or supporting distributed data or compute workloads using tools such as EMR, Spark, Glue, or Databricks
- 4+ years of experience designing, implementing, and operating real-time or streaming data pipelines
- 2+ years of experience working on data observability (e.g., Monte Carlo, Splunk) or data orchestration tools (e.g., Airflow, Dagster)
- 4+ years of experience working with unstructured or semistructured data using NoSQL databases (e.g., MongoDB, Cassandra, DynamoDB)
- 4+ years of experience designing and supporting data warehousing solutions (e.g., Snowflake, Redshift)
- 2+ years of experience working in an Agile development environment
- 2+ years of experience developing user-centric reusable data products
Immigration / Work Authorization Note: Capital One will not sponsor a new applicant for employment authorization, or offer any immigration related support for this position (including H1B, F-1 OPT, F-1 STEM OPT, F-1 CPT, J-1, TN, E-3, O-1, and any other forms of work authorization that require immigration support from an employer).