EngineerJobs.io
← Back to all jobs

Job Description

Capital One is seeking a Data Engineer 4 to design, develop, implement, and support cloud-first data solutions and pipelines. In this onsite role based in McLean, VA, you will help build scalable and secure data engineering platforms using Python, Spark, and AWS-related and data warehousing technologies such as Databricks and Snowflake.

This position emphasizes collaboration across Agile teams, technical leadership across full-stack systems, and independent delivery of high-quality cloud data solutions. You will also play an ambassador role, translating data engineering outcomes and concepts clearly for internal and external stakeholders.

Key Responsibilities

  • Collaborate with and across Agile teams to design, develop, test, implement, and support technical solutions using full-stack development tools and technologies
  • Influence a team of developers, data analysts, and data scientists, leveraging experience in machine learning, distributed microservices, lakehouse architecture, and full-stack systems
  • Use Python and Spark with open-source relational and NoSQL databases, along with cloud-based data warehousing platforms such as Databricks and Snowflake
  • Stay current with data trends by experimenting with and learning new technologies, participating in internal and external technology communities, and mentoring others
  • Work with product managers and software engineers to deliver robust cloud-first data solutions that support experiences for millions of Americans
  • Independently design, build, and deliver cloud data solutions and applications with limited or no support from supervisors or managers
  • Architect and enforce common data engineering design patterns to improve code quality, maintainability, and reusability
  • Design and build data pipelines and platforms focused on scalability, resilience, and operational efficiency to maintain performance as data volume and business demands grow
  • Implement data security standards, including encryption at rest and in transit and fine-grained access control to support compliance with data privacy regulations
  • Communicate technical concepts and data outcomes clearly to drive alignment and shared understanding

Required Qualifications

  • Bachelor’s Degree or higher in Computer Science or a related quantitative field (Statistics, Economics, Operations Research, Analytics, Mathematics, Engineering)
  • At least 4 years of experience in application development (internship experience does not apply)
  • At least 2 years of experience in distributed data
  • At least 2 years of experience with SQL
  • At least 2 years of experience with one programming language: Python, Java, or Scala
  • At least 2 years of experience in data pipeline design and development
  • At least 1 year of experience in data modeling and designing end-to-end data solutions using both relational and non-relational database systems

Technologies

  • Python, Spark, Databricks, Snowflake, SQL
  • EMR, Glue, Airflow, Dagster
  • NoSQL, MongoDB, Cassandra, DynamoDB
  • Redshift
  • Monte Carlo, Splunk
  • AWS, Microsoft Azure, Google Cloud

Benefits

  • Comprehensive, competitive, and inclusive set of health, financial, and other benefits

Work Location, Salary, and Experience

  • Location: McLean, VA (onsite)
  • Salary: USD 197,300 - 225,100 per yearly
  • Minimum experience: 4 years

Preferred Qualifications

  • 7+ years of experience in application development with demonstrated proficiency in Python, SQL, Scala, or Java
  • 4+ years of hands-on experience designing, deploying, and operating data workloads in at least one public cloud environment (AWS, Microsoft Azure, or Google Cloud)
  • 4+ years of experience building or supporting distributed data or compute workloads using tools such as EMR, Spark, Glue, or Databricks
  • 4+ years of experience designing, implementing, and operating real-time or streaming data pipelines
  • 2+ years of experience working on data observability (e.g., Monte Carlo, Splunk) or data orchestration tools (e.g., Airflow, Dagster)
  • 4+ years of experience working with unstructured or semistructured data using NoSQL databases (e.g., MongoDB, Cassandra, DynamoDB)
  • 4+ years of experience designing and supporting data warehousing solutions (e.g., Snowflake, Redshift)
  • 2+ years of experience working in an Agile development environment
  • 2+ years of experience developing user-centric reusable data products

Immigration / Work Authorization Note: Capital One will not sponsor a new applicant for employment authorization, or offer any immigration related support for this position (including H1B, F-1 OPT, F-1 STEM OPT, F-1 CPT, J-1, TN, E-3, O-1, and any other forms of work authorization that require immigration support from an employer).

Similar Jobs