Data Engineer 4 (Manager, IC)
Backend Developer
Manager
Amazon Web Services
Analytics
Apache Airflow
Artificial Intelligence
Automation
AWS
Aws Glue
Big Data
Bigdata
Cloud
Cloud Computing
Cloud Data Engineering
Cloud Data Platform
Cloud Data Warehouse
Cloud Data Warehouse
Cloud Platform
Cloud Platforms
Cyber Security
Cybersecurity Tools
Data
Data Analysis
Data Analytics
Data Architecture
Data Engineer
Data Engineering
Data Integration
Data Lake
Data Lakehouse
Data Pipeline
Data Pipelines
Data Platform
Data Processing
Data Warehouse
Data Warehousing
Database
Databases
Databricks
Dataops
Digital Marketing
ETL
Information Technology (IT)
Integration
IT Services
Log Management
Machine Learning Engineer
Programming
Programming Languages
Project Management
Security Information And Event Management
Security Operations
Snowflake
Spark
Splunk
SQL
Job Description
Capital One is hiring a Data Engineer 4 (Manager, IC) to design, build, and operate cloud-first data platforms and pipelines.
Responsibilities
- Partner with and across Agile teams to design, develop, test, implement, and support full-stack technical solutions
- Influence a team of developers, data analysts, and data scientists with experience in machine learning, distributed microservices, lakehouse architecture, and full-stack systems
- Build data solutions using Python and Spark, supported by open-source relational and NoSQL databases and cloud data warehousing platforms such as Databricks and Snowflake
- Experiment with and learn new technologies; participate in internal and external technology communities; mentor members of the data community
- Collaborate with product managers and software engineers to deliver robust cloud-first data solutions
- Independently design, build, and deliver cloud data solutions and applications with little or no support from supervisors or managers
- Architect and enforce common data engineering design patterns to improve code quality, maintainability, and reusability
- Serve as a data engineering ambassador by communicating technical concepts and data outcomes clearly to internal and external stakeholders
- Design and build pipelines and platforms focused on scalability, resilience, and operational efficiency
- Implement security standards including encryption at rest and encryption in transit plus fine-grained access control to support data privacy compliance
Requirements
- Bachelor’s Degree or higher in Computer Science or a related quantitative field (Statistics, Economics, Operations Research, Analytics, Mathematics, Engineering)
- At least 4 years of experience in application development (internship experience does not apply)
- At least 2 years of experience in distributed data
- At least 2 years of experience with SQL
- At least 2 years of experience with one programming language: Python, Java, or Scala
- At least 2 years of experience in data pipeline design and development
- At least 1 year of experience in data modeling and designing end-to-end data solutions using both relational and non-relational databases
Technologies
- Python, Spark, Databricks, Snowflake, SQL
- NoSQL, machine learning, distributed microservices, lakehouse architecture
- Encryption at rest, encryption in transit, fine-grained access control
- Agile
- EMR, Glue, Airflow, Dagster
- Monte Carlo, Splunk
- MongoDB, Cassandra, DynamoDB, Redshift
- Scala, Java
Benefits
- Performance based incentive compensation, which may include cash bonus(es) and/or long term incentives (LTI)
- Comprehensive, competitive, and inclusive set of health, financial, and other benefits
Preferred Qualifications
- 7+ years of application development experience with demonstrated proficiency in Python, SQL, Scala, or Java
- 4+ years hands-on experience designing, deploying, and operating data workloads in at least one public cloud environment: AWS, Microsoft Azure, or Google Cloud
- 4+ years experience building or supporting distributed data or compute workloads using tools such as EMR, Spark, Glue, or Databricks
- 4+ years experience designing, implementing, and operating real-time or streaming data pipelines
- 2+ years experience in data observability (e.g., Monte Carlo, Splunk) or data orchestration tools (e.g., Airflow, Dagster)
- 4+ years experience with unstructured or semistructured data using NoSQL databases (e.g., MongoDB, Cassandra, DynamoDB)
- 4+ years experience designing and supporting data warehousing solutions (e.g., Snowflake, Redshift)
- 2+ years experience working in an Agile development environment
- 2+ years experience developing user-centric reusable data products
Job Details
- Location: New York, NY (onsite)
- Type: Full-time
- Salary: USD $215,200 - $245,600 per year
- Posted date: 9/24/2026