Data Engineer II
Artificial Intelligence
Automation
Big Data
Bigdata
Bigquery
CI/CD
Cloud
Cloud Data Engineering
Cloud Data Warehouse
Cloud Infrastructure
Cloud Operations
Cloud Platform
Cloud Platforms
Cloud Technology
Data
Data Analysis
Data Analytics
Data Architecture
Data Engineer
Data Engineering
Data Integration
Data Pipeline
Data Platform
Data Processing
Data Warehouse
Data Warehousing
Database
Databases
Databricks
Databricks Pyspark
DevOps
Devops Tools
DevSecOps
ETL
Informatica
Information Technology (IT)
Infrastructure As Code
Integration
Lakehouse
Programming
Programming Languages
Security Automation
Snowflake
Software Development
Spark
SQL
Job Description
CoStar Group’s Data Engineer II role within Homes.com focuses on delivering sitewide tracking architecture and turning incoming data into KPI dashboards and day-to-day insights on consumer behavior. You will design and operate secure cloud data platforms, build scalable pipelines, and support AI and analytics solutions that help teams act on the data they need.
What you’ll do
- Design and oversee implementation of dimensional modeling, database design, and cloud data platform structures using platforms such as Databricks, Snowflake, and BigQuery.
- Build a secure data warehouse and data lakehouse framework applying Medallion Architecture principles.
- Design, develop, and maintain scalable data pipelines and ETL processes using Databricks, Snowflake, and other AWS services.
- Implement and optimize Spark jobs, data transformations, and data processing workflows in Databricks.
- Build and deploy AI/ML models by integrating machine learning into data pipelines, leveraging Databricks ML and AWS ML to develop predictive models and support business insights.
- Apply AI-assisted software development practices to improve engineering productivity and code quality.
- Work across data domains and translate them through data transformations to meet different business needs.
- Oversee automated data infrastructure by partnering with software engineering and product teams to streamline data tracking in products.
- Integrate data from disparate sources via cross-domain data stitching into existing data models.
- Implement data governance, lineage tracking, testing, and monitoring frameworks to help ensure data integrity.
- Perform hands-on data optimization and query performance improvements.
- Liaise with DBA, DevOps, SecOps, and other key teams to support data accessibility and analysis.
- Mentor and manage engineers and analysts to support normal execution of responsibilities.
- Deliver software development projects to specification within scheduled timelines and budget parameters.
- Use AWS DevOps and CI/CD best practices to automate deployment and manage data pipelines and infrastructure using terraform.
Requirements
- Bachelor’s degree required from an accredited, not-for-profit, in-person college or university.
- A track record of commitment to prior employers.
- 3+ years of hands-on experience in software development delivering high-quality solutions.
- Strength in data engineering with strong proficiency in Python, PySpark, and SQL.
- Solid foundation in Object-Oriented Programming principles and best practices.
- Demonstrated success building and launching data-driven products operating at terabyte scale.
- Proven record designing and implementing enterprise-level secure and accurate data platforms.
- Ability to translate technical requirements into robust architecture, data models, and ETL strategies.
- Practical experience with cloud-based databases, including both relational and non-relational systems.
- Knowledge of business intelligence software (for example, Power BI or similar tools).
- Understanding of AI technologies and hands-on experience leveraging AI tooling in development, including Claude Code, GitHub Copilot, Cursor, Gemini, and Grok.
- Ability to retrieve, synthesize, and present critical data in structures useful for answering ad-hoc questions.
- Hands-on data experience with optimization, data quality checks, data accuracy, and query performance improvement.
Technology stack
- Databricks, Snowflake, BigQuery
- Medallion Architecture, Spark
- Databricks ML, AWS ML
- Python, PySpark, SQL
- Power BI
- Claude Code, GitHub Copilot, Cursor, Gemini, Grok
- AWS, terraform, CI/CD
- Object-Oriented Programming
Location: Arlington, VA (onsite)
Compensation: USD 103,000 - 153,000 per year
Benefits
- Comprehensive healthcare coverage: Medical / Vision / Dental / Prescription Drug
- Life, legal, and supplementary insurance
- Virtual and in-person mental health counseling services for individuals and family
- Commuter and parking benefits
- 401(K) retirement plan with matching contributions
- Employee stock purchase plan
- Paid time off
- Tuition reimbursement
- On-site fitness center and/or reimbursed fitness center membership (location dependent), with amenities including yoga studio, Pelotons, personal training, and group exercise classes
- Access to CoStar Group’s Diversity, Equity, & Inclusion Employee Resource Groups
- Complimentary gourmet coffee, tea, hot chocolate, fresh fruit, and other healthy snacks