Vertage is seeking a Senior Data Engineer to design and build the cloud data foundation that securely and reliably supports NextBrain Off-Prem. The role focuses on building AWS-based pipelines, storage architecture, data models, APIs, and automated data quality controls to serve operational, engineering, analytical, and contextual information.
Responsibilities
- Design and implement the NextBrain cloud data architecture.
- Build scalable ingestion, transformation, storage, and serving pipelines on AWS.
- Develop pipelines for operational, historical, engineering, application, and contextual data.
- Design data models optimized for analytical applications and AI consumption.
- Establish storage patterns across relational, object, time-series, and analytical data stores.
- Develop APIs and services that enable NextBrain tools and agents to retrieve data consistently.
- Establish automated data-quality validation, reconciliation, and monitoring.
- Implement metadata management, lineage, cataloging, and data-governance practices.
- Design data segregation and access-control mechanisms.
- Optimize data pipelines for scalability, performance, reliability, and cost.
- Partner with operational subject-matter experts to validate the meaning and quality of source data.
- Support migration of relevant information from document-based repositories into structured cloud-hosted data services.
- Collaborate with AI/ML engineers to create reliable datasets, feature pipelines, retrieval mechanisms, and knowledge sources.
- Create reusable data integration patterns to accelerate onboarding of additional NextBrain use cases and sites.
Requirements
- Bachelor's degree in Computer Science, Data Engineering, Engineering, or a related field.
- 5+ years of data engineering experience.
- Strong Python and SQL skills.
- Hands-on experience building production data pipelines in AWS.
- Experience with AWS data technologies such as S3, Glue, Lambda, RDS/Aurora, Redshift, Athena, Kinesis, or equivalent technologies.
- Experience with ETL/ELT architecture and orchestration frameworks.
- Strong knowledge of relational and non-relational database design.
- Experience implementing automated data-quality processes.
- Experience building APIs or data services.
- Understanding of data governance, lineage, security, and access-control principles.
Preferred Skills
- Experience with time-series data.
- Experience with streaming and event-driven architectures.
- Experience with industrial IoT, SCADA, historians, or operational technology data.
- Experience with energy-generation assets such as solar, battery storage, wind, or conventional generation.
- Experience designing data architectures supporting AI/ML and generative AI applications.
- Experience with vector databases and retrieval architectures.
- Experience working with high-volume telemetry datasets.
Tech Stack
- Python, SQL
- Amazon S3, AWS Glue, AWS Lambda
- RDS/Aurora, Redshift, Athena
- Amazon Kinesis
- ETL/ELT
Role Scope and Purpose
This Senior Data Engineer role is responsible for the cloud foundation supporting NextBrain Off-Prem. NextBrain relies on large volumes of operational, engineering, analytical, and contextual data. The selected candidate will build the pipelines, storage architecture, data models, APIs, and quality controls required to make that information available to NextBrain applications and AI capabilities, with an emphasis on secure and reliable access.
Location and Salary
Location: United States (onsite)
Compensation: USD 80 - 85 per hourly
Minimum Education
Bachelor's degree in Computer Science, Data Engineering, Engineering, or a related field.
Languages
- English (professional working proficiency)