EngineerJobs.io
← Back to all jobs

Job Description

Startek Pro Technologies is looking for a Data Engineer in Dallas, TX to help design, build, and maintain scalable data pipelines and architectures. In this onsite role, you will focus on data integration and ETL, develop data warehouse and lake pipeline capabilities, and support analytics and machine learning use cases with reliable, governed data.

The position centers on transforming complex, multi-source information into unified datasets that teams can trust for reporting, insights, and downstream applications. You will collaborate with cross-functional partners to turn data requirements into practical, high-performing solutions across distributed and cloud-based environments.

Key Responsibilities

  • Develop, implement, and optimize ETL (Extract, Transform, Load) pipelines to move data efficiently from multiple sources into centralized data warehouses or lakes.
  • Design and maintain data models and schemas using dimensional modeling to support business intelligence tools and analytics workflows.
  • Manage large-scale big data systems using Hadoop, Spark, Hive, and other distributed processing frameworks for complex datasets.
  • Work with cross-functional teams to understand data requirements and implement scalable solutions on AWS, Azure Data Lake, or other public cloud services.
  • Integrate data from diverse systems including Microsoft SQL Server, Oracle, cloud databases, and linked data environments to build unified datasets.
  • Write efficient SQL queries and scripts in Python or Bash to automate, analyze, and troubleshoot data workflows.
  • Support the development of RESTful APIs to enable data access and integration with external applications or services.
  • Maintain data system performance while following data management best practices, security standards, and governance policies.

Requirements

  • Proven experience designing and implementing data warehousing solutions using Informatica, Talend, or similar ETL platforms.
  • Strong proficiency in SQL across databases such as Microsoft SQL Server, Oracle, and cloud-based systems like Azure Data Lake (or equivalents).
  • Hands-on experience with big data technologies including the Hadoop ecosystem (including HDFS), Spark, Hive, and related distributed frameworks.
  • Ability to develop scalable ETL pipelines using Python, Shell scripting (Bash), or other scripting languages in an Agile environment.
  • Familiarity with AWS or Azure Public Cloud for deploying and managing big data systems.
  • Knowledge of data modeling principles, specifically dimensional modeling, for building analytics-focused data warehouses.
  • Experience with business intelligence tools such as Looker (or similar) to convert complex datasets into actionable insights.
  • Strong analytical skills and the ability to troubleshoot issues quickly while maintaining high-quality standards in data management.

Tech Stack

  • ETL, SQL, Python, Bash, RESTful APIs
  • Hadoop, Spark, Hive, HDFS
  • AWS, Azure Data Lake, Azure Public Cloud
  • Informatica, Talend
  • Microsoft SQL Server, Oracle
  • Looker
  • Dimensional modeling

Compensation: USD 45 - 55 per hour. Startek Pro Technologies will have you build and maintain scalable data pipelines and architectures that support data integration across platforms, with high-quality data available for analytics, reporting, and machine learning applications.

Similar Jobs