EngineerJobs.io
← Back to all jobs

Job Description

Vforce Infotech is looking for a Data Engineer in Edison, NJ to design, build, and maintain scalable data pipelines and platforms that support advanced analytics, business intelligence, and machine learning initiatives. The role focuses on reliable data movement, modern data management practices, and usable data services across cloud environments.

Responsibilities

  • Develop, implement, and optimize ETL pipelines (Extract, Transform, Load) to move data from multiple source systems into centralized data warehouses and lakes, using tools such as Informatica, Talend, or custom scripting.
  • Design and maintain scalable data models and schemas using dimensional modeling principles to improve query performance in SQL databases, including Microsoft SQL Server and Oracle, as well as cloud platforms such as Azure Data Lake and AWS Redshift.
  • Build and manage big data systems with Hadoop ecosystem technologies, including Apache Hive, Spark, and related components for processing large volumes of structured and unstructured data.
  • Work with cross-functional teams to integrate linked data sources and apply data management best practices across Azure and AWS Cloud environments, including cloud-based storage and analytics services.
  • Create RESTful APIs and automate recurring workflows using Bash (Unix shell scripting) or Python to improve data accessibility and pipeline automation.
  • Support business intelligence efforts by using tools such as Looker to produce dashboards, reports, and visualizations that convert complex datasets into actionable insights for stakeholders.
  • Participate in agile development cycles to enhance data infrastructure with a focus on high availability, security, and compliance with organizational standards.

Requirements

  • Proven experience designing and implementing large-scale big data systems using Hadoop, Spark, Hive, or similar frameworks.
  • Strong proficiency in SQL programming, with extensive knowledge of SQL databases including Microsoft SQL Server, Oracle, and cloud-based options such as Azure Data Lake and other cloud databases.
  • Hands-on experience building ETL pipelines with Informatica or Talend; familiarity with cloud-native ETL solutions is a plus.
  • Solid understanding of data modeling practices, including dimensional modeling for data warehousing.
  • Experience with programming languages such as Python and Java to support software development tasks tied to data engineering.
  • Experience working with AWS and Azure, particularly around public cloud storage solutions and big data services.
  • Knowledge of analytics concepts, including model training, query management, analysis skills, and business intelligence tools such as Looker.
  • Familiarity with Bash, RESTful API development, and Linux/Unix environments to automate processes effectively.

Technologies

ETL, Informatica, Talend, custom scripting, dimensional modeling, SQL, Microsoft SQL Server, Oracle, Azure Data Lake, AWS Redshift, Hadoop, Apache Hive, Spark, RESTful APIs, Bash, Python, Looker, AWS, Azure, AWS Cloud, Linux/Unix

Job Details

Location: Edison, NJ (onsite)
Compensation: USD 60,000 - 80,000 per year

Role Summary

Vforce Infotech seeks a Data Engineer to drive data-driven decision-making by designing, developing, and maintaining scalable data pipelines and systems. This position supports advanced analytics, business intelligence, and machine learning, leveraging big data, cloud technologies, and modern data management practices to deliver actionable insights and reinforce a data-centric culture.

Experience

  • Proven experience designing and implementing large-scale big data systems using Hadoop, Spark, Hive, or similar frameworks.
  • Strong SQL programming skills and extensive experience with SQL databases such as Microsoft SQL Server, Oracle, and cloud-based options like Azure Data Lake.
  • Hands-on ETL pipeline development using Informatica or Talend, with familiarity of cloud-native ETL as a plus.
  • Experience with dimensional modeling and data warehousing data modeling techniques.
  • Programming experience using Python and Java for data engineering development tasks.
  • Experience working with AWS and Azure, with focus on public cloud storage and big data services.
  • Knowledge of analytics concepts, including model training, query management, analysis, and BI tooling such as Looker.
  • Familiarity with Bash, RESTful API development, and Linux/Unix environments for process automation.

Similar Jobs