DeveloperJobs.io
← Back to all jobs

Job Description

Vforce Infotech is hiring a Data Engineer for an onsite role in Edison, NJ. This position focuses on building reliable, scalable data pipelines and systems that support advanced analytics, business intelligence, and machine learning initiatives. You will help enable a more data-centric culture by turning large volumes of structured and unstructured information into actionable insights.

Compensation for this role is USD 60,000 to 80,000 per year.

What you’ll do

  • Develop, implement, and optimize ETL pipelines to move data from diverse sources into centralized data warehouses and lakes, using tools such as Informatica, Talend, or custom scripting.
  • Design and maintain scalable data models and schemas using dimensional modeling principles to support efficient query performance in SQL databases like Microsoft SQL Server and Oracle, as well as cloud environments such as Azure Data Lake and AWS Redshift.
  • Build and manage big data systems with the Hadoop ecosystem, including Apache Hive and Spark, to process large volumes of structured and unstructured data.
  • Collaborate with cross-functional teams to integrate linked data sources and apply data management best practices across public cloud platforms including Azure Data Lake and AWS Cloud.
  • Create RESTful APIs and automate workflows using Bash (Unix shell scripting) or Python to improve data accessibility and pipeline automation.
  • Support business intelligence with Looker by building dashboards, reports, and visualizations that help stakeholders understand complex datasets.
  • Work through agile development cycles to continuously improve data infrastructure while maintaining high availability, security, and compliance with organizational standards.

What you bring

  • Proven experience designing and implementing large-scale big data systems using Hadoop, Spark, Hive, or similar frameworks.
  • Strong SQL skills with extensive knowledge of SQL databases including Microsoft SQL Server and Oracle, and cloud-based options such as Azure Data Lake or cloud databases.
  • Hands-on ETL experience with Informatica or Talend, with familiarity with cloud-native ETL solutions considered a plus.
  • Solid understanding of data modeling, including dimensional modeling for data warehousing.
  • Expertise in Python and Java for data engineering software development tasks.
  • Experience with AWS and Azure, especially public cloud storage solutions and big data services.
  • Knowledge of analytics concepts, including model training, query management, analysis skills, and BI tools such as Looker.
  • Familiarity with Bash, RESTful API development, and Linux/Unix environments for process automation.

Technologies: ETL, Informatica, Talend, custom scripting, dimensional modeling, SQL, Microsoft SQL Server, Oracle, Azure Data Lake, AWS Redshift, Hadoop, Apache Hive, Spark, RESTful APIs, Bash, Python, Looker, AWS, Azure, AWS Cloud, Linux/Unix

Similar Jobs