Jr Data Engineer

Vforce Infotech
Edison, NJ, United States
30 days ago
Apply on www.indeed.com
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Starter
Compensation
$60,000.0 - $80,000.0
Working hours
Regular working hours
Job source

Tech stack

Java (Programming Language) Agile Methodology Amazon Web Services Amazon S3 Data Analysis Microsoft Azure Bash Shell Big Data Code Review Databases Information Engineering Extract Transform Load (ETL)
+24 more
Data Security Data Systems Data Warehousing Database Design Dimensional Modeling Apache Hadoop Hadoop Distributed File System Apache Hive Python (Programming Language) Scrum Methodology Query Optimization Azure Data Lake Shell Script SQL Databases Data Streaming Talend Cloud Platform System Apache Spark Data Strategy Apache Kafka Data Management Restful APIs Looker Analytics Data Pipelines

Job description

Join our innovative team as a Jr Data Engineer and play a pivotal role in transforming raw data into actionable insights that drive strategic decision-making. In this dynamic position, you will design, develop, and maintain robust data pipelines and architectures within cloud environments, empowering our organization to harness the full potential of big data systems. Your expertise will enable seamless data management, integration, and analysis across diverse platforms, fueling business intelligence and analytics initiatives that make a real impact., * Develop, implement, and optimize scalable ETL (Extract, Transform, Load) pipelines using tools like Informatica, Talend, and custom scripting in Python or Bash to ensure efficient data flow across systems.

  • Design and maintain data models and schemas for data warehouses using dimensional modeling techniques to support business intelligence activities.
  • Build and manage cloud-based data storage solutions such as Azure Data Lake, AWS S3, or other public cloud databases to facilitate secure and scalable data access.
  • Collaborate with cross-functional teams to integrate linked data sources and develop comprehensive data management strategies aligned with enterprise goals.
  • Utilize big data technologies including Hadoop, Apache Hive, Spark, and Kafka to process large datasets efficiently for analysis and model training purposes.
  • Implement best practices for query management, database design, and data warehousing to ensure high performance and reliability of data systems.
  • Support agile development processes by participating in sprint planning, code reviews, and continuous improvement initiatives related to data engineering workflows.

Requirements

  • Proven experience with cloud platforms such as AWS (Amazon Web Services), Azure (Microsoft Azure), or equivalent public cloud environments.
  • Strong proficiency in programming languages including Java, Python, SQL (Structured Query Language), and Shell Scripting for automation and development tasks.
  • Extensive knowledge of big data systems like Hadoop ecosystem components (HDFS, Hive), Spark, Kafka, and related technologies.
  • Hands-on experience designing and implementing ETL pipelines using tools such as Informatica or Talend; familiarity with RESTful API integrations is a plus.
  • Solid understanding of data modeling principles including dimensional modeling for data warehousing solutions.
  • Familiarity with business intelligence tools such as Looker or similar platforms for reporting and visualization purposes.
  • Ability to work within an Agile environment while managing multiple priorities effectively.
  • Excellent analysis skills with a focus on database design, query optimization, and analytics-driven decision-making.

Join us to leverage your technical expertise in a fast-paced environment where innovation meets impact! We’re committed to fostering growth through cutting-edge technology solutions that empower our organization’s success.

Benefits & conditions

$60,000 - $80,000 a year - Full-time, Contract

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.indeed.com
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

3:28 min

Defining big data and machine learning fundamentals

Ayon Roy · LIVE

3:43 min

The enduring legacy of the amazon S3 storage API

Chris Heilmann +3 · LIVE

3:03 min

Career evolution in data engineering and AI platforms

Maria Apazoglou · Coffee With Developers

2:10 min

Why organizations combine big data and machine learning

Ayon Roy · LIVE

3:09 min

Balancing data science skillings alongside systems engineering rigor

Nico Schmidt · LIVE

3:44 min

Automating storage savings with S3 intelligent tiering

Sébastien Stormacq · World Congress 2021

Videos

See all

Related articles

See all