Junior Data Engineer

Group Llc
Herndon, VA, United States
2 months ago

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Starter
Compensation
$83,200.0 - $104,000.0
Working hours
Shift work
Job source

Tech stack

Java (Programming Language) Agile Methodology Amazon Web Services Bash Shell Big Data Databases Continuous Delivery Data Infrastructure Data Integration Extract Transform Load (ETL) Relational Databases Database Design
+22 more
Apache Hadoop Apache Hive Unix Shell Machine Learning Microsoft SQL Server Oracle (Applications) Scrum Methodology Azure Data Lake Shell Script SQL Databases Data Streaming Talend Data Processing Scripting Informatica Powercenter Apache Spark Data Analytics Tools for Reporting Restful APIs Looker Analytics Data Pipelines Programming Languages

Job description

Join our innovative team as a Junior Data Engineer and become a vital contributor to our data-driven initiatives! In this energetic role, you will support the development, maintenance, and optimization of our data infrastructure, enabling insightful analysis and strategic decision-making. You’ll work with cutting-edge technologies to build scalable data pipelines, manage large datasets, and collaborate across teams to turn complex data into actionable insights. This position offers a fantastic opportunity for aspiring data professionals eager to grow their skills in a dynamic environment., * Assist in designing, developing, and maintaining data pipelines using ETL (Extract, Transform, Load) processes to ensure seamless data flow across systems.

  • Support the integration of diverse data sources including cloud platforms like AWS and Azure Data Lake, as well as on-premises databases such as Microsoft SQL Server and Oracle.
  • Collaborate with senior engineers to implement data models and optimize database design for efficiency and scalability.
  • Write and maintain SQL queries, Python scripts, Bash (Unix shell), and Shell Scripting to automate data workflows and perform analysis tasks.
  • Contribute to the development of dashboards and reports using Looker or similar analytics tools to visualize key metrics.
  • Assist in managing big data technologies such as Hadoop, Apache Hive, Spark, and Informatica for processing large datasets.
  • Support model training activities by preparing datasets and performing preliminary analysis to inform machine learning projects.
  • Participate in Agile development cycles, including sprint planning, stand-ups, and retrospectives to ensure continuous delivery of high-quality solutions.

Requirements

Do you have experience in Data-driven problem-solving?, * Basic understanding of cloud platforms such as AWS or Azure Data Lake is preferred.

  • Familiarity with programming languages like Java and Python for data manipulation and automation tasks.
  • Knowledge of big data frameworks including Hadoop, Spark, and Apache Hive is advantageous.
  • Experience working with relational databases such as Microsoft SQL Server or Oracle; knowledge of database design principles is a plus.
  • Exposure to ETL tools like Talend or Informatica for data integration projects.
  • Understanding of RESTful APIs for integrating external services into data workflows.
  • Strong analysis skills with the ability to interpret complex datasets and communicate findings clearly.
  • Familiarity with analytics tools like Looker or similar platforms for creating visualizations.
  • Experience working in an Agile environment enhances collaboration and project delivery efficiency. Embark on your journey as a Junior Data Engineer with us-where innovation meets opportunity! We’re committed to fostering your growth through hands-on experience with industry-leading tools and technologies while supporting your professional development every step of the way.

Benefits & conditions

$40 - $50 an hour - Contract, Pulled from the full job description

  • Flexible schedule

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on indeed.com

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

3:04 min

Database evolution and the funding behind vector databases

Erik Bamberg · LIVE

3:28 min

Defining big data and machine learning fundamentals

Ayon Roy · LIVE

3:03 min

Career evolution in data engineering and AI platforms

Maria Apazoglou · Coffee With Developers

4:01 min

Managing application isolation via pluggable database models

Wei Hu Wei Hu · WWC 2022

2:10 min

Why organizations combine big data and machine learning

Ayon Roy · LIVE

3:09 min

Balancing data science skillings alongside systems engineering rigor

Nico Schmidt · LIVE

Videos

See all

Related articles

See all