Junior Data Engineer

Techfield, LLC
Chicago, IL, United States
20 days ago

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Starter
Compensation
$70,000.0 - $80,000.0
Working hours
Regular working hours
Job source

Tech stack

Agile Methodology Amazon Web Services Amazon S3 Bash Shell Big Data Cloud Database Computer Programming Information Engineering Extract Transform Load (ETL) Data Transformation Data Security Data Warehousing
+18 more
Dimensional Modeling Apache Hadoop Apache Hive Python (Programming Language) Machine Learning Azure Data Lake Shell Script Software Engineering Data Streaming Talend Scripting Freeform SQL Apache Spark Vba Programming Language Data Management Restful APIs Looker Analytics Data Pipelines

Job description

  • Design, develop, and optimize large-scale ETL (Extract, Transform, Load) pipelines using tools like Informatica, Talend, or custom scripting with Python and Bash to ensure reliable data flow across systems.
  • Architect and maintain cloud-based data solutions utilizing AWS, Azure Data Lake, and other public cloud platforms to support scalable data warehousing and analytics.
  • Collaborate with cross-functional teams to understand business needs and translate them into effective data models using dimensional modeling techniques for data warehouses.
  • Develop and implement advanced SQL queries for query management, reporting, and business intelligence tools such as Looker.
  • Manage big data systems including Hadoop, Apache Hive, Spark, and related technologies to process vast datasets efficiently.
  • Ensure data quality, security, and compliance by establishing best practices for data management integration and governance.
  • Support model training and analysis efforts by providing clean, well-structured datasets for machine learning initiatives.

Requirements

We are seeking a dynamic and highly skilled Junior Data Engineer to join our innovative data team. In this role, you will lead the design, development, and management of complex data systems that empower data-driven decision-making across the organization. You will leverage your expertise in big data technologies, cloud platforms, and data modeling to build scalable, efficient, and secure data pipelines and warehouses. Your contributions will be instrumental in transforming raw data into actionable insights, supporting strategic initiatives, and driving business growth., * Proven experience as a Data Engineer or similar role with a strong background in software development and data engineering principles.

  • Extensive knowledge of cloud databases such as Azure Data Lake, AWS services (e.g., S3), and other public cloud environments.
  • Hands-on experience with big data systems including Hadoop ecosystem components like Hive and Spark for large-scale data processing.
  • Proficiency in SQL programming along with Python scripting for automation and pipeline development; familiarity with VBA or Shell Scripting is a plus.
  • Strong understanding of data modeling concepts including dimensional modeling for data warehousing design.
  • Experience working with business intelligence tools such as Looker or similar platforms to facilitate analytics reporting.
  • Familiarity with ETL pipeline development using tools like Talend or Informatica; knowledge of RESTful APIs for integrating external systems is advantageous.
  • Ability to work within Agile teams while managing multiple priorities effectively; excellent analysis skills to interpret complex datasets.

Join us to be at the forefront of transforming raw data into strategic assets! We are committed to fostering an energetic environment where your expertise drives innovation and growth across our organization.

Benefits & conditions

Pulled from the full job description 401(k) 401(k) matching Paid time off Vision insurance Dental insurance Life insurance, * 401(k)

  • 401(k) matching
  • Dental insurance
  • Life insurance
  • Paid time off
  • Vision insurance

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.indeed.com

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

3:28 min

Defining big data and machine learning fundamentals

Ayon Roy · LIVE

3:43 min

The enduring legacy of the amazon S3 storage API

Chris Heilmann +3 · LIVE

3:03 min

Career evolution in data engineering and AI platforms

Maria Apazoglou · Coffee With Developers

2:10 min

Why organizations combine big data and machine learning

Ayon Roy · LIVE

1:34 min

Bringing diverse skills to industrial data science roles

Katja Träumner

3:44 min

Automating storage savings with S3 intelligent tiering

Sébastien Stormacq · WWC 2021

Videos

See all

Related articles

See all