Data Engineer

Vsg Business Solutions
Tulsa, OK, United States
26 days ago
Apply on www.careerjet.com
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Compensation
$65,000.0 - $90,000.0
Working hours
Regular working hours

Tech stack

Java (Programming Language) Airflow Amazon Web Services Data Analysis Big Data Information Engineering Extract Transform Load (ETL) Data Structures Distributed Computing Environment Python (Programming Language) Systems Architecture Data Processing
+9 more
Scripting Snowflake Apache Spark Data Lakes Pyspark Production Code Software Coding Data Pipelines Control M

Job description

  • Build and manage ETL pipelines for Data Lake and Snowflake
  • Automate data analysis and aggregation processes
  • Design scalable data processing solutions
  • Develop and optimize large-scale Spark/PySpark applications
  • Contribute to system architecture and production-quality code

Additional Notes:

  • Strong Data Engineering background is mandatory.
  • Heavy hands-on experience with PySpark, ETL, and Control-M is required.
  • Java is NOT required. Any scripting language experience is acceptable.
  • Open to 601, 602, and 603 level candidates.

Requirements

  • Strong Python programming (core Python, not just PySpark)
  • Expert-level Apache Spark / PySpark
  • Hands-on ETL development with large-scale data pipelines
  • Strong Data Structures & Algorithms
  • Production experience with distributed data processing
  • AWS Big Data services (EMR, Glue)
  • Snowflake
  • Data Lake
  • Control-M
  • Airflow
  • Strong software engineering fundamentals
  • Excellent problem-solving and live coding skills
  • Ability to explain projects and technologies confidently

Benefits & conditions

  • $65,000-90,000 per year

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.careerjet.com
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

2:15 min

Empowering domain teams with an open data platform

Sandhya Menon Sandhya Menon · World Congress 2026 Europe

3:28 min

Defining big data and machine learning fundamentals

Ayon Roy · LIVE

2:57 min

Core technical practices for robust data engineering

Sandhya Menon Sandhya Menon · World Congress 2026 Europe

4:32 min

Harnessing Spark with Python using PySpark and Py4J

Ayon Roy · LIVE

2:10 min

Why organizations combine big data and machine learning

Ayon Roy · LIVE

2:34 min

Capabilities of the Apache Spark processing engine

Ayon Roy · LIVE

Videos

See all

Related articles

See all