> Markdown version of [/jobs/ext/2565670-data-engineer](https://www.wearedevelopers.com/jobs/ext/2565670-data-engineer). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Data Engineer - **Company:** Vsg Business Solutions - **Location:** Tulsa, OK, United States - **Salary:** $65,000.0 - $90,000.0 - **Contract:** Permanent contract - **Skills:** Java (Programming Language), Airflow, Amazon Web Services, Data Analysis, Big Data, Information Engineering, Extract Transform Load (ETL), Data Structures, Distributed Computing Environment, Python (Programming Language), Systems Architecture, Data Processing, Scripting, Snowflake, Apache Spark, Data Lakes, Pyspark, Production Code, Software Coding, Data Pipelines, Control M - **Published:** August 10, 2026 - **Apply:** https://www.careerjet.com/jobad/us6c24a581f845ac338a7938099a6acd1b ## About the Role * Strong Python programming (core Python, not just PySpark) * Expert-level Apache Spark / PySpark * Hands-on ETL development with large-scale data pipelines * Strong Data Structures & Algorithms * Production experience with distributed data processing * AWS Big Data services (EMR, Glue) * Snowflake * Data Lake * Control-M * Airflow * Strong software engineering fundamentals * Excellent problem-solving and live coding skills * Ability to explain projects and technologies confidently ## Description * Build and manage ETL pipelines for Data Lake and Snowflake * Automate data analysis and aggregation processes * Design scalable data processing solutions * Develop and optimize large-scale Spark/PySpark applications * Contribute to system architecture and production-quality code Additional Notes: * Strong Data Engineering background is mandatory. * Heavy hands-on experience with PySpark, ETL, and Control-M is required. * Java is NOT required. Any scripting language experience is acceptable. * Open to 601, 602, and 603 level candidates. ## Related Videos - [PySpark - Combining Machine Learning & Big Data](https://www.wearedevelopers.com/videos/44-pyspark-combining-machine-learning-big-data) - [From Messy Queries to Scalable Systems - How Data Engineering actually works](https://www.wearedevelopers.com/videos/100203-from-messy-queries-to-scalable-systems-how-data-engineering-actually-works) - [Alibaba Big Data and Machine Learning Technology](https://www.wearedevelopers.com/videos/37-alibaba-big-data-and-machine-learning-technology) - [Data Science in Retail](https://www.wearedevelopers.com/videos/586-data-science-in-retail) - [Enjoying SQL data pipelines with dbt](https://www.wearedevelopers.com/videos/823-enjoying-sql-data-pipelines-with-dbt) - [Empowering Retail Through Applied Machine Learning](https://www.wearedevelopers.com/videos/976-empowering-retail-through-applied-machine-learning) ## Related Articles - [Top Big Data Technologies That You Need to Know](https://www.wearedevelopers.com/magazine/108-top-big-data-technologies-that-you-need-to-know) - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [Data Engineer Salary UK](https://www.wearedevelopers.com/magazine/253-data-engineer-salary-uk) - [Making Data Warehouses Fast: A Developer’s Story](https://www.wearedevelopers.com/magazine/107-making-data-warehouses-fast-a-developer-s-story) - [How We Built a Worry-Free System That Runs for 10+ Years – And What We’d Do Again](https://www.wearedevelopers.com/magazine/751-how-we-built-a-worry-free-system-that-runs-for-10-years-and-what-we-d-do-again) - [The Most Popular IT Jobs on the Market](https://www.wearedevelopers.com/magazine/376-the-most-popular-it-jobs-on-the-market)