> Markdown version of [/jobs/ext/2642796-junior-data-engineer](https://www.wearedevelopers.com/jobs/ext/2642796-junior-data-engineer). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Junior Data Engineer - **Company:** Sequoia Connect - **Location:** United States (Remote available) - **Experience:** Starter - **Contract:** Internship / Graduate position - **Skills:** Application Programming Interfaces (APIs), Amazon Web Services, Amazon S3, Data Analysis, Cloud Computing, Cloud Database, Cloud Engineering, Program Optimization, Information Engineering, Extract Transform Load (ETL), Data Mining, JSON, Python (Programming Language), Oracle (Applications), Systems Development Life Cycle, DataOps, Amazon Simple Notification Service (SNS), Software Engineering, SQL Databases, Data Logging, Data Processing, Freeform SQL, GitHub Copilot, Gitlab, Pyspark, AWS Glue, Cloudwatch, Amazon Simple Queue Service (SQS), Terraform, Software Version Control, Data Pipelines - **Published:** August 10, 2026 - **Apply:** https://www.indeed.com/viewjob?jk=c4e6f2c9965b8f46 ## About the Role * Degree holders for the visa application process. * 0 to 4 years of software development or data engineering experience across relevant cloud platforms. * Good knowledge of Python and SQL. * Solid understanding of AWS services, specifically S3, SNS/SQS, EMR, Glue, Lambda, Redshift, and Step Functions. * Strong grasp of ETL concepts and data processing fundamentals. * Familiarity with GitLab, Terraform, and the Software Development Life Cycle (SDLC) from development to production. * Familiarity with PySpark, AWS managed services, data engineering best practices, and code optimization. * High-Performance Mindset: Resilience, emotional intelligence, and a focus on agile delivery. * Technologist DNA: A deep understanding of the difference between "coding" and "engineering." Desired * Exposure to PySpark, Athena, CloudWatch, SNS, and SQS. * Internship, project, or academic experience specifically in cloud computing, analytics, or data engineering. * Familiarity with cloud-native foundations or AI coding assistants (e.g., GitHub Copilot). Languages * Advanced Oral English: For seamless collaboration with global teams. * Advanced Spanish., * 0-4 years of software development experience across the appropriate platform. * Good knowledge on Python and SQL. * Good understanding to AWS services such as S3, SNS/SQS, EMR, Glue, Lambda, Redshift, and Step Functions. * Good understanding of ETL concepts and data processing fundamentals. * Familiarity with GitLab/Terraform and SDLC from development to production. * Familiarity to PySpark, AWS managed services, data engineering best practices, and code optimization. * Good analytical, problem-solving, and communication skills. Keywords: Phyton, SQL, PySpark, AWS Glue, ETL ## Description * Assist in building and maintaining ETL pipelines using Python and PySpark. * Support the development of workflows utilizing AWS Glue, Lambda, and Step Functions. * Work extensively with cloud data storage platforms, including S3, Redshift, RDS, and Oracle. * Write complex SQL queries for data extraction, transformation, validation, and reporting. * Help implement basic monitoring, logging, and error handling for data pipelines. * Support the ingestion and processing of data from APIs and JSON payloads. * Collaborate with software engineers, data analysts, and business stakeholders to understand requirements. * Contribute to code management, technical documentation, and deployment support activities. ## Related Videos - [PySpark - Combining Machine Learning & Big Data](https://www.wearedevelopers.com/videos/44-pyspark-combining-machine-learning-big-data) - [Shipping Faster with Less: Render on Cloud Hosting, AI Workloads, and the Future of DevOps](https://www.wearedevelopers.com/videos/1894-shipping-faster-with-less-render-on-cloud-hosting-ai-workloads-and-the-future-of-devops) - [WeAreDevelopers LIVE - Modern DevOps for IoT Devices and More](https://www.wearedevelopers.com/videos/1805-wearedevelopers-live-modern-devops-for-iot-devices-and-more) - [Tips and Tricks for Working with JSON](https://www.wearedevelopers.com/videos/1229-tips-and-tricks-for-working-with-json) - [Data Science in Retail](https://www.wearedevelopers.com/videos/586-data-science-in-retail) - [Empowering Retail Through Applied Machine Learning](https://www.wearedevelopers.com/videos/976-empowering-retail-through-applied-machine-learning) ## Related Articles - [Data Engineer Salary UK](https://www.wearedevelopers.com/magazine/253-data-engineer-salary-uk) - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [Fully Remote Software Engineer Jobs](https://www.wearedevelopers.com/magazine/447-fully-remote-software-engineer-jobs) - [Where to Find Entry-Level Software Engineering Jobs](https://www.wearedevelopers.com/magazine/397-where-to-find-entry-level-software-engineering-jobs) - [Software Engineer Salary London](https://www.wearedevelopers.com/magazine/252-software-engineer-salary-london) - [The Most Popular IT Jobs on the Market](https://www.wearedevelopers.com/magazine/376-the-most-popular-it-jobs-on-the-market)