> Markdown version of [/jobs/ext/586497-data-engineer-full-time-opportunity-not-c2c](https://www.wearedevelopers.com/jobs/ext/586497-data-engineer-full-time-opportunity-not-c2c). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Data Engineer- Full Time Opportunity, NOT C2C - **Company:** Codoxo, Inc. - **Location:** Duluth, GA, United States (Remote available) - **Experience:** Starter - **Contract:** Internship / Graduate position - **Skills:** Clean Code Principles, Artificial Intelligence, Airflow, Amazon Web Services, Amazon Elastic Compute Cloud, Amazon S3, Business Analytics Applications, Cloud Computing, Code Coverage, Software Quality, Information Systems, Data Validation, Information Engineering, Data Governance, Data Infrastructure, Data Integration, Extract Transform Load (ETL), Data Security, Data Warehousing, Relational Databases, Dimensional Modeling, Distributed Computing Environment, Document-Oriented Databases, Identity and Access Management, Python (Programming Language), PostgreSQL, Linux System Administration, Machine Learning, Automation of Marketing, Shell Script, Software Engineering, SQL Databases, Database Optimization, Apache Spark, Database Performance, Git, Pyspark, Information Technology, AWS Glue, Machine Learning Operations, Software Version Control, Data Pipelines - **Published:** June 22, 2026 - **Apply:** https://www.indeed.com/viewjob?jk=780ec295f6bacdf5 ## About the Role Do you have experience in Python?, Do you have a Bachelor's degree?, * Bachelor's degree in Computer Science, Data Engineering, Information Systems, or a related technical field (or equivalent practical experience). * 2+ years of experience in data engineering, software engineering, or related technical roles (internships included). * Proficiency in Python, PySpark and SQL. * Familiarity with ETL/ELT concepts and data pipeline architecture. * Experience working with relational databases such as PostgreSQL. * Basic understanding of cloud computing concepts, preferably AWS. * Exposure to distributed data processing frameworks such as Spark. * Experience working in Linux environments and basic shell scripting. * Strong analytical and problem-solving skills. * Ability to work collaboratively in a team environment under mentorship. * Strong written and verbal communication skills., * Experience working with medical claims data strongly preferred. * Hands-on experience with AWS services such as EC2, S3, Glue, and IAM. * Experience with workflow orchestration tools such as Apache Airflow. * Exposure to data warehousing concepts and dimensional modeling. * Familiarity with CI/CD pipelines and version control (e.g., Git). * Understanding of data security, governance, and compliance best practices. * Experience supporting machine learning pipelines or analytics platforms. * Demonstrated use of AI tools (e.g., code assistants, automation platforms) to improve development efficiency. * Physical Requirements: Work is performed in an office environment (either in our office or work-from home) and requires the ability to work on a computer, operate standard office equipment, and work at a desk. ## Description The Data Engineer supports the design, development, and maintenance of scalable data pipelines that power analytics, reporting, and machine learning initiatives. Working under the guidance of senior engineers, this role contributes to building reliable ETL workflows, optimizing database performance, and integrating structured and unstructured data sources. This position partners closely with data scientists, analysts, and cross-functional stakeholders to ensure timely, accurate, and secure data delivery. By strengthening foundational data infrastructure, the Junior Data Engineer helps advance analytics maturity, enable AI initiatives, and promote data-driven decision-making across the organization. The role consistently leverages AI tools to enhance productivity, code quality, and solution effectiveness., * Assist in designing, building, and maintaining scalable ETL/ELT data pipelines. * Develop and optimize batch and streaming workflows using tools such as AWS Glue, Spark, and Airflow. * Support data integration across multiple structured and unstructured data sources. * Write clean, efficient, and maintainable code in Python, PySpark and SQL. * Monitor, troubleshoot, and improve pipeline reliability and performance. * Optimize database performance, particularly in PostgreSQL and cloud-based environments. * Maintain and support AWS-based infrastructure (EC2, S3, Glue, etc.). * Implement data validation, quality checks, and monitoring processes. * Ensure compliance with data governance, security, and regulatory standards. * Collaborate with data scientists and analysts to translate data requirements into scalable engineering solutions. * Document data flows, architecture decisions, and technical processes. * Use AI-assisted development tools to improve speed, testing coverage, and code quality. ## Related Videos - [From Messy Queries to Scalable Systems - How Data Engineering actually works](https://www.wearedevelopers.com/videos/100203-from-messy-queries-to-scalable-systems-how-data-engineering-actually-works) - [Data Science in Retail](https://www.wearedevelopers.com/videos/586-data-science-in-retail) - [Modern Data Architectures need Software Engineering](https://www.wearedevelopers.com/videos/1030-modern-data-architectures-need-software-engineering) - [PySpark - Combining Machine Learning & Big Data](https://www.wearedevelopers.com/videos/44-pyspark-combining-machine-learning-big-data) - [How a Small Team Shrank a Microsoft Monorepo by 94%](https://www.wearedevelopers.com/videos/1236-how-a-small-team-shrank-a-microsoft-monorepo-by-94) - [Enjoying SQL data pipelines with dbt](https://www.wearedevelopers.com/videos/823-enjoying-sql-data-pipelines-with-dbt) ## Related Articles - [Fully Remote Software Engineer Jobs](https://www.wearedevelopers.com/magazine/447-fully-remote-software-engineer-jobs) - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [Making Data Warehouses Fast: A Developer’s Story](https://www.wearedevelopers.com/magazine/107-making-data-warehouses-fast-a-developer-s-story) - [Data Engineer Salary UK](https://www.wearedevelopers.com/magazine/253-data-engineer-salary-uk) - [The Most Popular IT Jobs on the Market](https://www.wearedevelopers.com/magazine/376-the-most-popular-it-jobs-on-the-market) - [Is Software Engineering Over-Saturated?](https://www.wearedevelopers.com/magazine/418-is-software-engineering-over-saturated)