> Markdown version of [/jobs/ext/2648743-data-engineer](https://www.wearedevelopers.com/jobs/ext/2648743-data-engineer). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # data engineer - **Company:** EDGESOURCE CORPORATION - **Location:** Chantilly, VA, United States - **Experience:** Experienced - **Contract:** Permanent contract - **Skills:** Training Data, Application Programming Interfaces (APIs), Amazon Web Services, Amazon S3, Business Analytics Applications, Data Analysis, Big Data, Cloud Computing, Configuration Management, Capability Maturity Model Integration, Code Review, Information Systems, Databases, Data Architecture, Data Centers, Data Validation, Information Engineering, Data Files, Data Integrity, Extract Transform Load (ETL), Database Schema, Linux, DOS (Disk Operating System), Internet Security, Python (Programming Language), Linux System Administration, Machine Learning, Microsoft Office, Performance Tuning, Systems Development Life Cycle, Cloud Services, Software Engineering, SQL Databases, Data Streaming, Unstructured Data, Management of Software Versions, Virtualization Technology, Data Processing, Scripting, Computer Network Operations, Data Storage Technologies, Real Time Systems, Feature Engineering, Git, Pyspark, Data Management, Machine Learning Operations, Data Pipelines - **Published:** August 21, 2026 - **Apply:** https://www.careerbuilder.com/job-details/data-engineer-chantilly-va--02e69b18-f6f9-406f-bea2-32ec0f98cad4 ## About the Role * 3-5 years + of professional experience in data engineering or related roles. * Strong collaboration skills to work effectively with Data Scientists, Analysts, and Engineering teams. * Ability to communicate complex technical concepts to non-technical stakeholders. * Detail-oriented, curious, and committed to data quality. * Capable of managing multiple priorities in a fast-paced environment. Technical Skills * Python, SQL, and PySpark (highly desired) for data processing and pipeline development. * Elastic/OpenSearch for search and analytics solutions. * Experience with AWS cloud services and Linux environments. * Git for version control and collaborative development. * Understanding of machine learning workflows and MLOps concepts. Working at Edgesource: As an ISO 9001:2015 certified and CMMI Level 3 appraised small business, Edgesource specializes in providing a variety of technical solutions to include software development, database services, enterprise networking, data center virtualization, and management support. We are always seeking top-talent to join our team in helping to address the most critical technical challenges facing our nation., Amazon Simple Storage Service (S3), Amazon Web Services (AWS), Application Programming Interface (API), Architectural Analysis, Best Practices, Big Data, CISA - Certified Information Systems Auditor, Capability Maturity Model Integration (CMMI), Cataloguing, Cloud Computing, Code Reviews, Communication Skills, Cross-Functional, DOS Operating System, Data Analysis, Data Management, Data Processing, Data Quality, Data Science, Data Sets, Data Storage, Database Extract Transform and Load (ETL), Detail Oriented, Develop and Maintain Customers, Documentation, Flexible Spending Accounts, Git, Homeland Security, ISO 9001, Intelligence Community, Internet Security, Law Enforcement, Linux Operating System, Machine Learning, Maintain Compliance, Microsoft Office, Multitasking, Network Operations Center, Performance Tuning/Optimization, Production Control, Python Programming/Scripting Language, Regulatory Compliance, SQL (Structured Query Language), Scalable System Development, Small Business, Software Development, Source Code/Configuration Management (SCM), Structured Data, Team Player, Technical Support, Training Data Sets, Tuition Fees, United States Department of Defense (DoD), Unstructured Data, Virtualization ## Description Client is seeking a data engineer to grow our team performing cutting edge client work in mission space. With overwhelming amounts of large data, client is seeking a data engineer to support data structuring for implementation into a larger enterprise system that is being custom developed for client., * Design, develop, and maintain ETL/ELT pipelines for batch and real-time processing using Python and SQL. * Integrate data from multiple sources, including databases, APIs, streaming platforms, PDFs, and MS Office files. * Build scalable data architectures to support analytics and machine learning workloads. * Optimize data processing and queries for performance and cost efficiency in AWS S3. * Exposure to PySpark or other big data frameworks is a plus for future pipeline scalability. Data Management & Optimization * Collect, clean, and validate large volumes of structured and unstructured data. * Track data versions, implement data quality checks, and ensure data reliability. * Design and optimize data storage in AWS S3, including raw, intermediate, and final datasets. * Implement data governance practices, including documentation, cataloging, lineage, and security. * Ensure compliance security standards. Collaboration with Data Science & Stakeholders * Work closely with Data Scientists, Analysts, and stakeholders to understand data requirements. * Prepare clean, structured, and feature-ready datasets for analytics and machine learning. * Support feature engineering, aggregations, and transformations at scale. * Assist in deploying ML models to production, ensuring monitoring, versioning, and performance optimization. Documentation & Communication * Document pipelines, data schemas, and transformations clearly. * Communicate technical concepts effectively with cross-functional teams. * Participate in code reviews and promote best practices across the team. ## Related Videos - [From Messy Queries to Scalable Systems - How Data Engineering actually works](https://www.wearedevelopers.com/videos/100203-from-messy-queries-to-scalable-systems-how-data-engineering-actually-works) - [PySpark - Combining Machine Learning & Big Data](https://www.wearedevelopers.com/videos/44-pyspark-combining-machine-learning-big-data) - [Docker network without Docker](https://www.wearedevelopers.com/videos/1418-docker-network-without-docker) - [How a Small Team Shrank a Microsoft Monorepo by 94%](https://www.wearedevelopers.com/videos/1236-how-a-small-team-shrank-a-microsoft-monorepo-by-94) - [Data Science in Retail](https://www.wearedevelopers.com/videos/586-data-science-in-retail) - [Empowering Retail Through Applied Machine Learning](https://www.wearedevelopers.com/videos/976-empowering-retail-through-applied-machine-learning) ## Related Articles - [Data Engineer Salary UK](https://www.wearedevelopers.com/magazine/253-data-engineer-salary-uk) - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [Top Big Data Technologies That You Need to Know](https://www.wearedevelopers.com/magazine/108-top-big-data-technologies-that-you-need-to-know) - [Top-Paying Tech Jobs (with Salaries)](https://www.wearedevelopers.com/magazine/372-top-paying-tech-jobs-with-salaries) - [The Most Popular IT Jobs on the Market](https://www.wearedevelopers.com/magazine/376-the-most-popular-it-jobs-on-the-market) - [Is Software Engineering Over-Saturated?](https://www.wearedevelopers.com/magazine/418-is-software-engineering-over-saturated)