> Markdown version of [/jobs/ext/1179595-lead-etl-engineer-expert](https://www.wearedevelopers.com/jobs/ext/1179595-lead-etl-engineer-expert). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Lead ETL Engineer (Expert) - **Company:** NIGHTWING LLC - **Location:** Sterling, VA, United States - **Experience:** Expert - **Contract:** Permanent contract - **Skills:** Agile Methodology, Amazon Web Services, Amazon Elastic Compute Cloud, Computer Vision, Microsoft Azure, Bash Shell, Big Data, Cloud Computing, Apache Lucene, Profiling, Code Review, Cyber Security, Information Systems, Continuous Integration, Data Architecture, Data Validation, Information Engineering, Data Files, Data Governance, Data Integrity, Extract Transform Load (ETL), Data Warehousing, Database Queries, Linux, Elasticsearch, Apache Hadoop, Python (Programming Language), NoSQL, Object-Oriented Software Development, Oracle (Applications), Windows PowerShell, Query Optimization, Raw Data, Cloudera, Search Technologies, Software Engineering, Apache Solr, SQL Databases, SQL Server Integration Services, Talend, Unstructured Data, Data Processing, Scripting, Data Ingestion, Azure Data Factory, Informatica Powercenter, Apache Spark, Parallel Computation, Git, Containerization, Information Technology, AWS Glue, Data Delivery, Data Inconsistencies, GPT, Software Version Control, Data Pipelines, Devsecops, Docker, Databricks - **Published:** July 4, 2026 - **Apply:** https://www.juju.com/job/00000000gduxhf ## About the Role Nightwing is currently seeking a detail-oriented and experienced ETL Engineer to join our data engineering team. The ideal candidate will be responsible for designing, building, and maintaining data pipelines and integration processes that enable reliable, timely, and high-quality data delivery across the organization. This role requires strong technical capabilities, a deep understanding of data architecture, and the ability to collaborate effectively with cross-functional stakeholders. The ideal candidate brings deep expertise in Natural Language Processing (NLP), large-scale data processing, and search/discovery systems, with the ability to lead technical teams and translate mission needs into scalable, production-ready capabilities that deliver operational impact. Development will take place in an iterative fashion using Agile development methodology with input from all levels of stakeholders. The candidate must have the ability to communicate with project team members, user community, and leadership to assess changes and demonstrate iterative progress., + Bachelor's degree in Computer Science, Information Systems, Engineering, or related field (or equivalent experience). + Minimum 6-8 years working in Linux Operating system with updating the system for efficient parallel processing, understanding memory, storage and processing data at scale + Minimum 6-8 years in Object Oriented programming. Python is preferred software development language + Minimum 6-8 years of demonstrated experience with applications in the Commercial Cloud Services (C2S) environment or an Amazon Web Services cloud environment. Willing to consider to substitute C2S if candidate has a minimum 4-6 years of cloud computing technology to include Azure, Oracle, Google, etc. + Minimum 4-6 years of demonstrated (Extract, Transform, Load - ETL) with large structured and unstructured raw data sets.Strong experience with ETL tools such as Informatica, Talend, SSIS, AWS Glue, or Azure Data Factory. + Proficiency in SQL, including complex queries and query optimization. + 6-8 Years of experience with AWS platform including understanding EC2, RCS instance types + Strong understanding of data warehousing concepts, data modeling, and schema design. + Hands-on experience with scripting languages such as Python, Bash, or PowerShell. + Familiarity with relational and NoSQL databases. + Experience using version control systems such as Git. Desired Qualifications + Experience working with big data technologies (e.g., Spark, Hadoop, Databricks). + Experience with transformer-based models (e.g., BERT) and modern NLP architectures + Background in document exploitation, e-discovery, or large-scale search platforms + Experience with multi-modal analytics (OCR, image recognition, text + image fusion) + Familiarity with search technologies (Solr, Elasticsearch, Lucene) + Experience with containerization and DevSecOps pipelines (Docker, CI/CD) + Cloudera or similar big data certifications + Experience developing risk scoring, anomaly detection, or predictive analytic models TS/SCI with Polygraph Required Day 1 ## Description + Design, develop, and maintain ETL/ELT pipelines that support data warehouse, analytics, and application needs. Must be experienced with large data sets [hundreds of thousands of records, GB and TB size data sets] + Extract, transform, and load data from various sources into centralized storage solutions. + Design and enhance search and discovery platforms across large volumes of structured and unstructured data + Perform data ingestion, ETL, and integration across enterprise and multi-source environments + Optimize ETL workflows for performance, scalability, and reliability. + Conduct data validation, profiling, and quality checks to ensure accuracy and completeness. + Troubleshoot and resolve data inconsistencies, pipeline failures, or performance bottlenecks. + Build and maintain cloud-native solutions (AWS) aligned to secure and resilient architecture patterns + Partner with mission operators, analysts, and senior stakeholders to define requirements and deliver mission-relevant analytics + Translate mission needs into technical designs, architectures, and implementation roadmaps, ensuring alignment to operational objectives + Deliver clear, compelling visualizations, dashboards, and executive-level briefings that communicate analytic insights and recommendations + Provide technical leadership and mentorship, including hands-on development, code review, and team development + Own delivery of analytic capabilities from concept through deployment, accreditation, and sustainment + Support system accreditation, data governance, and security architecture, ensuring data integrity and compliance within classified environments ## Related Videos - [How a Small Team Shrank a Microsoft Monorepo by 94%](https://www.wearedevelopers.com/videos/1236-how-a-small-team-shrank-a-microsoft-monorepo-by-94) - [ Evaluating AI models for code comprehension](https://www.wearedevelopers.com/videos/1462-evaluating-ai-models-for-code-comprehension) - [Docker network without Docker](https://www.wearedevelopers.com/videos/1418-docker-network-without-docker) - [From Messy Queries to Scalable Systems - How Data Engineering actually works](https://www.wearedevelopers.com/videos/100203-from-messy-queries-to-scalable-systems-how-data-engineering-actually-works) - [Git for Code Reviews](https://www.wearedevelopers.com/videos/429-git-for-code-reviews) - [Modern Data Architectures need Software Engineering](https://www.wearedevelopers.com/videos/1030-modern-data-architectures-need-software-engineering) ## Related Articles - [Data Engineer Salary UK](https://www.wearedevelopers.com/magazine/253-data-engineer-salary-uk) - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [Dev Digest 121 - AI goes offline](https://www.wearedevelopers.com/magazine/456-dev-digest-121-ai-goes-offline) - [Best Coding Boot Camps in Germany](https://www.wearedevelopers.com/magazine/237-best-coding-boot-camps-in-germany) - [Fully Remote Software Engineer Jobs](https://www.wearedevelopers.com/magazine/447-fully-remote-software-engineer-jobs) - [Software Engineer Salary London](https://www.wearedevelopers.com/magazine/252-software-engineer-salary-london)