> Markdown version of [/jobs/ext/613793-senior-data-engineer](https://www.wearedevelopers.com/jobs/ext/613793-senior-data-engineer). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Senior Data Engineer - **Company:** Pinnacle Inc. - **Location:** Arlington, VA, United States - **Experience:** Expert - **Contract:** Permanent contract - **Skills:** Agile Methodology, Artificial Intelligence, Apache HTTP Server, Cloud Computing, Information Engineering, Extract Transform Load (ETL), Data Mining, Data Stores, Elasticsearch, Information Lifecycle Management, Python (Programming Language), Cloud Services, SQL Databases, Unstructured Data, Data Processing, Data Storage Technologies, Data Ingestion, Large Language Models, Apache Spark, Kubernetes, Apache Nifi, Data Management, Data Pipelines - **Published:** June 19, 2026 - **Apply:** https://www.indeed.com/viewjob?jk=e5794b7694730c3f ## About the Role Do you have experience in Tooling?, 10+ years' experience with: · Data lifecycle engineering · Development and maintenance of extract, transform and load (ETL) tools and services · Cloud and on-prem data storage and processing solutions · Python, SQL, Spark and other data engineering programming · COTS and open-source data engineering tools such as ElasticSearch and NiFi · Processing data within the Agile Lifecycle, Desired Skills * Experience using AI/LLMs to accelerate the processing and transformation of data ## Description We are seeking a Data Engineer that will develop new tools, code, and services to execute data engineering activities involving data of varying types and in varying conditions. Activities include the following tasks: Movement of structure and unstructured data using approved methods. Execute data ingestion activities for storing data in a local or enterprise level location. Develop code to format data that supports exploration. Analyze source data formats and work with Data Scientists and partners to determine the formats and transforms that best meet mission objectives. Develop code and tools to provide one-time and on-going data extraction from various repositories, formatting and transformations into enterprise or standalone data models. Develop new ETL and perform O&M and enhancements on existing ETL code using best practices/standards. Develop and deliver documentation for each project including ETL mappings, code use guide, code location and access instructions., Responsibilities: · Design and optimize Data Pipelines using tools such as Spark, Apache Iceberg, Trino, OpenSearch, EMR cloud services, NiFi and Kubernetes containers · Ensure the pedigree and provenance of the data is maintained such that the access to data is protected · Clean and preprocess data to enable access for advanced analytics · Collaborate with enterprise working groups to advance the state of data standards· Collaborate with the engineering team, data stewards, and mission partners to aid in getting actionable value out of the data holdings · Collaborate with software engineers to update, configure, and maintain data services based on the requirements · Ensure data quality by working with the testing and data quality team to enhance standardization of data conditioning pipelines · Experience adapting to various types and formats of data, and working with development teams to integrate new data processing platforms ## Related Videos - [Data Science on Software Data](https://www.wearedevelopers.com/videos/162-data-science-on-software-data) - [Understanding Kubernetes in a visual way](https://www.wearedevelopers.com/videos/100085-understanding-kubernetes-in-a-visual-way) - [Alibaba Big Data and Machine Learning Technology](https://www.wearedevelopers.com/videos/37-alibaba-big-data-and-machine-learning-technology) - [Instant KAI Sandboxes with vCluster: Multi-Tenant, Multi-Scheduler GPU Sharing](https://www.wearedevelopers.com/videos/100333-instant-kai-sandboxes-with-vcluster-multi-tenant-multi-scheduler-gpu-sharing) - [Data Mining Accessibility](https://www.wearedevelopers.com/videos/802-data-mining-accessibility) - [From Messy Queries to Scalable Systems - How Data Engineering actually works](https://www.wearedevelopers.com/videos/100203-from-messy-queries-to-scalable-systems-how-data-engineering-actually-works) ## Related Articles - [Top Big Data Technologies That You Need to Know](https://www.wearedevelopers.com/magazine/108-top-big-data-technologies-that-you-need-to-know) - [Making Data Warehouses Fast: A Developer’s Story](https://www.wearedevelopers.com/magazine/107-making-data-warehouses-fast-a-developer-s-story) - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [How to Become an AI Engineer](https://www.wearedevelopers.com/magazine/331-how-to-become-an-ai-engineer) - [Data Engineer Salary UK](https://www.wearedevelopers.com/magazine/253-data-engineer-salary-uk) - [How We Built a Worry-Free System That Runs for 10+ Years – And What We’d Do Again](https://www.wearedevelopers.com/magazine/751-how-we-built-a-worry-free-system-that-runs-for-10-years-and-what-we-d-do-again)