> Markdown version of [/jobs/ext/221691-it-applications-data-engineer](https://www.wearedevelopers.com/jobs/ext/221691-it-applications-data-engineer). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # IT Applications Data Engineer - **Company:** Pantex, Inc. - **Location:** Amarillo, TX, United States - **Experience:** Expert - **Contract:** Permanent contract - **Skills:** Application Programming Interfaces (APIs), Artificial Intelligence, Amazon Web Services, Software Applications, Cloud Database, Data Architecture, Data Validation, Data Dictionary, Data Infrastructure, Data Integrity, Extract Transform Load (ETL), Data Systems, Data Vault Modeling, Data Warehousing, Dimensional Modeling, Distributed Computing Environment, Python (Programming Language), Machine Learning, Operational Databases, Performance Tuning, DataOps, Software Engineering, SQL Databases, Data Streaming, Data Processing, Data Storage Technologies, Feature Engineering, Data Ingestion, Azure Data Factory, Snowflake, Apache Spark, Software Security, Database Performance, Information Technology, Data Lineage, Production Code, Real Time Data, Apache Kafka, Data Management, Machine Learning Operations, Data Pipelines, Devsecops, Databricks - **Published:** May 19, 2026 - **Apply:** https://www.indeed.com/viewjob?jk=2ed2871270abe8a0 ## About the Role Do you have experience in Technical documentation?, * Bachelor's degree in engineering/science/information technology discipline: Minimum 2 years of relevant experience. * OR Master's degree in engineering/science/information technology discipline. * OR applicants without a bachelor's degree may be considered based on a combination of at least 10 years of completed education and/or relevant experience Department of Energy (DOE) Order 426.2A Requirements, * Minimum 2 years of relevant experience, with at least 1 year specifically focused on building and managing enterprise-grade data pipelines and ELT/ETL processes. * Expert proficiency in Structured Query Language (SQL) and Python or Scala for data manipulation, transformation, and automation. * Demonstrated experience with distributed processing frameworks (e.g., Apache Spark, Databricks, Snowflake) and cloud-based data services (e.g., Azure Data Factory, Amazon Web Services (AWS) Glue). * Master's degree (MS) in a relevant technical field with minimum 3 years of relevant experience. * Relevant certifications from modern data stack providers (e.g., Databricks Certified Data Engineer, Snowflake Certification) or cloud platforms (e.g., Azure Data Engineer Associate). * Experience implementing data streaming technologies (e.g., Kafka, Azure Event Hubs) for near real-time data ingestion. * Prior experience with data modeling concepts (e.g., Dimensional Modeling, Data Vault) as implemented in data warehousing solutions. ## Description A career at Pantex can offer you the opportunity to make a personal impact on our nation. We recognize that excellent employees are absolutely critical for mission success. We are seeking a Data Engineer to design, build, and optimize the high-performance data pipelines and data infrastructure necessary to power enterprise-level Artificial Intelligence (AI), machine learning, and advanced analytics capabilities. Primary duties include building and maintenance of data pipelines, hands-on development of robust, scalable Extract, Load, Transform (ELT)/Extract, Transform, Lead (ETL) processes, integrating diverse operational and analytical data sources, and ensuring the reliability and quality of all data assets used for modeling and reporting. Additional duties may include performance tuning of centralized data platforms, implementing DataOps methodologies, and collaborating closely with Data Architects and Data Scientists to translate data requirements into fully operationalized solutions. Core Responsibilities and Duties * Build, test, and maintain highly scalable data pipelines for both batch and real-time data consumption using cloud-native services and distributed processing frameworks. * Construct and manage data systems within the modern data stack (e.g., Databricks, Snowflake, or equivalent cloud data warehouse/lakehouse solutions) to ensure optimal data accessibility and performance. * Write production-grade code (primarily Python or Scala) to cleanse, transform, and load complex, high-volume data into structured data models defined by the Data Architect. * Implement and manage data quality checks, data lineage tools, and pipeline monitoring to ensure data integrity and immediate detection of failures across the data platform. * Optimize and fine-tune database performance, query efficiency, and cost management across data storage and processing services in a secure cloud environment. * Support Data Science teams by providing efficient access to data, developing feature engineering workflows, and integrating model-ready data into the Machine Learning Operations (MLOps) platform. Other Responsibilities and Duties * Participate in defining and enforcing Data Operations (DataOps) and Development, Security and Operations (DevSecOps) practices for data pipeline automation, testing, and secure deployment. * Maintain comprehensive documentation of data flows, data dictionaries, and operational runbooks for all production data infrastructure. * Collaborate with application development teams and Application Programming Interface (API) developers to establish secure and efficient data ingestion points from source systems. * Work with Data Architects on modeling new data sources to ensure new data assets adhere to established enterprise data standards and architectures. ## Related Videos - [From Messy Queries to Scalable Systems - How Data Engineering actually works](https://www.wearedevelopers.com/videos/100203-from-messy-queries-to-scalable-systems-how-data-engineering-actually-works) - [How Cisco embraced a DevOps culture within its network engineering team](https://www.wearedevelopers.com/videos/99-how-cisco-embraced-a-devops-culture-within-its-network-engineering-team) - [DevSecOps: Injecting Security into Mobile CI/CD Pipelines](https://www.wearedevelopers.com/videos/273-devsecops-injecting-security-into-mobile-ci-cd-pipelines) - [Hacking AI at the Edge of the Indian Ocean](https://www.wearedevelopers.com/videos/100177-hacking-ai-at-the-edge-of-the-indian-ocean) - [DevSecOps: Security in DevOps](https://www.wearedevelopers.com/videos/36-devsecops-security-in-devops) - [Data Science, ML & AI in the Oil and Gas Industry at NDT Global - Dr. Katja Träumner](https://www.wearedevelopers.com/videos/1308-data-science-ml-ai-in-the-oil-and-gas-industry-at-ndt-global-dr-katja-traumner) ## Related Articles - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [Top Big Data Technologies That You Need to Know](https://www.wearedevelopers.com/magazine/108-top-big-data-technologies-that-you-need-to-know) - [Data Engineer Salary UK](https://www.wearedevelopers.com/magazine/253-data-engineer-salary-uk) - [Making Data Warehouses Fast: A Developer’s Story](https://www.wearedevelopers.com/magazine/107-making-data-warehouses-fast-a-developer-s-story) - [The Most Popular IT Jobs on the Market](https://www.wearedevelopers.com/magazine/376-the-most-popular-it-jobs-on-the-market) - [How to Become an AI Engineer](https://www.wearedevelopers.com/magazine/331-how-to-become-an-ai-engineer)