> Markdown version of [/jobs/ext/3098939-data-engineer](https://www.wearedevelopers.com/jobs/ext/3098939-data-engineer). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Data Engineer - **Company:** Accenture - **Location:** San Antonio, TX, United States - **Experience:** Experienced - **Salary:** $106,300.0 - $206,200.0 - **Contract:** Permanent contract - **Skills:** Java (Programming Language), Apache Accumulo, Apache HTTP Server, Automation of Tests, Big Data, Cloud Engineering, Configuration Management, Continuous Integration, Information Engineering, Data Stores, Distributed Data Store, Distributed Systems, Apache Hadoop, Hadoop Distributed File System, Java Virtual Machine (JVM), Data Processing, Data Ingestion, Apache Yarn, System Availability, Delivery Pipeline, Apache Spark, SC Clearance, Containerization, Pyspark, Apache Nifi, Cyber Warfare, Data Pipelines - **Published:** September 26, 2026 - **Apply:** https://dejobs.org/x/x/A064AA9982C04E7EB0FB7072FD015414/job/ ## About the Role * 3 years of experience in CI/CD pipeline development and automation * 3 years of experience with cloud infrastructure management and monitoring * 3 years of experience with Infrastructure as Code and configuration management * 3 years of experience with containerization and automated testing Bonus Points if you have * Experience with troubleshooting and continuous process improvement * Knowledge of tuning JVM/GC, YARN, tablet servers, or NiFi performance-aligned with the types of stability and modernization work reflected in internal AFS/Recro big-data engagements * Experience supporting cyber operations or data engineering work in environments using Accumulo, NiFi, and Spark pipelines * Hands-on experience with distributed data platforms including Hadoop and HDFS * Practical experience designing, maintaining, or troubleshooting data ingestion flows using Apache NiFi * Experience building or optimizing Spark-based data processing jobs (PySpark, Scala, or Java) * Familiarity with Accumulo as a high-ingest, low-latency data store, including integration patterns, bulk loading, or iterator development Clearance Must have an active Secret clearance; TS/SCI preferred ## Description * Building, automating, and optimizing the infrastructure and processes that enable rapid, reliable software delivery * Working closely with developers to integrate automation, improve deployment workflows, and monitor application performance to maintain high availability across environments * Fostering a culture of collaboration, continuous improvement, and shared ownership across engineering teams * Proactively identifying bottlenecks, recommending architectural enhancements, and championing practices such as containerization, configuration-as-code, and automated testing * Aligning tooling, processes, and infrastructure with business needs to help teams deliver features faster, respond to issues quickly, and maintain a stable, secure operational posture * Designing and enhancing large-scale data pipelines leveraging Hadoop, HDFS, Apache NiFi, Apache Spark, and Accumulo to support mission-critical data ingestion, processing, and analytics workflows * Optimizing distributed systems performance, improving data reliability, and maintaining resilient ingest and analytic patterns across the Apache ecosystem