> Markdown version of [/jobs/ext/1429609-data-engineer-top-secret-clearance](https://www.wearedevelopers.com/jobs/ext/1429609-data-engineer-top-secret-clearance). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Data Engineer - Top Secret Clearance - **Company:** Malik Consulting, Inc. - **Location:** St. Louis, MO, United States - **Experience:** Expert - **Salary:** $140,000.0 - $170,000.0 - **Contract:** Permanent contract - **Skills:** Application Programming Interfaces (APIs), Agile Methodology, Artificial Intelligence, Amazon Web Services, System Configuration, Data as a Services, Data Architecture, Data Validation, Data Governance, Data Security, Elasticsearch, Fault Tolerance, Iterative and Incremental Development, Machine Learning, Meta-Data Management, MySQL, Node.Js, OpenShift, Scrum Methodology, Search Technologies, Software Deployment, Data Streaming, Software Vulnerability Management, Datadog, System Availability, Delivery Pipeline, Software Security, Kubernetes Helm Charts, Change Data Capture, Backend, Debezium, Kubernetes, Apache Flink, Apache Kafka, Data Management, NestJS, Stream Processing, Data Pipelines - **Published:** July 24, 2026 - **Apply:** https://www.disabledperson.com/jobs/73831769-data-engineer-top-secret-clearance ## About the Role * 7+ years of experience building scalable, real-time data streaming architectures and CDC patterns. * Hands-on experience with Apache Kafka and deep proficiency in writing complex stream processing jobs using Apache Flink. * Extensive experience with OpenSearch (or Elasticsearch), including cluster management, writing ingest pipelines, managing index templates, and writing complex, optimized search queries. * Applied knowledge of integrating AI/ML models into data pipelines, working with vector databases (e.g., OpenSearch k-NN), or building AI-driven data products. * Experience with Debezium for CDC and Vector (by Datadog) for observability and data routing. * Proven experience deploying applications in Kubernetes/OpenShift environments with strong familiarity with infrastructure-as-code and deployment workflows using Helm and ArgoCD. * Ability to work closely with software engineers (particularly those using Node.js/NestJS) to define data contracts and query patterns. * Ability to design, develop, and operate highly available data services across availability zones and regions. * Self-starter with strong problem-solving, analytical, decision-making, and verbal and written communication skills. Preferred Skills * Experience working with cloud platforms such as AWS. * Support for data quality, data validation, metadata management, and data governance practices. * Familiarity with secure software development practices, vulnerability remediation, access control, and compliance requirements in federal or classified environments. * Familiarity with Agile, Scrum, SAFe, or other iterative development methodologies, with experience delivering solutions to government customers. * Experience serving in an "on-call" role supporting emergency response to application or system issues on occasion. Certifications: * Security+ certification is preferred. * Other relevant certifications include CCNA, CCNP, CISA, CISSP, and CISM. Years of Experience: 5 years+ Education: Bachelors Degree ## Description As a Data Engineer, you will work closely with O&M, development, product, design, and client teams to maintain, enhance, and deliver secure data infrastructure and search capabilities across client domains. The role requires experience building scalable, real-time data streaming architectures, designing CDC pipelines, and ensuring data is accessible, reliable, and mission-ready. You will help build, modernize, and sustain data workflows using AWS Cloud, OpenShift, and related technologies while supporting secure deployment, system performance, and continuous improvement. Responsibilities * Design, build, and maintain robust Change Data Capture (CDC) pipelines extracting data from MySQL databases using Debezium and routing to Apache Kafka. * Develop and optimize Apache Flink applications to consume Debezium topics, perform complex multi-table joins, and output denormalized records back to Kafka. * Configure and manage Vector pipelines to consume Flink-processed Kafka topics, perform object remapping, and reliably sink data into OpenSearch. * Architect OpenSearch data management workflows, including the design and implementation of custom ingest pipelines, index templates, and lifecycle policies. * Leverage AI skills to enhance data enrichment processes, implement vector/semantic search capabilities within OpenSearch, and support advanced analytics. * Deploy, scale, and maintain data infrastructure on OpenShift using Helm charts and ArgoCD following GitOps best practices. * Monitor system health, tune performance across the entire data streaming lifecycle (Kafka, Flink, OpenSearch), and ensure high availability and fault tolerance. * Collaborate with backend engineers to ensure OpenSearch indexes are highly optimized for performant querying by downstream NestJS applications. * Perform root cause analysis for system, application, data pipeline, and end-user issues, assisting Tier 2 support teams with complex problem resolutions. * Maintain technical documentation related to data architecture, data flows, APIs, system configurations, and operational procedures while supporting system security coordination. * Support occasional after-hours and weekend work for operational issue resolution, production deployments, data migrations, and maintenance windows. ## Related Videos - [MySQL Protocol Features You Should Be Aware Of](https://www.wearedevelopers.com/videos/100267-mysql-protocol-features-you-should-be-aware-of) - [Nest.js - TypeScript in the backend can also be clean](https://www.wearedevelopers.com/videos/1033-nest-js-typescript-in-the-backend-can-also-be-clean) - [Debugging in the Dark](https://www.wearedevelopers.com/videos/1658-debugging-in-the-dark) - [Modern Data Architectures need Software Engineering](https://www.wearedevelopers.com/videos/1030-modern-data-architectures-need-software-engineering) - [Coding for Good: Achieving social change with an app](https://www.wearedevelopers.com/videos/1645-coding-for-good-achieving-social-change-with-an-app) - [From Messy Queries to Scalable Systems - How Data Engineering actually works](https://www.wearedevelopers.com/videos/100203-from-messy-queries-to-scalable-systems-how-data-engineering-actually-works) ## Related Articles - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [How We Built a Worry-Free System That Runs for 10+ Years – And What We’d Do Again](https://www.wearedevelopers.com/magazine/751-how-we-built-a-worry-free-system-that-runs-for-10-years-and-what-we-d-do-again) - [Data Engineer Salary UK](https://www.wearedevelopers.com/magazine/253-data-engineer-salary-uk) - [Top Big Data Technologies That You Need to Know](https://www.wearedevelopers.com/magazine/108-top-big-data-technologies-that-you-need-to-know) - [Dev Digest 134 - Where pixels sing?](https://www.wearedevelopers.com/magazine/477-dev-digest-134-where-pixels-sing) - [Is Software Engineering Over-Saturated?](https://www.wearedevelopers.com/magazine/418-is-software-engineering-over-saturated)