> Markdown version of [/jobs/ext/248197-software-engineer-ai-infrastructure-swe2-d-26-0123](https://www.wearedevelopers.com/jobs/ext/248197-software-engineer-ai-infrastructure-swe2-d-26-0123). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Software Engineer - AI Infrastructure (SWE2) [D.26.0123] - **Company:** Dover Networks LLC - **Location:** Dover, MD, United States - **Experience:** Expert - **Salary:** $176,000.0 - $192,000.0 - **Contract:** Permanent contract - **Skills:** Artificial Intelligence, Amazon Web Services, Cloud Engineering, Encodings, DevOps, Distributed Systems, Python (Programming Language), Performance Tuning, Prometheus, Systems Integration, Web Applications, AI Infrastructure, Data Logging, High Performance Computing, System Availability, Grafana, AI Platforms, Kubernetes, Infrastructure Automation Frameworks, BIG-IP Access Policy Manager (APM) - **Published:** May 22, 2026 - **Apply:** https://www.clearancejobs.com/jobs/8929880/software-engineer-ai-infrastructure-swe2-d260123 ## About the Role * Proven experience building and maintaining production systems at scale. * Experience with high-volume web application architecture and performance optimization. * Strong background in systems integration across diverse technologies and platforms. * Hands-on experience with cloud engineering in AWS. * Proficiency with Kubernetes administration and deployment patterns. * Strong Python programming skills. * Experience implementing observability solutions (APM, OpenTelemetry, Grafana, Prometheus). * Familiarity with CI/CD pipelines and DevOps practices. * Strong change management and organizational influence skills. * Ability to thrive in ambiguous environments and create structure where needed. * Excellent communication and collaboration skills. Nice to Haves: * Experience with AI inference serving technologies (vLLM, LiteLLM, etc.). * Previous experience with agentic frameworks (LangChain). * Knowledge of vector databases and embedding systems. * Experience with high-performance computing or distributed systems. YOE Requirement: 8 yrs., B.S. in a technical discipline or 4 additional yrs. in place of B.S. ## Description Description: Join us in building the next generation of AI infrastructure that will power innovation across the customer organization. We're seeking a full-stack software engineer to support our AI infrastructure team. In this role, you'll help build and maintain the platform that provides the foundation for the customer's AI capabilities, focusing on inference services while supporting the broader ecosystem of AI-enabled applications. This role is intended for experienced engineers who can independently design, implement, and operate scalable AI infrastructure components. Responsibilities: * Design, implement, and optimize infrastructure for AI model inference at scale. * Support the development and maintenance of production AI services and applications, including retrieval augmented generation (RAG), autonomous agents, and emerging technologies. * Navigate ambiguity and define solutions for underspecified systems and requirements . * Drive adoption of new technologies and practices across engineering teams. * Implement monitoring, logging, and observability solutions for AI services. * Automate infrastructure provisioning and configuration using IaC principles. * Ensure high availability, reliability, and performance of AI platform components * Contribute to security best practices for AI systems and data. * Provide technical guidance and informal mentorship to junior engineers. ## Related Videos - [From DevOps to Scaled DevOps: How We’re Rebuilding Continuous Delivery as a Platform](https://www.wearedevelopers.com/videos/100018-from-devops-to-scaled-devops-how-we-re-rebuilding-continuous-delivery-as-a-platform) - [5 steps for running a Kubernetes environment at scale](https://www.wearedevelopers.com/videos/88-5-steps-for-running-a-kubernetes-environment-at-scale) - [A Brief History of Data Storage](https://www.wearedevelopers.com/videos/974-a-brief-history-of-data-storage) - [Shipping Faster with Less: Render on Cloud Hosting, AI Workloads, and the Future of DevOps](https://www.wearedevelopers.com/videos/1894-shipping-faster-with-less-render-on-cloud-hosting-ai-workloads-and-the-future-of-devops) - [#90DaysOfDevOps - The DevOps Learning Journey](https://www.wearedevelopers.com/videos/548-90daysofdevops-the-devops-learning-journey) - [DevOps for AI: running LLMs in production with Kubernetes and KubeFlow](https://www.wearedevelopers.com/videos/1222-devops-for-ai-running-llms-in-production-with-kubernetes-and-kubeflow) ## Related Articles - [How to Become an AI Engineer](https://www.wearedevelopers.com/magazine/331-how-to-become-an-ai-engineer) - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [Dev Digest 120 - Apple and peers](https://www.wearedevelopers.com/magazine/455-dev-digest-120-apple-and-peers) - [Dev Digest 132 - Binging WADFlix?](https://www.wearedevelopers.com/magazine/473-dev-digest-132-binging-wadflix) - [Dev Digest 121 - AI goes offline](https://www.wearedevelopers.com/magazine/456-dev-digest-121-ai-goes-offline) - [Navigating the AI Shift](https://www.wearedevelopers.com/magazine/629-navigating-the-ai-shift)