> Markdown version of [/jobs/ext/918057-senior-software-engineer](https://www.wearedevelopers.com/jobs/ext/918057-senior-software-engineer). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Senior Software Engineer - **Company:** Belay Technologies, Inc - **Location:** Fort Meade, MD, United States - **Experience:** Expert - **Salary:** $190,000.0 - $240,000.0 - **Contract:** Permanent contract - **Skills:** Artificial Intelligence, Amazon Web Services, Cloud Engineering, Encodings, Distributed Systems, Python (Programming Language), Prometheus, Systems Integration, AI Infrastructure, Data Logging, High Performance Computing, System Availability, Grafana, AI Platforms, Kubernetes, BIG-IP Access Policy Manager (APM) - **Published:** June 4, 2026 - **Apply:** https://www.indeed.com/viewjob?jk=2772f398d2529a78 ## About the Role Do you have experience in Systems integration?, Do you have a Bachelor's degree?, * Extensive experience designing, building, and operating large-scale production systems. * Deep expertise in systems integration across diverse technologies and platforms. * Hands-on experience with cloud engineering in AWS. * Advanced proficiency with Kubernetes administration and deployment patterns * Strong Python programming skills. * Experience implementing and scaling observability solutions (APM, OpenTelemetry, Grafana, Prometheus.) * Proven ability to lead technical initiatives and influence organizational change. * Experience developing technical policies and governance frameworks. * Excellent communication, stakeholder management, and leadership skills. * Ability to balance hands-on engineering with leadership and coordination responsibilities. Nice to Haves: * Experience with AI inference serving technologies (vLLM, LiteLLM, etc.). * Previous experience with agentic frameworks (LangChain). * Knowledge of vector databases and embedding systems. * Experience with high-performance computing or distributed systems. * Track record of successfully driving technical and cultural change. ## Description * Design, implement, and optimize infrastructure for AI model inference at scale. * Lead the development and maintenance of production AI services and applications, including retrieval augmented generation (RAG), autonomous agents, and emerging technologies. * Serve as technical lead for AI infrastructure initiatives, coordinating work across integrated teams. * Conduct regular one-on-ones and provide coaching, feedback, and support for assigned team members. * Act as the team point of contact (POC) for contract administration functions. * Navigate ambiguity and define solutions for complex, underspecified systems and requirements. * Establish new technicalpolicies, standards, and governance frameworks where gaps exist. * Drive adoption of new technologies and practices across engineering teams. * Implement and oversee monitoring, logging, and observability solutions for AI services. * Ensure high availability, reliability, performance, and security of AI platform components. * Communicate effectively with stakeholders at multiple organizational levels. Candidates should have the following qualifications: * TS/SCI Clearance with polygraph * 12 yrs., B.S. in a technical discipline or 4 additional yrs. in place of B.S. ## Related Videos - [5 steps for running a Kubernetes environment at scale](https://www.wearedevelopers.com/videos/88-5-steps-for-running-a-kubernetes-environment-at-scale) - [A Brief History of Data Storage](https://www.wearedevelopers.com/videos/974-a-brief-history-of-data-storage) - [Data binning and understanding histograms](https://www.wearedevelopers.com/videos/2086-data-binning-and-understanding-histograms) - [Shipping Faster with Less: Render on Cloud Hosting, AI Workloads, and the Future of DevOps](https://www.wearedevelopers.com/videos/1894-shipping-faster-with-less-render-on-cloud-hosting-ai-workloads-and-the-future-of-devops) - [All your telemetry data from any source in one place](https://www.wearedevelopers.com/videos/57-all-your-telemetry-data-from-any-source-in-one-place) - [Keycloak case study: Making users happy with service level indicators and observability](https://www.wearedevelopers.com/videos/1599-keycloak-case-study-making-users-happy-with-service-level-indicators-and-observability) ## Related Articles - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [How to Become an AI Engineer](https://www.wearedevelopers.com/magazine/331-how-to-become-an-ai-engineer) - [Dev Digest 132 - Binging WADFlix?](https://www.wearedevelopers.com/magazine/473-dev-digest-132-binging-wadflix) - [MLOps And AI Driven Development](https://www.wearedevelopers.com/magazine/82-mlops-and-ai-driven-development) - [Navigating the AI Shift](https://www.wearedevelopers.com/magazine/629-navigating-the-ai-shift) - [Dev Digest 120 - Apple and peers](https://www.wearedevelopers.com/magazine/455-dev-digest-120-apple-and-peers)