> Markdown version of [/jobs/ext/194662-senior-software-engineer-3-ai-infrastructure-aws-kubernetes](https://www.wearedevelopers.com/jobs/ext/194662-senior-software-engineer-3-ai-infrastructure-aws-kubernetes). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Senior Software Engineer 3 - (AI Infrastructure, AWS, Kubernetes) - **Company:** Akina, Inc. - **Location:** Annapolis Junction, MD, United States - **Experience:** Expert - **Salary:** $232,000.0 - $283,000.0 - **Contract:** Permanent contract - **Skills:** Artificial Intelligence, Amazon Web Services, Cloud Engineering, Encodings, Distributed Systems, Python (Programming Language), Prometheus, Systems Integration, AI Infrastructure, Data Logging, High Performance Computing, System Availability, Grafana, AI Platforms, Kubernetes, BIG-IP Access Policy Manager (APM) - **Published:** May 29, 2026 - **Apply:** https://www.indeed.com/viewjob?jk=3285626b9f3db2bf ## About the Role Do you have experience in Systems integration?, Do you have a Bachelor's degree?, * Extensive experience designing, building, and operating large-scale production systems. * Deep expertise in systems integration across diverse technologies and platforms. * Hands-on experience with cloud engineering in AWS. * Advanced proficiency with Kubernetes administration and deployment patterns. * Strong Python programming skills. * Experience implementing and scaling observability solutions (APM, OpenTelemetry, Grafana, Prometheus). * Proven ability to lead technical initiatives and influence organizational change. * Experience developing technical policies and governance frameworks. * Excellent communication, stakeholder management, and leadership skills. * Ability to balance hands-on engineering with leadership and coordination responsibilities. Position Desired Skills: * Experience with AI inference serving technologies (vLLM, LiteLLM, etc.). * Previous experience with agentic frameworks (LangChain). * Knowledge of vector databases and embedding systems. * Experience with high-performance computing or distributed systems. * Track record of successfully driving technical and cultural change. ## Description Join us in building the next generation of AI infrastructure that will power innovation across the customer organization. We're seeking a senior full-stack software engineer to support our AI infrastructure team. In this role, you'll lead the development and operation of critical AI platform components, with a focus on scalable inference services and the broader AI application ecosystem. This role includes project leadership responsibilities and people care for a small, integrated team within a larger AI platform organization. * Design, implement, and optimize infrastructure for AI model inference at scale. * Lead the development and maintenance of production AI services and applications, including retrieval augmented generation (RAG), autonomous agents, and emerging technologies. * Serve as technical lead for AI infrastructure initiatives, coordinating work across integrated teams. * Conduct regular one-on-ones and provide coaching, feedback, and support for assigned team members. * Act as the team point of contact (POC) for contract administration functions. * Navigate ambiguity and define solutions for complex, underspecified systems and requirements. * Establish new technical policies, standards, and governance frameworks where gaps exist. * Drive adoption of new technologies and practices across engineering teams. * Implement and oversee monitoring, logging, and observability solutions for AI services. * Ensure high availability, reliability, performance, and security of AI platform components. * Communicate effectively with stakeholders at multiple organizational levels. ## Related Videos - [5 steps for running a Kubernetes environment at scale](https://www.wearedevelopers.com/videos/88-5-steps-for-running-a-kubernetes-environment-at-scale) - [A Brief History of Data Storage](https://www.wearedevelopers.com/videos/974-a-brief-history-of-data-storage) - [Data binning and understanding histograms](https://www.wearedevelopers.com/videos/2086-data-binning-and-understanding-histograms) - [Shipping Faster with Less: Render on Cloud Hosting, AI Workloads, and the Future of DevOps](https://www.wearedevelopers.com/videos/1894-shipping-faster-with-less-render-on-cloud-hosting-ai-workloads-and-the-future-of-devops) - [All your telemetry data from any source in one place](https://www.wearedevelopers.com/videos/57-all-your-telemetry-data-from-any-source-in-one-place) - [Keycloak case study: Making users happy with service level indicators and observability](https://www.wearedevelopers.com/videos/1599-keycloak-case-study-making-users-happy-with-service-level-indicators-and-observability) ## Related Articles - [How to Become an AI Engineer](https://www.wearedevelopers.com/magazine/331-how-to-become-an-ai-engineer) - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [MLOps And AI Driven Development](https://www.wearedevelopers.com/magazine/82-mlops-and-ai-driven-development) - [Navigating the AI Shift](https://www.wearedevelopers.com/magazine/629-navigating-the-ai-shift) - [Stephan Gillich - Bringing AI Everywhere](https://www.wearedevelopers.com/magazine/489-stephan-gillich-bringing-ai-everywhere) - [Coffee with Developers - Maria Apazoglou - Making AI understandable for all in production](https://www.wearedevelopers.com/magazine/475-coffee-with-developers-maria-apazoglou-making-ai-understandable-for-all-in-production)