> Markdown version of [/jobs/ext/2617156-senior-software-development-engineer-aws-mantle](https://www.wearedevelopers.com/jobs/ext/2617156-senior-software-development-engineer-aws-mantle). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Senior Software Development Engineer, AWS Mantle - **Company:** Amazon.com, Inc. - **Location:** New York, NY, United States - **Experience:** Expert - **Salary:** $168,100.0 - $227,400.0 - **Contract:** Internship / Graduate position - **Skills:** Microsoft Access, Java (Programming Language), Application Programming Interfaces (APIs), Artificial Intelligence, Amazon Web Services, C++ (Programming Language), Code Review, Computer Programming, Software Design Patterns, Distributed Systems, Python (Programming Language), Machine Learning, Zero Trust Network Access, Software Engineering, AI Infrastructure, Information Technology, Low Latency, Build Process, Machine Learning Operations, TensorRT, Software Coding, Software Version Control, Golang, Programming Languages - **Published:** August 31, 2026 - **Apply:** https://dejobs.org/x/x/D0924B718B0E49A88F1BAF28BF02ACA0/job/ ## About the Role * 5+ years of non-internship professional software development experience * 5+ years of programming with at least one software programming language experience * 5+ years of leading design or architecture (design patterns, reliability and scaling) of new and existing systems experience * Experience as a mentor, tech lead or leading an engineering team * + 5+ years of non-internship professional software development experience * + 5+ years of leading design or architecture (design patterns, reliability and scaling) of new and existing systems experience * + Bachelor's degree in Computer Science, Engineering, a related field, or equivalent experience * + 5+ years of programming experience with at least one modern language such as Java, C++, Python, Go, or Rust * + Experience driving cross-organizational technical strategy and delivering results in complex, ambiguous environments where the business problem and technical approach are not pre-defined, * 5+ years of full software development life cycle, including coding standards, code reviews, source control management, build processes, testing, and operations experience * Bachelor's degree in computer science or equivalent * + Master's degree or equivalent in computer science, machine learning, engineering, or related fields, or PhD * + Experience building large-scale machine learning and AI solutions at Internet scale * + Familiarity with inference frameworks such as vLLM, TensorRT, or Triton Inference Server ## Description Are you passionate about building the infrastructure that powers the next generation of AI? We are seeking a Sr. Software Development Engineer to join the AWS Mantle team and drive the technical vision for our distributed inference engine that serves millions of customers across Amazon Bedrock. In this role, you will define and execute on large-scale, ambiguous technical challenges at the intersection of machine learning systems, distributed computing, and security-shaping how the world accesses foundation models. Set the long-term technical direction for a globally distributed, high-performance ML inference platform serving models from industry-leading AI providers Own end-to-end system design decisions that directly impact latency, reliability, and scalability for millions of customers worldwide Influence engineering strategy across Amazon Bedrock, partnering with senior leadership to align technical investments with business outcomes Raise the engineering bar through exemplary system design, mentorship, and contributions to the broader AWS engineering community Navigate complex trade-offs across performance, security, and cost while maintaining the highest standards for operational excellence, As a Sr. SDE on the Mantle team, you will serve as the technical conscience and strategic thought leader for one of AWS's most critical AI infrastructure platforms. You will architect solutions that are reliable, scalable, and secure-operating at the cutting edge of distributed systems where millisecond-level latency and zero-trust security are non-negotiable. Design and evolve the architecture of Mantle's distributed inference engine, including capacity management, model onboarding pipelines, and quality-of-service controls Drive cross-organizational initiatives spanning multiple AWS teams to deliver seamless, OpenAI-compatible API experiences with Zero Operator Access (ZOA) security guarantees Lead technical strategy for scaling inference to support rapid onboarding of new foundation models while maintaining global availability and performance SLAs About the team The AWS Mantle team is building the next-generation inference engine that powers Amazon Bedrock-providing secure, enterprise-grade access to high-performing foundation models from the world's leading AI companies. Our mission is to simplify and accelerate how models are served at global scale, with an unwavering commitment to customer trust through innovations like our Zero Operator Access architecture, designed so that no person-whether from AWS, a customer, or a model provider-can ever access customer inference data. We operate at massive scale, serving inference requests across all major AWS regions with sophisticated automated capacity management and unified resource pools Our team values builders who thrive in ambiguity, think long-term, and are excited to define the future of AI infrastructure from the ground up We foster a collaborative, inclusive environment where diverse perspectives drive better solutions-and where the best ideas win regardless of where they originate We ship fast and iterate with purpose, having rapidly expanded from launch to supporting models from OpenAI, DeepSeek, Google, Mistral, NVIDIA, and more We believe work should be meaningful and fun-you'll join a team that takes pride in making history at the forefront of generative AI ## Related Videos - [Shipping Faster with Less: Render on Cloud Hosting, AI Workloads, and the Future of DevOps](https://www.wearedevelopers.com/videos/1894-shipping-faster-with-less-render-on-cloud-hosting-ai-workloads-and-the-future-of-devops) - [Tour de Force: Open-Source LLM Inference Optimization from Simple to Sophisticated](https://www.wearedevelopers.com/videos/100099-tour-de-force-open-source-llm-inference-optimization-from-simple-to-sophisticated) - [Go with the Flow: Stop the Leaks Before Your Memory's a Waterfall!](https://www.wearedevelopers.com/videos/100073-go-with-the-flow-stop-the-leaks-before-your-memory-s-a-waterfall) - [Are Code Reviews Worth It? Insights from 16 Years of Review Data](https://www.wearedevelopers.com/videos/1135-are-code-reviews-worth-it-insights-from-16-years-of-review-data) - [Fireside Chat - In conversation with Werner Vogels, CTO of Amazon.com](https://www.wearedevelopers.com/videos/100265-fireside-chat-in-conversation-with-werner-vogels-cto-of-amazon-com) - [Efficient deployment and inference of GPU-accelerated LLMs​](https://www.wearedevelopers.com/videos/929-efficient-deployment-and-inference-of-gpu-accelerated-llms) ## Related Articles - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [MLOps And AI Driven Development](https://www.wearedevelopers.com/magazine/82-mlops-and-ai-driven-development) - [What is Software Engineering in the Age of AI?](https://www.wearedevelopers.com/magazine/640-what-is-software-engineering-in-the-age-of-ai) - [Dev Digest 162: AI careers, MCP, AWS best practices & floppy sweaters](https://www.wearedevelopers.com/magazine/571-dev-digest-162-ai-careers-mcp-aws-best-practices-floppy-sweaters) - [Navigating the AI Shift](https://www.wearedevelopers.com/magazine/629-navigating-the-ai-shift) - [Dev Digest 121 - AI goes offline](https://www.wearedevelopers.com/magazine/456-dev-digest-121-ai-goes-offline)