> Markdown version of [/jobs/ext/3581107-director-of-engineering-infrastructure](https://www.wearedevelopers.com/jobs/ext/3581107-director-of-engineering-infrastructure). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Director of Engineering, Infrastructure - **Company:** OTTER, INC. - **Location:** United States - **Experience:** Experienced - **Salary:** $244,000.0 - $313,000.0 - **Contract:** Permanent contract - **Skills:** Artificial Intelligence, Amazon Web Services, Amazon Elastic Compute Cloud, Backup Devices, Cloud Computing, Cluster Analysis, Databases, Continuous Integration, DevOps, Disaster Recovery, Distributed Systems, Django Web Framework, Fault Tolerance, Identity and Access Management, Python (Programming Language), Machine Learning, MySQL, Operational Data Store, Queueing Systems, Reliability Engineering, Software Engineering, WebSocket, Load Balancing, Cloud Platform System, Autoscaling, Fastapi, Kubernetes, Information Technology, Model Inference, Terraform - **Published:** October 4, 2026 - **Apply:** https://startup.jobs/engineering-director-infrastructure-otterai-10269259 ## About the Role * Bachelor's degree or higher in Computer Science or a related technical field. * 10+ years of software development or engineering experience, including 4+ years of engineering management experience. * Significant experience leading infrastructure, DevOps, SRE, or platform teams responsible for production systems at scale. * Deep experience designing and operating highly available, fault-tolerant distributed systems. * Strong knowledge of AWS and cloud-native infrastructure, including Kubernetes/EKS, EC2, Auto Scaling Groups, load balancing, networking, storage, IAM, monitoring, and infrastructure-as-code technologies such as Terraform. * Deep understanding of MySQL operations, including performance, replication, clustering, backup, and disaster recovery. * Strong knowledge of asynchronous architectures, including message queues, worker systems, retries, idempotency, and failure recovery. * Experience building strong production reliability practices, including observability, incident management, on-call operations, disaster recovery, multi-region architectures, and zero-downtime migrations. * A track record of materially improving cloud efficiency and infrastructure unit economics without compromising reliability or performance. * Experience operating latency-sensitive, high-throughput AI inference workloads in production; GPU inference optimization experience is a plus. * Experience scaling infrastructure to support rapid product and customer growth, including both self-serve products and enterprise customers with demanding security, reliability, and compliance requirements. * Ability to move effectively between architecture, operational data, financial analysis, organizational design, and executive communication, with strong cross-functional influence. * Ability to use AI-assisted engineering and operational tools to improve team productivity, incident response, diagnostics, and infrastructure management. * Familiarity with relevant application technologies such as Python, Django, FastAPI, and WebSockets. ## Description Otter.ai is seeking a Director of Engineering, Infrastructure to lead the teams that power our products at scale. This leader will own our cloud infrastructure, CI/CD, DevOps, site reliability practices, and production operations. This is a hands-on, high-impact leadership role for someone who can set infrastructure strategy while going deep on critical technical and operational challenges. You will partner closely with Product, Product Engineering, AI/ML, Security, and Finance to deliver a platform that is reliable, scalable, secure, developer-friendly, and cost-efficient. Your Impact * Own the architecture and operation of cloud, compute, networking, database, production, and AI inference infrastructure. * Translate business and product plans into infrastructure roadmaps, capacity requirements, staffing plans, and investment decisions. * Ensure production infrastructure is reliable, scalable, secure, performant, and cost-efficient. * Establish infrastructure metrics, engineering standards, observability, and clear accountability, proactively surfacing risks and opportunities. * Improve developer velocity through automation, CI/CD, and self-service tooling. * Strengthen incident response, disaster recovery, and business continuity practices. * Partner with AI/ML teams to optimize inference performance, reliability, resource utilization, and cost. * Embed access control, compliance, and risk management into infrastructure architecture and operations. * Recruit, develop, and lead a high-performing infrastructure engineering organization as a technically credible player-coach. * Communicate infrastructure priorities, risks, investments, and tradeoffs clearly to the executive team, and partner with Engineering and Finance on cost and capacity decisions. ## Related Videos - [Shipping Faster with Less: Render on Cloud Hosting, AI Workloads, and the Future of DevOps](https://www.wearedevelopers.com/videos/1894-shipping-faster-with-less-render-on-cloud-hosting-ai-workloads-and-the-future-of-devops) - [Build your backend using FastAPI](https://www.wearedevelopers.com/videos/506-build-your-backend-using-fastapi) - [Shifting Stress to Progress— Understanding DevOps to do DevOps Better](https://www.wearedevelopers.com/videos/268-shifting-stress-to-progress-understanding-devops-to-do-devops-better) - [MySQL Protocol Features You Should Be Aware Of](https://www.wearedevelopers.com/videos/100267-mysql-protocol-features-you-should-be-aware-of) - [Remote Driving on Plant Grounds with State-of-the-Art Cloud Technologies](https://www.wearedevelopers.com/videos/251-remote-driving-on-plant-grounds-with-state-of-the-art-cloud-technologies) - [Intro to FastAPI](https://www.wearedevelopers.com/videos/462-intro-to-fastapi) ## Related Articles - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [How We Built a Worry-Free System That Runs for 10+ Years – And What We’d Do Again](https://www.wearedevelopers.com/magazine/751-how-we-built-a-worry-free-system-that-runs-for-10-years-and-what-we-d-do-again) - [Navigating the AI Shift](https://www.wearedevelopers.com/magazine/629-navigating-the-ai-shift) - [Is Software Engineering Over-Saturated?](https://www.wearedevelopers.com/magazine/418-is-software-engineering-over-saturated) - [How to Become an AI Engineer](https://www.wearedevelopers.com/magazine/331-how-to-become-an-ai-engineer) - [What is Software Engineering in the Age of AI?](https://www.wearedevelopers.com/magazine/640-what-is-software-engineering-in-the-age-of-ai)