> Markdown version of [/jobs/ext/2706446-staff-infrastructure-engineer](https://www.wearedevelopers.com/jobs/ext/2706446-staff-infrastructure-engineer). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Staff Infrastructure Engineer - **Company:** Ai, Inc - **Location:** San Francisco, CA, United States - **Experience:** Expert - **Salary:** $307,000.0 - **Contract:** Permanent contract - **Skills:** Artificial Intelligence, Amazon Web Services, Amazon S3, Cloud Computing, Code Review, Continuous Integration, Data Infrastructure, Identity and Access Management, Python (Programming Language), Ruby, Software Engineering, TypeScript, Pulumi, Amazon Virtual Private Cloud (VPC), Amazon Relational Database Service, Kubernetes, AWS Fargate, Functional Programming, Software Coding, Golang - **Published:** September 4, 2026 - **Apply:** https://startup.jobs/staff-infrastructure-engineer-kikoff-company-8998567 ## About the Role * 7+ years of infrastructure, platform, or software engineering experience, or an equivalent record of Staff-level technical impact. * Sustained ownership of a consequential production system. You have operated what you built, handled incidents and unfamiliar failure modes, and improved the system afterward. * Strong coding skills in TypeScript, Python, Go, Ruby, or a similar language. You can build production software, not only configure tools. * Production experience with AWS, infrastructure as code, containers, CI/CD, observability, and the security boundaries around them. * Experienced generalist judgment. You have depth in at least one infrastructure domain and can learn quickly across adjacent ones. * A track record of building platform capabilities that engineers adopted because they solved real problems and were easier to use than one-off alternatives. * Clear, direct communication. You can explain a technical decision, its evidence, and its tradeoffs to engineers and non-technical partners., * Deep experience in one or more of observability, developer productivity, cloud compute, networking, storage, data infrastructure, or security. * Experience with Pulumi and TypeScript on AWS, including ECS or Fargate, k8s, RDS, S3, Lambda, VPC, and IAM. * Experience building and operating an early-stage platform through rapid growth in fintech or another regulated environment. ## Description Kikoff's infrastructure team builds the systems that enable engineering teams to move quickly without sacrificing reliability, security, or cost discipline. The team owns five connected areas: Observability, Developer Productivity, Compute Infrastructure, Networking and Storage, and Data Infrastructure. As a Staff Infrastructure Engineer, you will own high-leverage infrastructure problems from design through production operation and help run infrastructure as a product for Kikoff engineers. You will build internal products and paved paths that turn ambiguous problems into durable systems. Your work should improve delivery speed and reliability, reduce cost, and reduce the amount of operational work product teams carry. This is a hands-on Staff IC role. You will foster relationships, write code, review designs, make architecture decisions, lead through incidents, and help engineers solve problems outside the runbook. You will have a primary area of depth and enough range to follow production problems across infrastructure boundaries., * Design and implement self-service infrastructure on AWS using reusable code and infrastructure-as-code patterns. We use Pulumi, HCL, and TypeScript heavily. The company runs on Ruby. * Own the systems you build in production, including reliability, security, capacity, cost, upgrades, incidents, and recovery. * Give critical services an SLO, actionable alerts, a useful dashboard, a runbook, and a tested recovery path. * Automate recurring operational work and eliminate failure-prone manual steps rather than allowing them to become permanent processes. Run Infrastructure as a Product * Work directly with engineers to turn recurring friction into paved paths, self-service tools, and automated workflows that are faster and safer than one-off solutions. * Measure outcomes for internal customers through adoption, developer feedback, delivery speed, reliability, cost, toil, and on-call load. * Use the fastest responsible path when a team is blocked, then turn recurring friction into automation or a durable platform capability. Set Technical Direction * Set technical direction for ambiguous infrastructure work, then carry it from problem framing and design through implementation, rollout, and production ownership. * Make clear trade-offs among delivery speed, reliability, security, cost, and long-term operational complexity. * Partner across Infrastructure, Security, Data, and Product Engineering to build security and compliance controls into normal engineering workflows. Raise the Engineering Bar * Review code and designs, challenge weak assumptions, help other engineers make better technical decisions, and uphold high standards for reliability, testing, safe deployments, security, and maintainability. * Lead through incidents and unfamiliar failure modes, and ensure the system is better after service is restored. * Create standards, reusable patterns, and durable documentation that improve how teams build, deploy, observe, and operate software. ## Related Videos - [Why segmenting your infrastructure into tiers makes your infrastructure design better](https://www.wearedevelopers.com/videos/1960-why-segmenting-your-infrastructure-into-tiers-makes-your-infrastructure-design-better) - [Shipping Faster with Less: Render on Cloud Hosting, AI Workloads, and the Future of DevOps](https://www.wearedevelopers.com/videos/1894-shipping-faster-with-less-render-on-cloud-hosting-ai-workloads-and-the-future-of-devops) - [Go with the Flow: Stop the Leaks Before Your Memory's a Waterfall!](https://www.wearedevelopers.com/videos/100073-go-with-the-flow-stop-the-leaks-before-your-memory-s-a-waterfall) - [Coffee with Developers: David Heinemeier Hansson](https://www.wearedevelopers.com/videos/875-coffee-with-developers-david-heinemeier-hansson) - [Unleashing Potential Across Teams: The Power of Infrastructure as Code](https://www.wearedevelopers.com/videos/930-unleashing-potential-across-teams-the-power-of-infrastructure-as-code) - [Retooling and refactoring - an investment in people.](https://www.wearedevelopers.com/videos/371-retooling-and-refactoring-an-investment-in-people) ## Related Articles - [Navigating the AI Shift](https://www.wearedevelopers.com/magazine/629-navigating-the-ai-shift) - [How to Become an AI Engineer](https://www.wearedevelopers.com/magazine/331-how-to-become-an-ai-engineer) - [What is Software Engineering in the Age of AI?](https://www.wearedevelopers.com/magazine/640-what-is-software-engineering-in-the-age-of-ai) - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [How We Built a Worry-Free System That Runs for 10+ Years – And What We’d Do Again](https://www.wearedevelopers.com/magazine/751-how-we-built-a-worry-free-system-that-runs-for-10-years-and-what-we-d-do-again) - [Trustworthy AI Starts at Deployment: 5 Checks Before You Ship](https://www.wearedevelopers.com/magazine/753-trustworthy-ai-starts-at-deployment-5-checks-before-you-ship)