> Markdown version of [/jobs/ext/2590353-staff-infrastructure-engineer-technical-staff](https://www.wearedevelopers.com/jobs/ext/2590353-staff-infrastructure-engineer-technical-staff). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Staff Infrastructure Engineer - Technical Staff - **Company:** Savvy Wealth - **Location:** New York, NY, United States - **Experience:** Experienced - **Salary:** $251,000.0 - $272,500.0 - **Contract:** Permanent contract - **Skills:** Java (Programming Language), Amazon Web Services, Continuous Integration, Django Web Framework, Identity and Access Management, PostgreSQL, Network Control, Platform as a Service (PAAS), Redis, Software Engineering, Cloud Platform System, Large Language Models, Kubernetes, Terraform - **Published:** August 25, 2026 - **Apply:** https://www.indeed.com/viewjob?jk=421fe8d561667310 ## About the Role * Experience operating Rails, Sidekiq, or another stateful monolithic application (Django, Java, and similar all count). * Fintech or regulated-industry background: financial services, wealth management, or compliance-constrained infrastructure. * Experience migrating off a managed PaaS onto self-managed infrastructure. ## Description Savvy is hiring our first dedicated infrastructure engineer. Our platform runs on AWS, EKS, and OpenTofu, and we've built it with product engineers who picked up infrastructure work alongside shipping features. That got us here. Going forward we want someone who has run production infrastructure for years, can look hard at the decisions we've made, and set the direction from here. You'll own the cloud platform: cluster architecture, networking, IAM, cost, IaC standards, CI/CD, observability, and the operational side of Postgres, Redis, and our background job tiers. You'll set the deploy contract and the standards that 35+ engineers build within, and you'll be the person who decides what our infrastructure looks like in a year. The other half of the job is enablement. We're running coding agents across the engineering org and building the infrastructure to support them, which means compute, isolation, and observability problems that don't have settled answers yet. You'll work directly with product engineers on the problems in front of them rather than from behind a ticket queue., * Own the platform: Set architecture and technical direction for our AWS and EKS footprint, including cluster topology, networking, IAM boundaries, and cost. * Audit and course-correct: Review what we've built, identify what won't hold, and sequence the remediation against a team that has to keep shipping. * Set the standards: Define our OpenTofu module structure, state management, deploy contract, and the patterns every engineer works within. * Raise our operational bar: Build out SLOs, alerting that means something, incident response, and on-call practice. * Run our stateful systems: Own Postgres, Redis, and job processing in production, including upgrades, failover, and capacity. * Build the paved road: Ship the tooling and internal services that make the platform easy for product engineers to use correctly. * Support agent infrastructure: Help design and run the compute layer for our coding agents and LLM workloads. * Mentor: Grow the engineers on the team who've been doing infrastructure work part-time into stronger operators., * Production Kubernetes ownership: 3+ years as the escalation point for a cluster you owned, including upgrades, control plane, networking, and incidents. Deploying to a cluster someone else runs isn't the same job. * Experience inheriting infrastructure: You've taken over someone else's platform, formed a view on what to keep and what to replace, and executed it without stopping the business. * Small team, real consequences: You've worked on an infrastructure team of five or fewer where mistakes cost money, customers, or compliance standing. You couldn't specialize and you didn't have anyone to escalate to. * You build software: A controller, a CLI, a deploy system, an internal service, something other engineers depended on. Infrastructure work here is engineering, and IaC is one of the tools. * Terraform or OpenTofu depth: Module design, state management, and a real answer for drift. * Stateful systems in production: Postgres specifically, including connection management, upgrades, and failover. * CI/CD and observability as a practice: Not just standing up the tools, but making the signal useful. * Fluent with agentic coding tools: You use Claude Code, Codex, or similar in your own work, and you have opinions about what infrastructure for LLM and agent workloads should look like. ## Related Videos - [Shipping Faster with Less: Render on Cloud Hosting, AI Workloads, and the Future of DevOps](https://www.wearedevelopers.com/videos/1894-shipping-faster-with-less-render-on-cloud-hosting-ai-workloads-and-the-future-of-devops) - [Optimizing Discovery: PostgreSQL's Role in Transforming GetYourGuide's Search](https://www.wearedevelopers.com/videos/1647-optimizing-discovery-postgresql-s-role-in-transforming-getyourguide-s-search) - [Infrastructure as Code: The Developer's Secret Weapon](https://www.wearedevelopers.com/videos/1221-infrastructure-as-code-the-developer-s-secret-weapon) - [Reducing LLM Calls with Vector Search Patterns - Raphael De Lio (Redis)](https://www.wearedevelopers.com/videos/1714-reducing-llm-calls-with-vector-search-patterns-raphael-de-lio-redis) - [Navigating the Corporate Jungle: Life as a Developer in a large Company](https://www.wearedevelopers.com/videos/621-navigating-the-corporate-jungle-life-as-a-developer-in-a-large-company) - [Discover the open source trio you didn’t expect: .NET and PostgreSQL on Linux](https://www.wearedevelopers.com/videos/2042-discover-the-open-source-trio-you-didn-t-expect-net-and-postgresql-on-linux) ## Related Articles - [How We Built a Worry-Free System That Runs for 10+ Years – And What We’d Do Again](https://www.wearedevelopers.com/magazine/751-how-we-built-a-worry-free-system-that-runs-for-10-years-and-what-we-d-do-again) - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [Why Upskilling And Reskilling is Important For Developers](https://www.wearedevelopers.com/magazine/428-why-upskilling-and-reskilling-is-important-for-developers) - [Dev Digest 120 - Apple and peers](https://www.wearedevelopers.com/magazine/455-dev-digest-120-apple-and-peers) - [Is Software Engineering Over-Saturated?](https://www.wearedevelopers.com/magazine/418-is-software-engineering-over-saturated) - [Fully Remote Software Engineer Jobs](https://www.wearedevelopers.com/magazine/447-fully-remote-software-engineer-jobs)