> Markdown version of [/jobs/ext/2193318-infrastructure-engineer-database](https://www.wearedevelopers.com/jobs/ext/2193318-infrastructure-engineer-database). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Infrastructure Engineer, Database - **Company:** Langchain, Inc - **Location:** San Francisco, CA, United States - **Experience:** Expert - **Salary:** $180,000.0 - $230,000.0 - **Contract:** Permanent contract - **Skills:** Artificial Intelligence, Amazon Web Services, Microsoft Azure, Backup Devices, Cloud Computing, Cloud Storage, Databases, Continuous Integration, Data Systems, Disaster Recovery, Distributed Data Store, Fault Tolerance, Python (Programming Language), PostgreSQL, Operational Databases, Redis, Reliability Engineering, Runbook, Database Engines, System Programming, Pulumi, Google Cloud, Delivery Pipeline, Kubernetes, Low Latency, Vertica, Terraform - **Published:** August 23, 2026 - **Apply:** https://www.dice.com/job-detail/c3ff8f2c-ff39-46b1-8496-a3ee26a09443 ## About the Role * 5+ years of experience in infrastructure, platform engineering, or SRE with hands-on * Strong hands-on experience with Kubernetes and cloud infrastructure (AWS/Google Cloud Platform/Azure) * Solid scripting/systems programming ability (Go, Python, or similar); * Experience with infrastructure-as-code and CI/CD tooling (Terraform, Helm, ArgoCD, or similar) * Deep familiarity with at least one major cloud provider (AWS, Google Cloud Platform, or Azure) and the primitives used to run stateful workloads reliably - persistent volumes, managed node groups, cloud storage, etc. * Infrastructure-as-code fluency - you write Terraform (or Pulumi/CDK) as your primary language, not an afterthought * Strong operational instincts - you've been on-call for high-traffic data systems, you know how to triage under pressure, and you write runbooks that actually get used * Experience with container orchestration (Kubernetes) and deploying stateful workloads in production * A bias for automation - if you've done something manual twice, you're already thinking about how to make it never happen again * Strong written and oral communication skills, with the ability to translate infrastructure health into language product and business stakeholders understand * The DNA to thrive in a fast-moving, high-autonomy environment - you see gaps as opportunities and own them end to end Nice to Have * Ownership of production database systems (Postgres, ClickHouse, Redis, or similar) * Comfort reading and reasoning about Rust is a plus, as it's the language our database is written in * Understanding of database reliability concepts - replication, backups, point-in-time recovery, connection pooling, and graceful degradation under load ## Description We're building a database specifically designed for AI observability and evaluation, and we need someone to own the infrastructure layer that keeps it running reliably at scale. As a Database Infra Engineer on the SmithDB team, you won't be designing the storage engine - you'll be making sure the engine never goes down, scales seamlessly as our customer base grows, and is operationally excellent across cloud environments., * Own the deployment and operations of SmithDB across cloud environments - including cluster lifecycle management, blue/green and rolling upgrades, and automated failover * Build and maintain the infrastructure tooling (Terraform, Kubernetes, Helm, or equivalent) that provisions, configures, and scales SmithDB nodes * Own the Kubernetes infrastructure that runs our distributed database services (multi-tenant, high throughput, low latency) * Build and improve deployment pipelines, rollout strategies, and infrastructure-as-code for the storage layer * Drive reliability engineering efforts: incident response, postmortems, SLOs, and disaster recovery for a system operating at massive scale * Manage capacity planning and cost efficiency - model growth, rightsize resources, and ensure SmithDB can absorb traffic spikes from our largest customers without manual intervention * Build the CI/CD pipeline for database infrastructure changes - safe, tested, and fast promotion from dev through staging to production * Collaborate closely with SmithDB internals engineers to translate new engine features into production-ready infrastructure changes and ensure safe, low-risk rollouts ## Related Videos - [How building an industry DBMS differs from building a research one](https://www.wearedevelopers.com/videos/768-how-building-an-industry-dbms-differs-from-building-a-research-one) - [Why segmenting your infrastructure into tiers makes your infrastructure design better](https://www.wearedevelopers.com/videos/1960-why-segmenting-your-infrastructure-into-tiers-makes-your-infrastructure-design-better) - [Reducing LLM Calls with Vector Search Patterns - Raphael De Lio (Redis)](https://www.wearedevelopers.com/videos/1714-reducing-llm-calls-with-vector-search-patterns-raphael-de-lio-redis) - [Unleashing Potential Across Teams: The Power of Infrastructure as Code](https://www.wearedevelopers.com/videos/930-unleashing-potential-across-teams-the-power-of-infrastructure-as-code) - [Shipping Faster with Less: Render on Cloud Hosting, AI Workloads, and the Future of DevOps](https://www.wearedevelopers.com/videos/1894-shipping-faster-with-less-render-on-cloud-hosting-ai-workloads-and-the-future-of-devops) - [Accelerating Authentication Architecture: Taking Passwordless to the Next Level](https://www.wearedevelopers.com/videos/733-accelerating-authentication-architecture-taking-passwordless-to-the-next-level) ## Related Articles - [How We Built a Worry-Free System That Runs for 10+ Years – And What We’d Do Again](https://www.wearedevelopers.com/magazine/751-how-we-built-a-worry-free-system-that-runs-for-10-years-and-what-we-d-do-again) - [Dev Digest 120 - Apple and peers](https://www.wearedevelopers.com/magazine/455-dev-digest-120-apple-and-peers) - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [Building AI Solutions with Rust and Docker](https://www.wearedevelopers.com/magazine/494-building-ai-solutions-with-rust-and-docker) - [Dev Digest 121 - AI goes offline](https://www.wearedevelopers.com/magazine/456-dev-digest-121-ai-goes-offline) - [Why Upskilling And Reskilling is Important For Developers](https://www.wearedevelopers.com/magazine/428-why-upskilling-and-reskilling-is-important-for-developers)