> Markdown version of [/jobs/ext/2715265-cloud-infrastructure-engineer](https://www.wearedevelopers.com/jobs/ext/2715265-cloud-infrastructure-engineer). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Cloud Infrastructure Engineer - **Company:** McClure Engineering Co. - **Location:** San Francisco, CA, United States - **Experience:** Expert - **Contract:** Permanent contract - **Skills:** Artificial Intelligence, Amazon Web Services, Cloud Computing, Cursor (Graphical User Interface Elements), Programming Tools, Distributed Systems, Domain Name System (DNS), Network Architecture, Reliability Engineering, Blockchain, Prometheus, Istio, Grafana, Multi-Cloud, Amazon Virtual Private Cloud (VPC), Kubernetes, Cloudflare, Terraform - **Published:** September 4, 2026 - **Apply:** https://startup.jobs/cloud-infrastructure-engineer-alchemy-com-8288726 ## About the Role * 5+ years as an Infrastructure Engineer focused on reliability (SRE, Production Engineer, Platform Engineer). * Experience driving company-wide reliability efforts, including SLO frameworks and error budget policies. * Strong proficiency with observability stacks: OpenTelemetry, Prometheus/Grafana. * Deep experience with cloud infrastructure (AWS/GCP), Kubernetes, and multi-region architectures. * Skilled with Terraform, Helm, and GitOps workflows (e.g., ArgoCD) with an automation-first mindset. * Experience leveraging agentic development tools (Claude Code, Cursor, Codex) and workflow automation (n8n) to accelerate IaC and build internal tooling is a strong plus. * Solid networking fundamentals - VPC design, DNS, IPAM, security groups, cross-cloud connectivity, and service mesh (e.g., Istio) experience is a plus. * Strong cross-functional communicator across SRE, security, and product engineering. * Blockchain infrastructure, distributed systems, or high-throughput RPC experience - not required but a plus. ## Description As an engineer in the Infrastructure department at Alchemy, you will design, deploy, and continuously improve the infrastructure powering our blockchain developer platform - serving 100+ chains, billions of daily requests, and over $150B in annual transactions. The Infrastructure team provides the infrastructure, tooling, and expertise needed to allow Alchemy engineers to ship, scale, and operate high-quality products in a fast, safe, and cost-efficient manner. What You'll Do * Architect and operate scalable, self-healing infrastructure leveraging Kubernetes, Terraform, and cloud-native tools across multi-region deployments. * Drive AI enablement across engineering - ensuring repos, tooling, and workflows are optimized for agentic development with tools like Claude Code, Cursor, and Codex. * Build AI-powered infrastructure tooling and automation (e.g., automated K8s upgrades, IaC plan analysis, cost optimization advisors, MCP servers, n8n workflows). * Build and maintain internal developer platform (IDP) capabilities for self-service deployments, observability, and reliability. * Develop observability frameworks using Prometheus and Grafana for metrics, dashboards, and alerting. * Lead incident management with blameless post-mortems; define and enforce SLIs, SLOs, and error budgets across services. * Design and manage multi-cloud, multi-region network architecture - VPC design, IPAM, DNS (Cloudflare), cross-cloud connectivity, security groups, and edge-proxy/istio gateway configuration. * Collaborate with security teams to embed compliance into infrastructure, including IaC scanning and runtime protection. * Provide technical leadership and mentorship to elevate the team's operational capabilities. ## Related Videos - [Infrastructure as Code: The Developer's Secret Weapon](https://www.wearedevelopers.com/videos/1221-infrastructure-as-code-the-developer-s-secret-weapon) - [Shipping Faster with Less: Render on Cloud Hosting, AI Workloads, and the Future of DevOps](https://www.wearedevelopers.com/videos/1894-shipping-faster-with-less-render-on-cloud-hosting-ai-workloads-and-the-future-of-devops) - [5 steps for running a Kubernetes environment at scale](https://www.wearedevelopers.com/videos/88-5-steps-for-running-a-kubernetes-environment-at-scale) - [Rate-limiting using eBPF and Istio: How to protect your SaaS customers from themselves](https://www.wearedevelopers.com/videos/100220-rate-limiting-using-ebpf-and-istio-how-to-protect-your-saas-customers-from-themselves) - [Building a Cloud Platform Where Everything is Just Another Kubernetes Resource](https://www.wearedevelopers.com/videos/100137-building-a-cloud-platform-where-everything-is-just-another-kubernetes-resource) - [Terraform for Developers](https://www.wearedevelopers.com/videos/3-terraform-for-developers) ## Related Articles - [7 Cloud Computing Trends Coming in 2025 for Developers](https://www.wearedevelopers.com/magazine/412-7-cloud-computing-trends-coming-in-2025-for-developers) - [How We Built a Worry-Free System That Runs for 10+ Years – And What We’d Do Again](https://www.wearedevelopers.com/magazine/751-how-we-built-a-worry-free-system-that-runs-for-10-years-and-what-we-d-do-again) - [What Are The Top Skills Required For Azure Developers?](https://www.wearedevelopers.com/magazine/77-what-are-the-top-skills-required-for-azure-developers) - [Dev Digest 120 - Apple and peers](https://www.wearedevelopers.com/magazine/455-dev-digest-120-apple-and-peers) - [Dev Digest 162: AI careers, MCP, AWS best practices & floppy sweaters](https://www.wearedevelopers.com/magazine/571-dev-digest-162-ai-careers-mcp-aws-best-practices-floppy-sweaters) - [Everything a Developer Needs to Know About MCP with Neo4j](https://www.wearedevelopers.com/magazine/604-everything-a-developer-needs-to-know-about-mcp-with-neo4j)