Cloud Infrastructure Engineer

The Collective
San Francisco, CA, United States
21 days ago
Apply on startup.jobs
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
5 years minimum
Working hours
Regular working hours
Job source

Tech stack

Clean Code Principles Artificial Intelligence Amazon Web Services Amazon Elastic Compute Cloud Systems Engineering BigQuery Command-Line Interface Software as a Service Cloud Computing Computer Networks Continuous Integration Software Debugging
+26 more
Linux DevOps Domain Name System (DNS) Github Identity and Access Management Python (Programming Language) Key Management Role-Based Access Control Zero Trust Network Access Software Engineering TCP/IP TypeScript Datadog Data Logging Load Balancing ReactJS Amazon ElastiCache Multi-Cloud Infrastructure as Code (IaC) Amazon Virtual Private Cloud (VPC) AI Platforms Sentry Amplitude Analytics AWS Fargate Codebase Terraform

Job description

You’ll join Collective’s Infrastructure team, a small group that owns the platform every engineer here builds on: AWS and GCP foundations, CI/CD, IaC, observability, secrets management, and identity. We partner with product engineering to unblock shipping, with the security team to keep the platform safe, and with the AI Platform team on the infra behind Collective’s growing AI investments. As a Senior Cloud Infrastructure Engineer on this team, you’ll pave new paths, test recovery plans, and manage quiet, stable systems. You’ll be the person justifying design choices with production experience and the one others turn to when a system needs rethinking.

What you’ll do:

  • Use Infrastructure as Code (IaC) with Terraform to provision, deploy, and manage cloud resources on AWS and GCP as well as other SaaS vendors.
  • Embed security best practices into the infrastructure by enforcing zero-trust architecture principles like least privilege and identity-based access to protect systems and data.
  • Build scalable, reliable, and cost-effective systems that hold up as Collective grows.
  • Develop and test disaster recovery plans.
  • Own the CI/CD system engineering teams ship on. Set standards, drive reliability and speed improvements, and mentor teams on best practices.
  • Reduce operational toil across the platform through automation, leveraging AI tooling where it accelerates safe, high-quality work.
  • Work closely with product engineering teams to understand application needs and translate them into scalable infrastructure solutions.
  • Own the observability stack (monitoring, logging, alerting) and use it to proactively identify and remediate performance bottlenecks.
  • Participate in the on-call rotation to respond to outages, recover systems, own incident response and post-mortem.
  • Stay current with emerging technologies and best practices in Cloud Infrastructure, DevOps, and Platform Engineering., * Cloud: AWS (EC2, IAM, VPC, ECS, Fargate, Lambda, RDS, Elasticache, Opensearch), GCP (BigQuery)
  • Monitoring/Observability: Datadog, Sentry, Amplitude
  • Github, GHA
  • IAC: Terraform, HCP
  • Security tooling
  • Codebase: Python/Typescript/React

Requirements

  • At least 5 years of hands-on experience as a Cloud Infrastructure Engineer, DevOps, or SRE with a proven track record of operating production cloud environments at scale.
  • You operate effectively in ambiguous, fast-changing environments. You can pick up a half-defined problem, define the path forward, and drive it to production without waiting for a playbook.
  • Cloud Platforms: Proficiency in multi-cloud operations. AWS is highly preferred; GCP is a plus.
  • Experience implementing infrastructure and security policy as code.
  • Strong software development skills, preferably in Python or another high level language
  • Strong written and verbal communication skills for driving cross-team alignment. You must be able to clearly and persuasively communicate complex concepts and risks in an engineering-driven environment.
  • Experience mentoring engineers, leading post-incident reviews, or driving cross-team infra initiatives to completion. You’re comfortable being the person other engineers ask when something breaks.
  • Ability to write clean, maintainable code for automation and tooling. Experience building internal tools or services to eliminate manual work is a plus.
  • Familiarity with foundational networking concepts and protocols (TCP/IP, DNS, load balancing, VPCs, firewalls) and their application in cloud and hybrid environments.
  • Strong hands-on skills with Linux and command-line tools; you are comfortable using terminals and utilities to manage and debug systems efficiently.
  • Comfort using AI tooling as leverage for infra automation, tooling, and debugging. Bonus if you’ve built or contributed to AI-assisted DevOps workflows.

Benefits & conditions

  • Commuter Support: $150 monthly reimbursement for transit expenses.
  • Health & Wellness: $200 quarterly reimbursement to support your well-being.
  • Time Off: Flexible PTO plus 14 company holidays.
  • Comprehensive Coverage: 100% medical, dental, and vision for employees; 75% coverage for dependents.
  • Parental Leave: 16 weeks fully paid.
  • Retirement & Ownership: 401k plan plus an equity package.
  • Team Connection: Quarterly virtual events and an annual in-person summit.

About the company

Collective is on a mission to redefine the way businesses-of-one work. Our technology and team of trusted advisors help members achieve financial independence by taking care of everything from business incorporation to accounting, bookkeeping, tax services, and access to a thriving community, all in one integrated platform. We believe in empowering self-employed people to enjoy the same tax savings that big companies get, so they can focus on their passion, not paperwork.

Featured in Forbes, Business Insider, Yahoo, Bloomberg, Financial Times, TechCrunch, and more. We are backed by General Catalyst, Sound Ventures, QED Investors, Google’s Gradient Ventures, Expa, and other investors who have financed iconic companies like YouTube, Substack, Twitch, Box, Hims, Instacart, and Lyft.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on startup.jobs
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

1:34 min

Pivoting careers into specialized platform engineering roles

Xavier Portilla Edo · LIVE

6:36 min

Funding open source through GitHub Accelerator and Sponsors

Stormy Peters · World Congress 2023

2:59 min

Designing HTTP and HTML for familiar document sharing

Tim Berners-Lee · World Congress 2023

52 sec

Running persistent Linux environments directly on Windows

Ben Breard Ben Breard · World Congress 2025

2:40 min

Using GitHub primitives for internal documentation and corporate operations

Kyle Daigle · Coffee With Developers

5:02 min

Mapping distributed compute paradigms to modern vehicles

Joachim Werner · LIVE

Videos

See all

Related articles

See all