SITE RELIABILITY ENGINEER
- Discuss this with your agent
- Open in Claude
- Open in ChatGPT
Prepare application
- Draft this with your agent
- Open in Claude
- Open in ChatGPT
Role details
Tech stack
+1 more
Requirements
Looking for 7-10 years of experience, ideally with a strong technical focus.
Must be proficient in Python and Typescript, with familiarity in Kubernetes, Helm, Terraform, and Terragrunt., Experience driving systematic application improvements across complex, distributed systems with high- availability requirements for a high- growth start- up
Extended period working at high- growth start- ups (no recent Big Tech)
Prior SWE experience and strong coding abilities (Python and TypeScript)
Hard skills
Hands- on experience automating single- tenant infrastructure provisioning and optimising deployment pipelines for application code or machine learning models
IaC and Orchestration tools including Kubernetes, Helm, Terraform, and Terragrunt
Experience with the rest of our tech stack (e. g. , AWS, PostgreSQL, Redis, and Kafka)
Soft skills
Ability to own complex systems and solve complex problems independently
Defines standards for operational excellence, e. g. , automating manual workflows
Miscellaneous
Experience in regulated industries (health tech given HIPAA complications)
Open- source contributions and LinkedIn recommendations for others
Benefits & conditions
Traits to avoid
Bootcamps and no university degrees, recent consulting, or traditional finance
Frequent career moves with multiple short, 1- year stints
Compensation and Logistics
Open to candidates in San Francisco or New York, with occasional travel required.
Timeline and Urgency
Looking to hire one person in Q1 and potentially another in Q2.
Three-stage interview process: team chat, technical interview, and on-site assessment.
Pain Points
Need for automation to reduce manual workload associated with onboarding new customers.
The current team is small, underscoring the need for individuals who can handle diverse tasks.
Ideal Candidate Profile
Prefer candidates with startup experience and a generalist, SRE background.
Background in backend development transitioning to infrastructure is highly valued. Infrastructure Responsibilities
- Infrastructure Ownership: Design, implement, and maintain the production environment, having previously handled 500+ machine deployments.
About the company
Client is solving complex challenges in healthcare through robust and reliable Medical Intelligence. Our AI-driven solutions significantly enhance workflows for hospitals and clinics, enabling patients to access medication faster while substantially boosting healthcare providers’ revenues. We are developing a suite of software solutions in conjunction with major healthcare systems to create fast, effective, and affordable access to a provider for every patient in the United States.
“Based in New York or San Francisco, offering a strong compensation range of $200-300k. The hiring team has a strong track record of successful placements (c.5 in recent months).
Position requires a balance of operational and software engineering skills.
Role involves optimizing database accesses, handling production issues, and improving application performance.
Apply for this position
This job is hosted externally. Click below to view the full posting and apply.
Prepare application
- Draft this with your agent
- Open in Claude
- Open in ChatGPT
Good distractions
Talks and stories from around this role — technically off-topic, practically not.
Moments
Explore playlistsVideos
See allRelated articles
See all
Fully Remote Software Engineer Jobs
Is Software Engineering Over-Saturated?
Find a Developer Job: 12 Best Job Sites For Developers
Highest Paying Tech Companies for Developers