Software Engineering SMTS
- Discuss this with your agent
- Open in Claude
- Open in ChatGPT
Prepare application
- Draft this with your agent
- Open in Claude
- Open in ChatGPT
Role details
Tech stack
+6 more
Job description
- Design, build, and maintain cross-region continuity workflows and orchestration systems that automate site-switch and disaster recovery (DR) activation sequences for large-scale, mission-critical infrastructure.
- Build validation frameworks and observability tooling (dashboards, telemetry, readiness signals) that answer “Are we ready right now?” - not just “Were we ready during the last review?”
- Partner with service-owning teams (SDB Engine, LCM, Archival, SCRT2, HBase, VegaCache) to define cross-system continuity contracts, integration boundaries, and readiness standards.
- Identify and retire manual steps through principled automation - documenting why each step exists, what can go wrong, and the signal that proves success.
- Build software with operational excellence in mind - including clear runbooks, alerting, observability hooks, and deployment patterns that Site Reliability Engineering (SRE) and DevOps teams can own and operate with confidence.
Requirements
We are looking for a Software Engineer who is passionate about distributed systems, automation-first thinking, and making complex operational outcomes reliable and repeatable. You will join a team whose goal is workflow-as-code and continuous readiness validation - not manual run books., * 5+ years of software engineering experience in distributed systems, infrastructure automation, or platform reliability engineering.
- Proven experience building workflow orchestration, automation pipelines, or control-plane systems at scale (e.g., site-switch automation, DR tooling, deployment orchestration).
- Strong proficiency in at least one systems or backend language (Java, Go, Python, or equivalent) with experience building production-quality services.
- Experience with observability tooling, telemetry design, and building operational dashboards that drive engineering decisions.
Preferred Qualifications
- Familiarity with SDB, HBase, ZooKeeper, BookKeeper, or comparable distributed storage systems.
- Experience with chaos engineering, fault injection, or large-scale resilience testing programs.
- Prior work contributing to RTO/RPO programs, disaster recovery automation, or business continuity engineering for Tier-1 production services.
- Understanding of multi-region cell architectures and cross-region dependency management.
Benefits & conditions
In the United States, compensation offered will be determined by factors such as location, job level, job-related knowledge, skills, and experience. Certain roles may be eligible for incentive compensation, equity, and benefits. Salesforce offers a variety of benefits to help you live well including: time off programs, medical, dental, vision, mental health support, paid parental leave, life and disability insurance, 401(k), and an employee stock purchasing program. More details about company benefits can be found at the following link: https://www.salesforcebenefits.com.Pursuant to the San Francisco Fair Chance Ordinance and the Los Angeles Fair Chance Initiative for Hiring, Salesforce will consider for employment qualified applicants with arrest and conviction records. At Salesforce, we believe in equitable compensation practices that reflect the dynamic nature of labor markets across various regions. The typical base salary range for this position is, $148,500 -
About the company
Salesforce is the #1 AI CRM, where humans with agents drive customer success together. Here, ambition meets action. Tech meets trust. And innovation isn’t a buzzword - it’s a way of life. The world of work as we know it is changing and we’re looking for Trailblazers who are passionate about bettering business and the world through AI, driving innovation, and keeping Salesforce’s core values at the heart of it all.
Ready to level-up your career at the company leading workforce transformation in the agentic era? You’re in the right place! Agentforce is the future of AI, and you are the future of Salesforce.
The Software Engineer, Member of Technical Staff (SMTS) role is part of the Advanced Cross-Region Continuity (ACRC) team within Salesforce’s Core Infrastructure Engineering organization. ACRC is the engineering team responsible for building, automating, validating, and continuously hardening the workflows that ensure Salesforce can recover critical services across regions with defined RTO/RPO behavior.
Apply for this position
This job is hosted externally. Click below to view the full posting and apply.
Prepare application
- Draft this with your agent
- Open in Claude
- Open in ChatGPT
Good distractions
Talks and stories from around this role — technically off-topic, practically not.
Moments
Explore playlistsVideos
See allRelated articles
See all
How We Built a Worry-Free System That Runs for 10+ Years – And What We’d Do Again
Now is the time for industrialized software development
Is Software Engineering Over-Saturated?
The 8 Best Code Testing Tools