> Markdown version of [/jobs/ext/3002057-senior-cloud-platform-engineer-smts](https://www.wearedevelopers.com/jobs/ext/3002057-senior-cloud-platform-engineer-smts). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Senior Cloud Platform Engineer (SMTS) - **Company:** Salesforce.com, Inc. - **Location:** Bellevue, WA, United States - **Experience:** Expert - **Salary:** $148,500.0 - $246,000.0 - **Contract:** Permanent contract - **Skills:** Java (Programming Language), Application Programming Interfaces (APIs), Artificial Intelligence, Amazon Web Services, Cloud Computing, Computer Programming, DevOps, Distributed Systems, Elasticsearch, Monitoring of Systems, Python (Programming Language), Enterprise Messaging Systems, Performance Tuning, Salesforce.Com, Cloud Platform System, Grafana, Kubernetes Helm Charts, Caching, Infrastructure as Code (IaC), Kubernetes, Apache Kafka, Virtual Agents, Terraform, Code Restructuring, Service Stack - **Published:** September 19, 2026 - **Apply:** https://jobs.localjobnetwork.com/apply/add/88389416/1 ## About the Role 5+ years Proven track record in Distributed systems, API platforms, Infrastructure Engineering, Observability or DevOps at scale. Proficiency with Kubernetes (K8s) and Terraform. Hands-on experience managing infrastructure in AWS and/or GCP. Proficiency in programming languages(eg: java, python etc) Experience managing or extending monitoring tools (e.g., Grafana), messaging systems (kafka etc), elastic search, caching frameworks Security First: Understanding of authN/authZ security protocols, particularly in managing isolated or restricted network environments. AI-assisted development fluency: demonstrated use of AI coding assistants (e.g., Claude Code) as part of a daily engineering workflow, able to prompt effectively, critically evaluate generated code, and integrate AI into IaC, testing, and automation pipelines. ## Description As a Senior Member of Technical Staff (SMTS) within our Monitoring Cloud team, you will be a key owner and operator of the systems that keep Salesforce reliable. You won't just be "using" tools; you will be productizing infrastructure to ensure our monitoring capabilities evolve at the scale of our multi-cloud footprint. Your mission is to bridge the gap between high-level feature design and deep-system stability. From automating the "paved path" across AWS and GCP to securing air-gap environments for our most sensitive customers, you will ensure our monitoring stack is invisible, resilient, and intelligent. This is an AI-first engineering role. You will use AI-assisted development tools (e.g., Claude Code) as the default for every inner-loop activity, code authoring, Terraform and Kubernetes scaffolding, test generation, refactoring, log/trace analysis, runbook drafting, and documentation. We expect AI to compound your throughput on routine implementation so you can focus your human judgment on architecture, security, on-call response, and customer outcomes. Core Responsibilities 1. Infrastructure as Code (IaC) & Automation Design and implement automation frameworks using Terraform and Kubernetes to manage monitoring infrastructure. Standardize "paved path" deployments across AWS and GCP, eliminating manual configuration errors and ensuring global consistency. Use AI-assisted tooling as the default for authoring, refactoring, and reviewing IaC modules, Helm charts, and automation scripts while directing intent, validating output, and owning the final result. 2. Infrastructure Upkeep & Productization Own the lifecycle of the Monitoring Cloud stack, including version upgrades and performance tuning. Productize core components (e.g., Grafana, custom Terraform providers) to make them consumable as reliable services by internal engineering teams. Leverage AI for upgrade planning, release-note analysis, migration scaffolding, and boilerplate-heavy productization work (API wiring, schema plumbing, SDK generation), while retaining accountability for design and rollout. 3. Secure & Air-Gapped Operations Deploy and manage the full monitoring stack within highly isolated, air-gapped environments. Ensure that our most secure customer segments receive the same level of observability and reliability as our public cloud offerings. Apply AI assistance during development of the artifacts that ship into these environments; operate them in-network with the disciplined, human-driven workflows these environments require. 4. Operational Excellence & Health Participate in the team's on-call rotation, providing the deep technical expertise required to maintain strict SLAs and availability targets. Conduct root-cause analysis (RCA) for complex system failures and implement long-term preventative fixes. Address support requests with a "customer first" mindset Use AI as a co-pilot during incident response and RCA: summarizing logs, correlating traces, proposing hypotheses, and drafting status updates and postmortem while the engineer remains the accountable responder and decision-maker. 5. Next-Gen Feature Delivery Design and deliver platform features that adhere to enterprise standards while pioneering AI-driven development practices to accelerate delivery and enhance system intelligence. Contribute to and evolve the team's AI-assisted development playbook: prompts, agents, skills, evaluation harnesses, and guardrails that let the team ship faster without sacrificing quality or security. ## Related Videos - [The Open-source Java SDK for Multi-Cloud Development - Sandeep Pal](https://www.wearedevelopers.com/videos/2113-the-open-source-java-sdk-for-multi-cloud-development-sandeep-pal) - [5 steps for running a Kubernetes environment at scale](https://www.wearedevelopers.com/videos/88-5-steps-for-running-a-kubernetes-environment-at-scale) - [HTTP headers that make your website go faster](https://www.wearedevelopers.com/videos/1676-http-headers-that-make-your-website-go-faster) - [From DevOps to Scaled DevOps: How We’re Rebuilding Continuous Delivery as a Platform](https://www.wearedevelopers.com/videos/100018-from-devops-to-scaled-devops-how-we-re-rebuilding-continuous-delivery-as-a-platform) - [#90DaysOfDevOps - The DevOps Learning Journey](https://www.wearedevelopers.com/videos/548-90daysofdevops-the-devops-learning-journey) - [Shipping Faster with Less: Render on Cloud Hosting, AI Workloads, and the Future of DevOps](https://www.wearedevelopers.com/videos/1894-shipping-faster-with-less-render-on-cloud-hosting-ai-workloads-and-the-future-of-devops) ## Related Articles - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [7 Cloud Computing Trends Coming in 2025 for Developers](https://www.wearedevelopers.com/magazine/412-7-cloud-computing-trends-coming-in-2025-for-developers) - [What Are The Top Skills Required For Azure Developers?](https://www.wearedevelopers.com/magazine/77-what-are-the-top-skills-required-for-azure-developers) - [Best US AI Conferences for CTOs in 2026: Build vs. Buy, Vendor Evaluation, and Peer Intelligence](https://www.wearedevelopers.com/magazine/736-best-us-ai-conferences-for-ctos-in-2026-build-vs-buy-vendor-evaluation-and-peer-intelligence) - [How We Built a Worry-Free System That Runs for 10+ Years – And What We’d Do Again](https://www.wearedevelopers.com/magazine/751-how-we-built-a-worry-free-system-that-runs-for-10-years-and-what-we-d-do-again) - [Why Upskilling And Reskilling is Important For Developers](https://www.wearedevelopers.com/magazine/428-why-upskilling-and-reskilling-is-important-for-developers)