> Markdown version of [/jobs/ext/2278893-slack-proactive-monitoring-engineer](https://www.wearedevelopers.com/jobs/ext/2278893-slack-proactive-monitoring-engineer). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Slack Proactive Monitoring Engineer - **Company:** Salesforce Inc. - **Location:** Indianapolis, IN, United States - **Experience:** Experienced - **Salary:** $75,000.0 - $113,500.0 - **Contract:** Permanent contract - **Skills:** JavaScript (Programming Language), Application Programming Interfaces (APIs), Bash Shell, Software as a Service, Databases, Monitoring of Systems, Python (Programming Language), Key Management, Reliability Engineering, Salesforce.Com, Datadog, Scripting, Grafana, Slack, Performance Monitor, Splunk, Webhooks, Pagerduty - **Published:** August 28, 2026 - **Apply:** https://salesforce.wd12.myworkdayjobs.com/External_Career_Site/job/Indiana---Indianapolis/Slack-Proactive-Monitoring-Engineer_JR343368 ## About the Role * 2+ years of experience in technical support, site reliability engineering, or a related operations role. * Hands-on experience with observability and monitoring tools (e.g., Grafana, Splunk, Datadog, PagerDuty, or equivalent). * Strong understanding of cloud-based SaaS architecture, APIs, and common failure modes. * Proficiency in reading and analyzing logs, metrics, and traces. * Excellent written and verbal communication skills; ability to clearly convey technical findings to both technical and non-technical audiences. * Demonstrated ability to leverage modern AI tools to optimize workflows, conduct research, and enhance daily productivity. Preferred Requirements: * Experience working with Slack platform (Slack API, Slack workflows, Bolt framework). * Familiarity with Salesforce Service Cloud / OrgCS case management. * Scripting or automation experience (Python, JavaScript, Bash). * Experience in a customer-facing support engineering or reliability role at a SaaS company. * ITIL, SRE, or similar certification. ## Description * Continuously monitor dashboards, alerting systems, and telemetry data (error rates, latency spikes, API failures, deployment anomalies) for early signals of degradation. * Triage and correlate alerts from multiple sources (Splunk, internal tools, etc) to identify patterns before customers report issues. * Actively monitor Slack platform health dashboards, network latency signals, message delivery queues, and database capacities for high-frequency workspaces. * Monitor critical custom automations, Slack Workflow Builder runs, Enterprise Key Management (EKM) operations, and Identity Provider (IDP) authentication syncs. * Identify customers potentially affected by degraded service conditions and coordinate proactive outreach with Customer Success and Support teams. * Partner with the Incident Management team to escalate signals that meet incident-threshold criteria. * Technical Advisory: Partner with Customer Success Managers and Success Architects to deliver annual technical health check reviews, assessing platform metrics, configuration limits, and custom integration health. * Perform root cause analysis (RCA) on proactively detected issues, documenting findings in internal case and incident management systems. * Work closely with Engineering and SRE teams to drive rapid remediation of identified issues * Intervene in low-risk system exceptions (e.g., advising clients on misconfigured Slack Webhooks, API rate limit exhaustion, or broken Salesforce-Slack app connections) before they trigger widespread downtime. * Build and maintain Slack-based automations and workflows to streamline proactive monitoring operations. ## Related Videos - [Applying Agile Principles to Incident Management ](https://www.wearedevelopers.com/videos/101-applying-agile-principles-to-incident-management) - [5 steps for running a Kubernetes environment at scale](https://www.wearedevelopers.com/videos/88-5-steps-for-running-a-kubernetes-environment-at-scale) - [Our journey with Spring Boot in a microservice architecture](https://www.wearedevelopers.com/videos/511-our-journey-with-spring-boot-in-a-microservice-architecture) - [Stack Overflow: Community and AI](https://www.wearedevelopers.com/videos/600-stack-overflow-community-and-ai) - [Handling incidents collaboratively is like solving a rubix cube](https://www.wearedevelopers.com/videos/680-handling-incidents-collaboratively-is-like-solving-a-rubix-cube) - [Keycloak case study: Making users happy with service level indicators and observability](https://www.wearedevelopers.com/videos/1599-keycloak-case-study-making-users-happy-with-service-level-indicators-and-observability) ## Related Articles - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [Fully Remote Software Engineer Jobs](https://www.wearedevelopers.com/magazine/447-fully-remote-software-engineer-jobs) - [Dev Digest 120 - Apple and peers](https://www.wearedevelopers.com/magazine/455-dev-digest-120-apple-and-peers) - [Is Software Engineering Over-Saturated?](https://www.wearedevelopers.com/magazine/418-is-software-engineering-over-saturated) - [Why Upskilling And Reskilling is Important For Developers](https://www.wearedevelopers.com/magazine/428-why-upskilling-and-reskilling-is-important-for-developers) - [DevOps Engineer Salary [2023]](https://www.wearedevelopers.com/magazine/203-devops-engineer-salary-2023)