> Markdown version of [/jobs/ext/539606-sr-platform-engineer-1](https://www.wearedevelopers.com/jobs/ext/539606-sr-platform-engineer-1). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Sr Platform Engineer-1 - **Company:** Flexential Corp. - **Location:** Centennial, CO, United States - **Experience:** Expert - **Salary:** $150,000.0 - $165,000.0 - **Contract:** Permanent contract - **Skills:** Artificial Intelligence, Amazon Web Services, Application Layers, Software Applications, Automation of Tests, Code Review, Continuous Integration, Data Centers, Noise Reduction, Linux, DevOps, JSON, Python (Programming Language), Key Management, Knowledge Management, Linux System Administration, NetApp Applications, Networking Basics, Platform as a Service (PAAS), Role-Based Access Control, Ansible, Prometheus, Service Discovery, Simple Network Management Protocols, Systems Integration, TCP/IP, Zabbix, Scripting, Data Storage Management, Cloud Platform System, Cyberark, System Availability, Grafana, Boomi, Backend, Gitlab, Gitlab-ci, Kubernetes, Infrastructure Automation Frameworks, Hashicorp, ArcSight Event Correlation, Cloudwatch, Restful APIs, Terraform, Docker, Servicenow, Vmware - **Published:** June 14, 2026 - **Apply:** https://www.indeed.com/viewjob?jk=c7e7de43eaf74c1c ## About the Role Do you have experience in vCenter?, * DevOps / Automation - 5+ years in a production environment, Kubernetes (RKE2/k3s), Helm chart deployment, system services, Docker/container * LGTM Stack Development and Configuration - 4+ years: Grafana, Mimir, Loki, Tempo configuration, tuning, dash-boarding and production operations; Prometheus required * Senior-level Python / Scripting frameworks - 5+ years, Automation scripts, exporter development, GitLab pipeline scripting, REST API integrations * GitOps / CI/CD - 5+ years, GitLab CI/CD pipeline authoring; Terraform and Ansible as primary IaC tools; ArgoCD or Flux preferred * AIOps / Observability Engineering - 2+ years, Alertmanager rule authoring, anomaly detection integration, event correlation, noise reduction techniques * Working infrastructure (Linux/VM) management knowledge - 5+ years, Linux administration, VMware vCenter/VCF experience, Netapp storage management, network fundamentals (SNMP, TCP/IP) * Secrets Management - 2+ years, CyberArk/Conjur, HashiCorp Vault, or equivalent - runtime secret injection patterns * Minimal travel may be required Preferred Skills: * Experience and/or knowledge of ITSM processes and workflow automation e.g. Incident & Response Mgmt (IRM), Release mgmt., ServiceNow ITSM integration, alert routing, escalation policy design, SLA-driven on-call workflows * Hands-on experience or working knowledge of Boomi integrations PaaS(iPaaS) technologies * Experience working with BAS / BMS systems in a Datacenter / OT environment. * Hands-on experience working with AWS products in a Well-architected Framework and multi-account model to develop various compute, storage, network iaaS and PaaS services for IT applications. ## Description The Senior Platform Engineer is a hands-on engineering role on a platform development team responsible for building and operating Flexential's IT platforms including observability, devops, ITSM incident and release mgmt, and Integrations technologies. This role develops and manages critical platform subsystems for high availability, operational resiliency, security and scalability utilizing native-AI enablement for all outcomes. This is an individual contributor role with significant technical ownership and direct impact on critical Flexential technology roadmap You will work across infrastructure, automation, and application layers - deploying Kubernetes workloads, authoring Terraform modules, building Ansible playbooks, and building GitLab pipelines that other engineers depend on daily. Key Responsibilities and Essential Job Functions: * Design, develop and operationally manage automated, resilient, high availability, self-healing, secure platforms with native-AI capabilities for IT needs, serving both internal as well as customer business capabilities * Develop, and manage the Observability OpenTelemetry Central Backend Stack: Grafana Enterprise, Mimir, Loki, Tempo, and Alertmanager on Kubernetes/RKE2 via Helm and GitLab CI-CD. * Build and manage iaC and CI-CD for automated provisiong and deployment, including Terraform modules for Infra/VM/storage provisioning, Ansible AWX playbooks for OS/App bootstrap, ArgoCD and Helm for Kubernetes configuration. * Develop and manage OpenTelemetry Prometheus scrape profile library including SNMP exporters, REST API exporters, and cloud provider exporters (CloudWatch, Azure Monitor, GCP) for multiple device classes. * Develop AIOps capabilities on platforms for e.g Observability use-cases: anomaly detection integrations, event correlation rules in Alertmanager, and synthetic monitoring patterns to reduce alert noise. * Configure and maintain Zabbix auto-discovery: network range scanning, device classification, and Prometheus service discovery integration. * Build and harden Edge Stack deployments (Prometheus + OTel collector) per data center site using GitOps templates. * Integrate Alertmanager with ServiceNow: webhook routing, ticket enrichment, auto-close logic, and escalation policy configuration. * Maintain platform security: Conjur/CyberArk secret injection at runtime, mTLS between stack components, RBAC in Grafana Enterprise. * Author and maintain Grafana dashboards in JSON/GitLab - facility overview, network health, RED metrics, application telemetry. * Mentor mid-level engineers, lead code reviews, and establish engineering standards for the team. Represent platform engineering in cross-functional architecture reviews and executive-level program updates. * Perform other duties as required and assigned ## Related Videos - [Docker network without Docker](https://www.wearedevelopers.com/videos/1418-docker-network-without-docker) - [Tips and Tricks for Working with JSON](https://www.wearedevelopers.com/videos/1229-tips-and-tricks-for-working-with-json) - [Shipping Faster with Less: Render on Cloud Hosting, AI Workloads, and the Future of DevOps](https://www.wearedevelopers.com/videos/1894-shipping-faster-with-less-render-on-cloud-hosting-ai-workloads-and-the-future-of-devops) - [Docker exec without Docker](https://www.wearedevelopers.com/videos/1094-docker-exec-without-docker) - [Introducing JSON Structure](https://www.wearedevelopers.com/videos/100219-introducing-json-structure) - [My journey into DevOps world - How it all started!](https://www.wearedevelopers.com/videos/545-my-journey-into-devops-world-how-it-all-started) ## Related Articles - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [What Are The Top Skills Required For Azure Developers?](https://www.wearedevelopers.com/magazine/77-what-are-the-top-skills-required-for-azure-developers) - [Best US AI Conferences for CTOs in 2026: Build vs. Buy, Vendor Evaluation, and Peer Intelligence](https://www.wearedevelopers.com/magazine/736-best-us-ai-conferences-for-ctos-in-2026-build-vs-buy-vendor-evaluation-and-peer-intelligence) - [Is Software Engineering Over-Saturated?](https://www.wearedevelopers.com/magazine/418-is-software-engineering-over-saturated) - [Top Must-Visit Developer Conferences in the US in 2026](https://www.wearedevelopers.com/magazine/679-top-must-visit-developer-conferences-in-the-us-in-2026) - [The Best Job Search Websites of 2025](https://www.wearedevelopers.com/magazine/368-the-best-job-search-websites-of-2025)