Site Reliability Engineer
- Discuss this with your agent
- Open in Claude
- Open in ChatGPT
Prepare application
- Draft this with your agent
- Open in Claude
- Open in ChatGPT
Role details
Tech stack
+25 more
Job description
We are seeking a SRE, Intermediate to join our team. The ideal candidate will have strong experience with SRE Concepts, Kubernetes, Docker, Monitoring tools and CI/CD pipelines.
Requirements
- Proficient applied experience with SRE concepts
- Strong hands-on experience with Kubernetes (cluster administration, workload management, CNI, CSI, security policies)
- Proficient in Docker containerization (optimized Dockerfiles, registries, container security best practices)
- Deep expertise in observability/monitoring with Prometheus, Grafana, Grafana Tempo, Dynatrace, and/or Splunk (dashboards, alerting, SLO-based monitoring)
- Experience implementing CI/CD pipelines and GitOps workflows (Jenkins, GitLab, ArgoCD, Flux, Spinnaker)
- Strong understanding of cloud platforms (AWS, Azure, or Google Cloud Platform) and their container services (EKS, AKS, GKE)
- Proficiency with infrastructure as code tooling (Terraform, Helm, Kustomize, CloudFormation), including Terraform scripting
- Hands-on Ansible skills for automation/configuration management in support of reliability and standardization
- Understanding of OAuth flows and how identity integrations work in practice (including ID Anywhere patterns where applicable)
- Working understanding of APIs (design/consumption, authentication/authorization considerations, operational concerns like rate limits and observability)
- Strong scripting and automation skills (Python, Go, Bash, or similar)
- Strong collaboration and communication skills; ability to work independently with minimal supervision
- Demonstrated ability to proactively identify and resolve complex technical challenges
- Note on networking: Networking knowledge is a plus, but not a must-have; this role is intended to be more automation/coding-oriented than pure network operations
Preferred Skills:
- Advanced certifications (CKA, CKAD, and/or cloud platform certifications)
- Deep expertise in SRE practices (SLA/SLO management, error budgets, MTTR/MTTD optimization, chaos engineering)
- Experience with service mesh technologies (Istio, Linkerd, Consul) and advanced Kubernetes patterns
- Knowledge of distributed tracing tooling (Grafana Tempo, OpenTelemetry)
- Experience with log aggregation/analysis platforms (ELK Stack, Loki, Splunk)
- Familiarity with container security scanning/compliance tools (Aqua, Twistlock)
- Experience with multi-cluster management, disaster recovery, and high-availability architectures
- Understanding of FinTech/financial services regulatory requirements and operational standards
- Contributions to open-source projects related to Kubernetes, observability, or SRE tooling
- Experience with Databus onboarding patterns and operational telemetry enablement across platforms
- Strong hands-on experience standing up and operationalizing Dynatrace monitoring at scale
- Experience automating pipelines using Jules/AIM (where applicable)
Benefits & conditions
- Competitive compensation and benefits.
- Opportunities for growth with global clients.
- A supportive, inclusive culture that values innovation and people.
- Exposure to cutting-edge technologies and projects.
About the company
BCforward is a leading global IT consulting and workforce solutions firm providing services and support to Fortune 500 and government clients. Founded in 1998, BCforward has grown with our customers needs into a full-service business solutions provider. With delivery centers and offices across North America and India, we take pride in building long-term relationships and delivering excellence through innovation, collaboration, and integrity.
Apply for this position
This job is hosted externally. Click below to view the full posting and apply.
Prepare application
- Draft this with your agent
- Open in Claude
- Open in ChatGPT
Good distractions
Talks and stories from around this role — technically off-topic, practically not.
Moments
Explore playlistsVideos
See allRelated articles
See all
Is Software Engineering Over-Saturated?
Fully Remote Software Engineer Jobs
Highest Paying Tech Companies for Developers
Why Upskilling And Reskilling is Important For Developers