Python Site Reliability / Platform Engineer

BCforward
Pennington, NJ, United States
1 day ago

Role details

Contract type
Temporary contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Compensation
$132,101.0
Working hours
Regular working hours
Job source

Tech stack

Amazon Web Services Microsoft Azure Cloud Computing Quartz (Graphics Layer) Relational Databases Linux DevOps Django Web Framework Python (Programming Language) Linux System Administration MySQL Performance Tuning
+6 more
Reliability Engineering Ansible Cloud Platform System Grafana Restful APIs Dynatrace

Job description

  • Own the reliability and operational health of the Quartz platform across Global Markets.
  • Actively monitor, triage, and resolve infrastructure and application issues in large-scale production environments.
  • Design and implement automation to improve operational efficiency and reduce toil.
  • Build and enhance internal tools to support platform operations and incident response.
  • Lead incident response, root cause analysis, and post-incident reviews to drive reliability improvements.
  • Partner with Quartz core teams on platform modernization and migration to Cloud 1.0 and 2.0.
  • Harden production systems and evolve observability and support tooling.

Requirements

We are seeking a Site Reliability Engineer III to join our dynamic team. The ideal candidate will have strong experience in Python, Django, REST APIs, MySQL, Linux, automation, observability, and cloud platforms and a proven ability to improve platform reliability, enhance observability, and drive operational efficiency at scale., * Strong Python development, including hands-on Django and REST API development.

  • Proficiency with MySQL and relational database design and tuning.
  • Deep Linux administration and troubleshooting experience.
  • Infrastructure engineering and automation background with experience in Ansible and CI/CD pipelines.
  • Observability and monitoring experience, with Dynatrace preferred.
  • Experience operating in large-scale production environments.
  • Cloud experience with Azure and/or AWS.
  • Knowledge of SRE, Reliability Engineering, and DevOps practices.
  • Familiarity with OpenTelemetry and modern observability tools.

Preferred Skills:

  • Experience with Infrastructure-as-Code and automation frameworks.
  • Background in Financial Services or Trading Platforms.

Benefits & conditions

  • Competitive compensation and benefits.
  • Opportunities for growth with global clients.
  • A supportive, inclusive culture that values innovation and people.
  • Exposure to cutting-edge technologies and projects.

About the company

BCforward is a leading global IT consulting and workforce solutions firm providing services and support to Fortune 500 and government clients. Founded in 1998, BCforward has grown with our customers needs into a full-service business solutions provider. With delivery centers and offices across North America and India, we take pride in building long-term relationships and delivering excellence through innovation, collaboration, and integrity.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.dice.com

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

52 sec

Running persistent Linux environments directly on Windows

Ben Breard Ben Breard · WWC 2025

2:18 min

Scaling MySQL databases for massive user growth

Johannes Nicolai Johannes Nicolai +1 · LIVE

2:17 min

Mapping the maturity roadmap for scaled devops adoption

Dominik Krichbaum Dominik Krichbaum · WWC Europe 2026

1:34 min

Pivoting careers into specialized platform engineering roles

Xavier Portilla Edo · LIVE

8:22 min

Simulating a Linux terminal and running Spring Boot

Jakov Semenski · LIVE

3:55 min

Demonstrating .NET installation on Debian and Azure Linux

Silvano Coriani Silvano Coriani · Europe 2026 Virtual

Videos

See all

Related articles

See all