Site Reliability Engineer

American IT Systems
Parsippany-Troy Hills, NJ, United States
7 days ago
Apply on www.careerjet.com
Prepare application

Role details

Contract type
Temporary to permanent
Employment type
Full-time (> 32 hours)
Experience level
Expert
Compensation
$98,000.0 - $164,000.0
Working hours
Regular working hours

Tech stack

Artificial Intelligence Bash Shell BigQuery Cloud Computing Cloud Storage Databases System Configuration Python (Programming Language) PostgreSQL Performance Tuning Reliability Engineering Ansible
+14 more
SQL Databases Scripting Google Cloud Reliability of Systems Firebase Gitlab-ci Kubernetes Machine Learning Operations Puppet Terraform Looker Analytics Atlassian Bamboo Jenkins Databricks

Job description

Site Reliability Engineer:We are looking for a talented Site Reliability Engineer (SRE) with a strong background in Google Cloud Platform (GCP) and kubernetes. The ideal candidate will be responsible for ensuring the reliability, performance, and scalability of our on premise and cloud based systems along with focus on reducing costs for Google Cloud. System Reliability: Ensure the reliability and uptime of critical services and infrastructure. Google Cloud Expertise: Design, implement, and manage cloud infrastructure using Google Cloud services. Automation: Develop and maintain automation scripts and tools to improve system efficiency and reduce manual intervention. Monitoring and Incident Response: Implement monitoring solutions and respond to incidents to minimize downtime and ensure quick recovery. Collaboration: Work closely with development and operations teams to improve system reliability and performance. Capacity Planning: Conduct capacity planning and performance tuning to ensure systems can handle future growth.Documentation: Create and maintain comprehensive documentation for system configurations, processes, and procedures.

Requirements

Proficiency in Google Cloud services (Compute Engine, Kubernetes Engine, Cloud Storage, Pub/Sub, etc.). Experience with database technologies (SQL & no SQL AlloyDB (PostgreSQL), DataBricks, Firestore, BigQuery, etc.). Familiarity with Google BI and AI/ML tools (Looker, BigQuery ML, Vertex AI, etc.). Experience with automation tools (Terraform, Ansible, Puppet). Familiarity with CI/CD pipelines and tools (Azure pipelines Jenkins, GitLab CI, etc.). Strong scripting skills (Python, Bash, etc.).

About the company

LHC Group

  • Basking Ridge, NJ
  • $112,700-193,200 per year Optum Tech is a global leader in health care innovation. Our teams develop cutting-edge solutions that help people live healthier lives and help make the health system work better …

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.careerjet.com
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

2:26 min

Understanding Puppeteer and its underlying architectural design

Miki Lombardi · JS Congress

1:02 min

Applying an ETL methodology to infrastructure configuration management

Axel Barbier · World Congress 2023

1:42 min

Automating Skupper deployments using Ansible

Alex Soto Alex Soto · World Congress 2024

3:36 min

Critical infrastructure and performance skills for modern developers

Andrew Holway · LIVE

4:01 min

Comparing Terraform to popular configuration management tools

Devlin Duldulao · LIVE

1:34 min

Pivoting careers into specialized platform engineering roles

Xavier Portilla Edo · LIVE

Videos

See all

Related articles

See all