Site Reliability Engineer

IBA InfoTech Inc.
Beaverton, OR, United States
about 2 months ago

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
5 years minimum
Working hours
Regular working hours

Tech stack

Java (Programming Language) Microsoft Windows Application Programming Interfaces (APIs) Artificial Intelligence Microsoft Azure Bash Shell Big Data C Sharp (Programming Language) Computer Programming Linux DevOps Distributed Systems
+35 more
Django Web Framework Amazon DynamoDB Apache Hadoop Monitoring of Systems Python (Programming Language) Memcached MongoDB Windows PowerShell Redis Release Management Reliability Engineering Logstash Ansible Newrelic Scala (Programming Language) Software Engineering Web Services Jetty Scripting Chatbots Flask (Web Framework) AWS Lambda Git Kubernetes Information Technology SolarWinds (Software) Apache Kafka Build Tools Bitbucket Splunk New Relic (SaaS) Docker Jenkins Servicenow Artifactory

Job description

  • Track record of succinctly documenting processes, procedures, and best practices

Requirements

  • Integration expert: able to wire systems together via their APIs
  • Fluent coder in Python, Java, Bash, C#, PowerShell or similar
  • Comfortable with Linux and Windows
  • Experience working on complex 24x7 available distributed systems
  • Familiarity with build tools, particularly Jenkins
  • Understanding and experience with Build Repositories, ideally Artifactory
  • Possession of a deep knowledge of developer workflows with Git (BitBucket)
  • Experience setting up monitoring using ScienceLogic, NewRelic or Solarwinds
  • Experience transforming Big Data into Operations Insights, ideally with Splunk
  • Comfortable leveraging AI and Machine Learning for Predictive Analysis of Failures and Correlation, ideally with Splunk
  • Experience migrating systems between technologies
  • Participation in on-call rotation
  • Runtime Infrastructure : Docker, Kubernetes, Lambdas
  • Storage Systems : DynamoDB, MongoDB, Redis, Hadoop, Memcached
  • Messaging : Kafka, Rsyslog, Logstash, Splunk
  • Programming/Scripting : Java / Jetty, Python, Scala, PowerShell, C#, Bash
  • Build Tools/Repositories : Jenkins, Artifactory
  • Web Services Framework: Django, Flask
  • API Framework: Gunicorn, * 5+ years of relevant experience focused on site reliability, TechOps, DevOps, systems administration, application development, build, release and deployment
  • Experience deploying systems with Kubernets, Docker, or Azure Containers
  • Automation experience with ScienceLogic or Ansible
  • Some experience with monitoring and anomaly detection systems
  • Hands-on experience on monitoring tools such as New Relic, Splunk, SignalFx, Solarwinds, etc.
  • Well versed with ITIL concepts (Event, Incident, Knowledge, Change, Problem Management)
  • Familiarity with Chatbots such as Slackbot
  • End user knowledge on ITSM tools such as ServiceNow

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on ibainfotech.com

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

3:55 min

Demonstrating semantic routing thresholds with the Redis vector library

6:21 min

Investigating push inefficiencies with upstream Git experts

Jonathan Creamer · Coffee With Developers

8:02 min

Integrating service level objectives into incident management

Diana Todea · LIVE

3:42 min

Comparing in-memory and Redis storage for cache scalability

Simone Sanfratello · WWC 2022

56 sec

Favorite git commands and the importance of patch commits

Eileen Uchitelle Eileen Uchitelle +1 · Coffee With Developers

1:06 min

Developer experience and project variety at scale

Alexandra Petri · WWC 2023

Videos

See all

Related articles

See all