Site Reliability Engineer

BINGHAMTOM UNIVERSITY
United States
8 days ago
Apply on www.themuse.com
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
5 years minimum
Working hours
Regular working hours

Tech stack

Artificial Intelligence Amazon Web Services Amazon Elastic Compute Cloud Amazon S3 Bash Shell Cloud Computing Continuous Integration Amazon DynamoDB Perl (Programming Language) Monitoring of Systems Identity and Access Management Python (Programming Language)
+27 more
PostgreSQL Log Analysis Microsoft SQL Server MySQL Standard Sql Scripting Grafana Amazon Virtual Private Cloud (VPC) Git Cloudformation Amazon Relational Database Service Containerization Kubernetes Infrastructure Automation Frameworks Information Technology Graphql Cloudwatch Terraform Splunk New Relic (SaaS) Autodesk Autocad Dynatrace Docker Pagerduty Jenkins Servicenow Artifactory

Job description

An exciting new opportunity has opened for a Site Reliability Engineer within the Autodesk Customer success Technical Advisory team. The successful candidate will wear multiple hats: first responder, performance analyst, system architect, capacity planner, and monitoring expert. Good technical and communication skills are required for you to be successful in this role. A passion for learning new technologies and a desire to solve problems you come across are the keys to being successful in this role.

Roles and Responsibilities

  • Architect and implement hosting solutions for highly dynamic SaaS web applications, ensuring reliability and performance at scale
  • Design, implement, and maintain Infrastructure-as-Code solutions to support scalable, reliable, and secure global environments
  • Ensure efficient and well-documented standards, processes, and best practices
  • Implement security and compliance with best practices across infrastructure and applications, including hardening, enforcing least privileges
  • Use modern administration tools like Docker, Terraform, AWS CloudFormation/CDK to manage and deploy containers and virtual machines
  • Collaborate with development, testing, and documentation teams during the product development cycle to ensure quality control
  • Automate processes and integrate new technologies as needed
  • Define and monitor Service Level Objectives (SLOs), Service Level Indicators (SLIs), and manage error budgets to ensure reliability goals are met
  • Work with stakeholders to align technical strategy with business requirements
  • Participate in on-call support and incident management, ensuring timely resolution and clear communication
  • Conduct post-incident blameless postmortems to identify root causes and drive continuous improvement
  • Self-starter mindset with a strong drive to learn and own initiatives to promote a culture of continuous improvement, and excellence
  • Design and maintain monitoring, logging, and observability frameworks to ensure full visibility into system health
  • Use observability platforms such as CloudWatch, Dynatrace, Splunk, New Relic, or Grafana to identify performance bottlenecks and optimize system performance

Requirements

  • 5+ years DevOps/SRE experience with cloud-based applications
  • Advanced hands-on experience Linux administration skills, including monitoring, troubleshooting, reliability and security
  • Experience with incident management platforms (PagerDuty, Opsgenie, ServiceNow)
  • Familiarity with AI-assisted operations (AIOps) for intelligent alerting, anomaly detection, and incident triage
  • Experience managing large-scale cloud infrastructure (AWS preferred)
  • Strong scripting abilities (eg. Bash, Python, Perl, etc)
  • Expert-level knowledge of AWS services (EC2, ECS, EKS, Lambda, ELB, S3, IAM, VPC, Dynamo DB, RDS, etc)
  • Hands-on experience with Docker, Kubernetes and container technologies
  • Proficiency with infrastructure-as-code tools (Terraform, CloudFormation)
  • Experience with CI/CD tools. (Jenkins, Artifactory, GIT, etc)
  • Skilled in log analysis and monitoring tools (CloudWatch, Splunk, Dynatrace, New Relic, Grafana)
  • Experience with relational/vector databases (MySQL, PostgreSQL, MSSQL), RAG and GraphQL, SQL
  • Excellent problem-solving skills and ability to work independently
  • Excellent written and verbal communication skills
  • Bachelor’s degree in computer science or related field

Benefits & conditions

Salary is one part of Autodesk’s competitive compensation package. Offers are based on the candidate’s experience and geographic location. In addition to base salaries, our compensation package may include annual cash bonuses, commissions for sales roles, stock grants, and a comprehensive benefits package.

About the company

Welcome to Autodesk! Amazing things are created every day with our software - from the greenest buildings and cleanest cars to the smartest factories and biggest hit movies. We help innovators turn their ideas into reality, transforming not only how things are made, but what can be made.

We take great pride in our culture here at Autodesk - it’s at the core of everything we do. Our culture guides the way we work and treat each other, informs how we connect with customers and partners, and defines how we show up in the world.

When you’re an Autodesker, you can do meaningful work that helps build a better world designed and made for all. Ready to shape the world and your future? Join us!, Working in sales at Autodesk allows you to build meaningful relationships with customers while growing your career. Join us and help make a better, more sustainable world. Learn more here: https://www.autodesk.com/careers/sales

Diversity & Belonging We take pride in cultivating a culture of belonging where everyone can thrive. Learn more here: https://www.autodesk.com/company/diversity-and-belonging

Are you an existing contractor or consultant with Autodesk?

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.themuse.com
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

1:34 min

Pivoting careers into specialized platform engineering roles

Xavier Portilla Edo · LIVE

2:38 min

Establishing comprehensive monitoring and log management

Michael Eder +1 · LIVE

2:18 min

Scaling MySQL databases for massive user growth

Johannes Nicolai Johannes Nicolai +1 · LIVE

6:21 min

Investigating push inefficiencies with upstream Git experts

Jonathan Creamer · Coffee With Developers

3:10 min

Correlating dispersed logs using structured request tracing

Michael Eder +1 · LIVE

1:48 min

Analyzing network packets with database protocol tools

Daniël van Eeden Daniël van Eeden · World Congress 2026 Europe

Videos

See all

Related articles

See all