SRE Leader with AI Experience

Talent Vista Inc.
United States
4 days ago

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
10 years minimum
Working hours
Regular working hours
Job source

Tech stack

Java (Programming Language) Agile Methodology Artificial Intelligence Amazon Web Services Applications Architecture Application Lifecycle Management Application Performance Management Business Process Modeling DevOps Spring Framework Newrelic Software Engineering
+11 more
Datadog Cloudbees Cloud Platform System Grafana Software Application Programming Information Technology Presto Splunk Dynatrace Jenkins Microservices

Job description

  • Collaborate with cross-functional teams to design, develop, test and deploy scalable, reliable, and high-performance applications.
  • Implement best practices for application design, testing, deployment, monitoring, and maintenance to ensure optimal performance and availability.
  • Develop automation tools and scripts to streamline processes and improve efficiency in application lifecycle management.
  • Conduct thorough analysis of application performance metrics and system health to identify areas for optimization and enhancement.
  • Proactively identify and address potential issues and bottlenecks in the application architecture to prevent downtime and service disruptions.
  • Stay updated on emerging technologies and industry trends to drive continuous improvement and innovation in application design and development practices.
  • Familiarity with Agile methodologies and DevOps practices

Requirements

(Experience, Qualification, Knowledge & Skills)

  • 10 + years of IT experience with Java Web/Enterprise projects
  • Good Understanding of Runtime Support and Application Development lifecycle
  • Good understanding about the Concepts / Design for reliability (i.e. Automation, Scaling, Auto Recovery, Redundancy, Availability)
  • Should have a good understanding of App/Infra capacity planning, SLA/SLOs
  • Strong proficiency in programming languages such as Java, sprint boot, microservices and Java Frameworks
  • Solid understanding of AWS cloud computing platforms, AWS Certification preferred.
  • Proficiency in implementing and maintaining CI/CD pipelines (preferred Cloudbees / Jenkins)
  • Excellent understanding on APM/ observability tools like NewRelic/ Datadog/ Dynatrace, Splunk etc.
  • Hands on with PCF - App-Pilot & Presto
  • Should have skills to doing POCs and reflect POC to enterprise scale
  • Excellent problem-solving skills and attention to detail.
  • Good understanding of AI technologies
  • Knowledge and experience on the Gen AI tools and technology like GHCP, claud, AWS bedrock
  • focuses on reliability, scalability, and performance of large-scale
  • Strong communication and collaboration skills

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.dice.com

Good distractions

Talks and stories from around this role β€” technically off-topic, practically not.

2:14 min

Exploring internal AI product initiatives and global engineering roles

Maria Apazoglou Β· Coffee With Developers

2:38 min

Establishing comprehensive monitoring and log management

Michael Eder +1 Β· LIVE

2:17 min

Mapping the maturity roadmap for scaled devops adoption

Dominik Krichbaum Dominik Krichbaum Β· WWC Europe 2026

3:10 min

Correlating dispersed logs using structured request tracing

Michael Eder +1 Β· LIVE

3:18 min

Scaling global network engineering through DevOps culture

Stuart Clark Β· LIVE

1:34 min

Transitioning from traditional software development to artificial intelligence consulting

Patrick Schnell Patrick Schnell Β· Coffee With Developers

Videos

See all

Related articles

See all