Java Performance & Reliability Engineer

ESPIRE
United States
8 days ago
Apply on www.thejobnetwork.com
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
5 years minimum
Working hours
Regular working hours

Tech stack

Java (Programming Language) Application Programming Interfaces (APIs) Amazon Web Services Applications Architecture Application Performance Management Cloud Computing Cloud Engineering Profiling Software Quality Databases Continuous Integration Data Structures
+38 more
Software Debugging DevOps Distributed Systems Memory Management Monitoring of Systems Java Virtual Machine (JVM) Apache JMeter Load Testing Log Analysis Query Optimization Redis Big O Memory Leaks Software Engineering SQL Databases Data Streaming Diagnostic Tools Multithreading Java Application Server Enterprise Software Applications Load Balancing Grafana Concurrency Spring-boot Gatling Caching Parallel Computation Indexer Containerization Kubernetes Performance Monitor Splunk Code Restructuring Appdynamics Dynatrace Docker Legacy Systems Microservices

Job description

We are seeking an experienced Senior Java Engineer with deep expertise in Java application performance, distributed systems, databases, concurrency, and production troubleshooting. The ideal candidate will be highly hands-on and comfortable working with both modern and legacy Java applications, identifying performance bottlenecks, optimizing resource utilization, and resolving complex production issues.

This role requires someone who can go beyond application development and understand how applications behave at scale across Java, databases, containers, Kubernetes, networks, caching, and infrastructure.

Responsibilities

  • Read, debug, refactor, and optimize existing Java applications across Java 8, 11, 17, and 21.
  • Diagnose and resolve complex application performance and scalability issues.
  • Analyze CPU spikes, memory leaks, high resource utilization, thread contention, and application latency.
  • Perform root cause analysis by reviewing application logs, thread dumps, heap dumps, and system metrics.
  • Optimize SQL queries and database interactions, including indexes, connection pooling, database locks, and transaction management.
  • Analyze application algorithms and identify opportunities to improve time and space complexity.
  • Troubleshoot applications operating in distributed and high-volume environments.
  • Understand and analyze traffic and data flow across applications, APIs, networks, and load balancers.
  • Design and troubleshoot multi-threaded and concurrent Java applications, including thread safety, synchronization, and resource management.
  • Work with containerized applications and Kubernetes environments.
  • Design and execute performance and load-testing strategies using JMeter, Gatling, or equivalent tools.
  • Use profiling tools such as JProfiler, YourKit, or equivalent to identify application-level performance bottlenecks.
  • Use monitoring and observability platforms such as Splunk, Dynatrace, AppDynamics, Grafana, or equivalent tools for proactive monitoring and troubleshooting.
  • Work with caching technologies such as Redis to improve application performance and scalability.
  • Collaborate with development, architecture, infrastructure, database, and production-support teams to resolve complex issues.
  • Recommend improvements to application architecture, code quality, scalability, reliability, and performance., worldwide, spread across domains like Higher Education, Insurance, Banking & Finance, Energy & Utilities, Digital Communications, Logistics, Manufacturing, Healthcare, Recruitment & Staffing, Digital Marketing Agencies & ISVs, etc.

Requirements

  • 5+ years of professional Java development experience, preferably supporting large-scale enterprise applications.
  • Strong hands-on expertise with Java 8, 11, 17, and 21.
  • Ability to understand and refactor code written by other developers, including legacy applications.
  • Deep understanding of Java memory management, JVM behavior, garbage collection, memory leaks, and CPU utilization.
  • Strong understanding of multi-threading, concurrency, synchronization, thread safety, and parallel processing.
  • Strong understanding of data structures, algorithms, and Big O time/space complexity.
  • Deep knowledge of relational databases and SQL optimization, including indexing, query optimization, connection pooling, locking, and transactions.
  • Strong experience working with distributed systems and microservices architectures.
  • Understanding of networking, load balancing, application traffic, and data flow across distributed environments.
  • Hands-on experience with Docker/containers and Kubernetes.
  • Strong production troubleshooting skills, including log analysis and root cause analysis.
  • Experience using application profiling and diagnostic tools.
  • Experience with monitoring and observability platforms.
  • Knowledge and practical experience with caching mechanisms such as Redis.

Preferred Qualifications

  • Experience supporting applications with large user bases and high transaction volumes.
  • Hands-on experience with performance and load-testing tools such as JMeter or Gatling.
  • Experience troubleshooting performance issues in cloud-native and Kubernetes environments.
  • Experience with Spring Boot and microservices-based Java applications.
  • Experience with cloud platforms such as AWS.
  • Strong understanding of application resiliency, scalability, availability, and performance engineering.
  • Experience with CI/CD and modern DevOps practices.
  • Ability to work across development and production environments and take ownership of issues through resolution.

About the company

With over 2 decades of experience behind us, and global operations spread across 11 locations worldwide - our Agile Digital Transformation Services help brands to be resilient to market disruptions and focus on business outcomes and returns.

We believe that true Digital Transformation can only be achieved with Total Experience (TX), and it is the sum of Multi-Experience (MX), User Experience (UX), Customer Experience (CX), and Employee Experience (EX).

To make this possible, we adopt a cross-enterprise approach, backed by robust operations systems - leading to meaningful customer engagements, retentions and increase in new customer acquisitions for businesses. Thereby, we are the preferred partner for our customers, and we aim to become a TX leader with end-to-end services of MX, US, CX and EX.

With 25+ Global Technology Partnerships, Espire has been empowering businesses to drive growth and customer engagement. We have served 200+ businesses worldwide, spread across domains like Higher Education, Insurance, Banking & Finance, Energy & Utilities, Digital Communications, Logistics, Manufacturing, Healthcare, Recruitment & Staffing, Digital Marketing Agencies & ISVs, etc., With over 2 decades of experience behind us, and global operations spread across 11 locations worldwide - our Agile Digital Transformation Services help brands to be resilient to market disruptions and focus on business outcomes and returns.\n\nWe believe that true Digital Transformation can only be achieved with Total Experience (TX), and it is the sum of Multi-Experience (MX), User Experience (UX), Customer Experience (CX), and Employee Experience (EX).\n\nTo make this possible, we adopt a cross-enterprise approach, backed by robust operations systems - leading to meaningful customer engagements, retentions and increase in new customer acquisitions for businesses. Thereby, we are the preferred partner for our customers, and we aim to become a TX leader with end-to-end services of MX, US, CX and EX.\n\nWith 25+ Global Technology Partnerships, Espire has been empowering businesses to drive growth and customer engagement. We have served 200+ businesses

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.thejobnetwork.com
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

1:06 min

Developer experience and project variety at scale

Alexandra Petri · World Congress 2023

2:07 min

Inspecting default bridge architectures and custom Docker networks

Oliver Seitz Oliver Seitz · World Congress 2025

3:01 min

Evaluating generated sorting algorithms and application memory complexity

Markus Walker Markus Walker · World Congress 2023

3:55 min

Demonstrating semantic routing thresholds with the Redis vector library

2:34 min

Docker sandbox architecture and microVM environment integration

Manuel de la Peña Manuel de la Peña · World Congress 2026 Europe

2:00 min

Time complexity in face verification versus face recognition

Sefik Serengil · LIVE

Videos

See all

Related articles

See all