SRE/Linux Engineer

Insight Global
Hopewell Township, NJ, United States
1 day ago
Apply on jobs.insightglobal.com
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
5 years minimum
Compensation
$106,080.0 - $133,120.0
Working hours
Regular working hours

Tech stack

Multitier Architecture Active Directory Amazon Web Services Apache HTTP Server Apache Tomcat Microsoft Azure Oracle WebLogic Server Cloud Computing Databases Continuous Integration Quartz (Graphics Layer) Data Centers
+25 more
Linux DevOps Disaster Recovery Django Web Framework Middleware Monitoring of Systems Web Servers IBM Websphere Application Server WildFly (JBoss AS) Job Scheduling Python (Programming Language) Linux System Administration Linux Servers Enterprise Messaging Systems MySQL Release Management Reliability Engineering Ansible Software Configuration Management Scripting Grid Computing Restful APIs Splunk Network Server Dynatrace

Job description

Insight Global is seeking a SRE/Linux Engineer to join one of our largest financial clients in New York City supporting the Markets APSE Support Team. This individual will be focused on operations and daily support of the Linux Platform this team oversees. The ideal candidates will possess extensive experience in scripting with Python and Shell and will install, maintain, monitor and manage Infrastructure support, Production support and system Administration tasks of the Linux servers. This role is vital for overseeing and supporting significant projects, including data center migration, disaster recovery, and capacity management., * Managing core Infrastructure of Quartz Operation environment and components

  • Managing and implementing end-to-end Linux based infrastructure that hosts many critical core components such as Database, GRID Computing, Enterprise Job Schedulers Messaging Platforms. Maintain a vast network of 10,000 Linux servers. Ensure system performance and reliability through routine maintenance and upgrades.
  • Supporting and monitoring application infrastructure and various internal core components and dependent technologies.
  • Ensures application issues on supported infrastructures are addressed, proactively fixed even before user impact.
  • Will support release management, troubleshooting issues, Incident problem management etc.
  • Support coverage for migration projects related to servers, Active Directory OU , Datacenter, Non Permitted technology ,Upgrade of Applications , LEAPP, etc
  • Handle all Tier II escalations related to Ecommerce infrastructure hosted on Linux Servers on both On-Premise and Cloud along with Middleware technologies like Apache/Tomcat, WebLogic and WebSphere Servers.
  • Developing the Scripts in Shell and Python to automate the day-to-day activities performed within the environment (DEV/UAT & PROD). Collaborate with development teams to integrate new scripts into existing systems.
  • Supporting and monitoring in-house developed components like DB-Sandra, Grid Computing-Hugs, Job Scheduler-BOB, Authorization- Quack, Messaging tool- AMPS.
  • Good knowledge on organization of Data Center infrastructure (Server, Storage, Network) and DC facility such as cooling, racks Management.
  • Experience with Software Configuration Management, Build and Release Management, continuous Integration and Deployment.
  • Hands-on resolution of complex tickets in a large-scale monitoring enterprise environment Remedy. Develop and implement disaster recovery plans to ensure business continuity.
  • Capacity Management: Monitor system performance and optimize capacity management. Utilize monitoring tools like Splunk and Dynatrace to identify and address performance bottlenecks.
  • ITIL Processes: Utilize Remedy for ITIL-based incident and problem management. Maintain adherence to ITIL best practices and document properly.

Requirements

Must Haves:

  • 5+ years of experience as a SRE/Linux Engineer
  • Strong Python development experience
  • Hands-on Django and REST API development
  • Strong MySQL/database skills
  • Deep Linux administration and troubleshooting experience
  • Infrastructure engineering and automation background
  • Experience with Ansible and CI/CD pipelines
  • Observability and monitoring experience (Dynatrace preferred)
  • Experience working in large-scale production environments
  • Azure and/or AWS cloud experience
  • SRE, Reliability Engineering, and DevOps practices
  • OpenTelemetry and modern observability tools
  • Infrastructure-as-Code and automation frameworks
  • Financial Services / Trading Platform experience

Nice to Have Skills & Experience

Plusses:

  • Previous Bank of America experience
  • Experience with Middleware products app servers and web servers (Tomcat, Apache, WebSphere, JBOSS, Oracle WebLogic, etc.)
  • Experience in data center migrations, capacity management, and disaster recovery.
  • Cloud experience is an optional plus

Benefits & conditions

Benefit packages for this role will start on the 1st day of employment and include medical, dental, and vision insurance, as well as HSA, FSA, and DCFSA account options, and 401k retirement account access with employer matching. Employees in this role are also entitled to paid sick leave and/or other paid time off as provided by applicable law.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on jobs.insightglobal.com
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

1:34 min

Pivoting careers into specialized platform engineering roles

Xavier Portilla Edo · LIVE

2:38 min

Establishing comprehensive monitoring and log management

Michael Eder +1 · LIVE

52 sec

Running persistent Linux environments directly on Windows

Ben Breard Ben Breard · World Congress 2025

2:18 min

Scaling MySQL databases for massive user growth

Johannes Nicolai Johannes Nicolai +1 · LIVE

8:22 min

Simulating a Linux terminal and running Spring Boot

Jakov Semenski · LIVE

3:10 min

Correlating dispersed logs using structured request tracing

Michael Eder +1 · LIVE

Videos

See all

Related articles

See all