Software Engineer (Platform)

SnapLogic, Inc.
San Mateo, CA, United States
20 days ago
Apply on arc.dev
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Working hours
Regular working hours
Job source

Tech stack

Java (Programming Language) Computer Engineering DevOps Monitoring of Systems Python (Programming Language) Reliability Engineering Data Logging System Availability Grafana

Job description

We are seeking a highly motivated and dedicated Software Developer (Platform) to join our Engineering team. Your primary focus will be on maintaining and significantly enhancing the reliability, scalability, and performance of our core platform. Our system is critical, processing over 1 trillion documents monthly for more than 1,000 global enterprises. Your direct contributions will be crucial to ensuring platform robustness and a seamless experience for our users at this massive scale. This role involves hands-on system monitoring, deep-dive investigations into production issues, defect resolution, and implementing key optimizations to boost overall platform stability and high availability.

What You’ll Do:

  • Proactively monitor and analyze platform performance, availability, and critical metrics using advanced observability tools to detect anomalies and prevent service incidents
  • Rapidly diagnose and troubleshoot complex production issues, and conduct thorough root cause analysis
  • Address and resolve bugs and system defects promptly
  • Design and implement system optimizations and reliability enhancements to guarantee a scalable, stable, and highly performant platform
  • Collaborate with engineers, DevOps, and product teams

Requirements

  • Bachelor’s Degree in Computer Engineering or related fields
  • Proficiency in Python is required. Experience with Java is an advantage
  • Strong troubleshooting skills with a sense of end-to-end ownership - from issue detection through resolution and implementation of preventative measures
  • Understanding of fundamental system performance, scalability, and reliability engineering principles
  • Eagerness and capacity to quickly learn and adopt new technologies
  • Strong self-management skills and a high sense of responsibility
  • Strong attention to detail
  • Experience with system monitoring, logging, and observability tools
  • Experience with enterprise-grade production systems.

About the company

At SnapLogic, we believe in empowering people - customers and employees alike - to integrate everything and create anything. From competitive salaries and equity packages to global wellness benefits, we’re committed to your success and well-being.

A Few Reasons You’ll Love it Here:

We’re Innovators

SnapLogic pioneered the first generative integration solution, SnapGPT, and continues to lead with a full suite of AI-powered tools - making integration faster, smarter, and accessible to more people.

We’re Recognized Leaders

From being named a Visionary in multiple Gartner Magic Quadrants, leading the market in innovative AI reports from Aragon Research, or being recognized for AI in the Cloud Awards, we’re setting the pace in a rapidly evolving market.

We’re Growing Fast

Named one of Inc. 5000’s Fastest Growing Private Companies in 2024, SnapLogic is scaling globally - and we want you to grow with us.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on arc.dev
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

2:17 min

Mapping the maturity roadmap for scaled devops adoption

Dominik Krichbaum Dominik Krichbaum · World Congress 2026 Europe

10:40 min

Visualizing Prometheus open metrics using custom Grafana dashboards

Stijn Polfliet · LIVE

1:10 min

Exposing sensitive information through partial search logs

Dennis Schulz Dennis Schulz +1 · World Congress 2026 Europe

4:18 min

Prioritizing communication and structural awareness over strict tool mastery

Liam Hurrel +1 · World Congress 2021

1:56 min

Discovering incidents using logs, metrics, and traces

Nele Uhlemann · World Congress 2023

3:18 min

Scaling global network engineering through DevOps culture

Stuart Clark · LIVE

Videos

See all

Related articles

See all