REMOTE Lead Cassandra Engineer

Insight Global
Beverly Hills, CA, United States
19 days ago

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Working hours
Regular working hours
Job source

Tech stack

Artificial Intelligence Amazon Web Services Amazon Elastic Compute Cloud Amazon S3 Software as a Service Cloud Computing Continuous Integration Distributed Systems Monitoring of Systems Identity and Access Management Python (Programming Language) Performance Tuning
+16 more
Query Optimization Ansible Prometheus Scripting Apache Cassandra System Availability Large Language Models Grafana Infrastructure as Code (IaC) Amazon Virtual Private Cloud (VPC) Gitlab-ci Kubernetes Cassandra Cloudwatch Terraform Splunk

Job description

A high-scale technology client is seeking a Lead Cassandra Engineer to join their team! This individual will lead a Cassandra migration initiative, working closely with engineering teams to design and build scalable infrastructure while ensuring proper configuration and performance. They will develop automation solutions, monitor system health, troubleshoot latency and outages, and run experiments to validate upgrades and changes in production environments. The role also involves influencing architectural decisions, mentoring team members, and ensuring high availability in a distributed system. This is a remote opportunity supporting U.S. time zones with on-call responsibilities; compensation is flexible and the role is full-time.

Requirements

  • Deep, hands-on Apache Cassandra production experience (not just application-level usage)

  • Strong expertise in Cassandra schema design, query optimization, performance tuning, compaction, and repair processes

  • Experience operating Cassandra at scale (multi-region, high-availability environments, large clusters)

  • Infrastructure as Code (IaC) experience with Ansible and/or Terraform

  • Strong automation mindset with proven experience building tooling and automation workflows

  • Cloud experience (AWS preferred: EC2, S3, IAM, VPC, CloudWatch)

  • Scripting experience (e.g., Python or similar)

  • Monitoring and observability experience (e.g., Prometheus, Grafana) with ability to interpret Cassandra metrics

  • Ability to contribute to architecture discussions and influence design decisions

  • Experience working in regulated environments (PCI, SOX, or similar)

  • Experience leveraging AI/LLM tools for development productivity or automation

  • Experience with CI/CD tools (e.g., GitLab CI/CD)

  • Strong written communication and ownership mindset

  • Willingness to participate in on-call rotation * Experience with ScyllaDB and/or migration off ScyllaDB

  • Experience with EKS, Kubernetes, and Helm

  • Exposure to tools like Vault, Splunk, Vector, or Medusa/Reaper

  • Enterprise-scale environment experience (ticketing, fintech, e-commerce, gaming, SaaS)

  • Background in DBA or DBRE/SRE roles

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on dejobs.org

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

4:36 min

Hiring passionate software engineers to tackle unprecedented scaling challenges

Dana Lawson Dana Lawson +1 · World Congress 2026 Europe

1:42 min

Automating Skupper deployments using Ansible

Alex Soto Alex Soto · World Congress 2024

10:40 min

Visualizing Prometheus open metrics using custom Grafana dashboards

Stijn Polfliet · LIVE

2:38 min

Establishing comprehensive monitoring and log management

Michael Eder +1 · LIVE

3:19 min

Executing complex workflows using Ansible Automation Platform

Goetz Rieger Goetz Rieger · World Congress 2025

1:04 min

Visualizing Keycloak performance via standard Grafana troubleshooting dashboards

Alexander Schwartz Alexander Schwartz · World Congress 2025

Videos

See all

Related articles

See all