Distributed Systems Engineer
- Discuss this with your agent
- Open in Claude
- Open in ChatGPT
Prepare application
- Draft this with your agent
- Open in Claude
- Open in ChatGPT
Role details
Tech stack
+11 more
Job description
best fits their depth. The work spans query engines, distributed runtimes, storage layers, streaming systems, and metadata services at petabyte and enterprise scale, with ample scope for upstream contribution. What You’ll Do Contribute production-grade code to Apache big data projects. Debug and optimize engine internals - query planning, distributed execution, scheduling, state management, replication, storage layers, and metadata services - at petabyte scale. Influence architectural direction for performance and scalability at the engine layer. Profile and tune JVM behavior (GC, memory layout, concurrency). Collaborate with cross-functional engineering teams and open source committers on integrations and ecosystem work. Mentor senior engineers and raise the engineering bar through code reviews and design critiques. What We Are Looking For 6+ years of experience in software development. Strong Java and/or Scala skills., Experience with distributed systems and concurrent or parallel
Requirements
programming, Working knowledge of internals of at least one Apache big data project: Spark, Flink, Trino, Ozone, Iceberg, Hive, NiFi, Kafka, Hadoop, HBase, Impala, or Kudu. Familiarity with JVM performance characteristics (GC, memory, threading). Advanced level of English. Nice To Have Upstream contributions to Apache big data projects; committer or PMC status is a strong plus. Experience operating distributed systems at petabyte scale in production. Experience with adjacent or comparable engines (PrestoDB, Impala, Druid, Pinot, ClickHouse, CockroachDB). Kubernetes and cloud-native deployment experience. Public technical presence (talks, blogs, OSS community leadership). How we do make your work (and your life) easier: Remote Work. Excellent compensation in USD or your local currency if preferred. Hardware and software setup for you to work from home. Flexible hours: create your own schedule. Paid parental leaves, vacations, and national holidays. Innovative and multicultural work
About the company
At BairesDev®, we’ve been leading the way in technology projects for over 15 years. We deliver cutting-edge solutions to giants like Google and the most innovative startups in Silicon Valley. Our diverse 4,000+ team, composed of the world’s Top 1% of tech talent, works remotely on roles that drive significant impact worldwide. When you apply for this position, you’re taking the first step in a process that goes beyond the ordinary. We aim to align your passions and skills with our vacancies, setting you on a path to exceptional career development and success. Distributed Systems Engineer (Apache Big Data Internals) at BairesDev We are seeking a Distributed Systems Engineer with experience in the internals of Apache big data projects to work on the engines themselves, not applications built on top of them. We have multiple openings across the ecosystem, including Spark, Flink, Trino, Ozone, Iceberg, Hive, NiFi, Kafka, Hadoop, and adjacent projects, and we match candidates to the role that
Apply for this position
This job is hosted externally. Click below to view the full posting and apply.
Prepare application
- Draft this with your agent
- Open in Claude
- Open in ChatGPT
Good distractions
Talks and stories from around this role — technically off-topic, practically not.
Moments
Explore playlistsVideos
See allRelated articles
See all
Fully Remote Software Engineer Jobs
Making Data Warehouses Fast: A Developer’s Story
Highest Paying Tech Companies for Developers
Dev Digest 121 - AI goes offline