> Markdown version of [/jobs/ext/2609119-senior-software-engineer](https://www.wearedevelopers.com/jobs/ext/2609119-senior-software-engineer). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Senior software engineer - **Company:** Alluxio, Inc. - **Location:** Foster City, CA, United States - **Experience:** Expert - **Salary:** $190,000.0 - $260,000.0 - **Contract:** Permanent contract - **Skills:** Java (Programming Language), Adobe InDesign, Artificial Intelligence, Amazon S3, Systems Engineering, Profiling, Software Quality, Distributed Data Store, Distributed Systems, Fault Tolerance, Apache Hadoop, Hadoop Distributed File System, Meta-Data Management, Open Source Technology, Performance Tuning, Posix, Concurrency, Apache Spark, Multi-Cloud, Caching, Kubernetes, Information Technology, Presto, Data Pipelines - **Published:** August 31, 2026 - **Apply:** https://www.careerboard.com/us/en/find-jobs-in-United-States/-91DA26DFC50CEA6822/ ## About the Role * Strong computer-science fundamentals and a passion for large-scale distributed systems. * Professional experience developing in Java, C+, or Go. * Practical knowledge of concurrency, replication, distributed coordination, and performance tuning. * Experience with distributed storage, caching, or data-access layers (eg, Spark, Presto, Hadoop, Kubernetes). * Bachelor's or advanced degree in Computer Science or related technical field (or equivalent experience). ## Description We're looking for experienced distributed-systems engineers to join our Core Product team and advance the next generation of Alluxio's data-orchestration engine - the foundation for AI and analytics at global scale. As a Senior Software Engineer, you'll work on high-impact systems problems such as: 1. Optimizing metadata management, caching, and replication across thousands of nodes. 2. Designing concurrent, fault-tolerant services for multi-region and multi-cloud environments. 3. Evolving Alluxio's storage abstraction and scheduling layer to support large-scale AI/ML data pipelines. 4. Collaborating with internal product teams to push the limits of distributed I/O performance. This is a hands-on, architecture-plus-implementation role for engineers who love deep systems work and want visible impact in a small, senior, highly technical team. What You'll Own * Cache and metadata enhancements - design and implement improvements to caching policies, eviction logic, and metadata scalability to increase performance and reliability. * Data path optimization - refine I/O pipelines for S3/GCS/HDFS/Posix to reduce latency and improve throughput using concurrency and scheduling techniques. * Distributed systems reliability - strengthen consistency, replication, and fault-tolerance mechanisms across large-scale clusters. * Feature development and integration - collaborate with product and solution-engineering teams to deliver features that support AI and analytics workloads. * Code quality and peer collaboration - participate in design reviews, provide constructive feedback, and ensure robust testing and observability in production systems. What You'll Do * Design, build, and optimize distributed components within Alluxio's orchestration layer. * Investigate performance bottlenecks and propose scalable solutions using profiling, tracing, and benchmarking tools. * Collaborate cross-functionally with fellow engineers, architects, and the open-source community to drive improvements. * Contribute to releases and stability efforts, ensuring enterprise-grade reliability across global deployments. ## Related Videos - [Using WebAssembly to run, extend, and secure your application](https://www.wearedevelopers.com/videos/652-using-webassembly-to-run-extend-and-secure-your-application) - [HTTP headers that make your website go faster](https://www.wearedevelopers.com/videos/1676-http-headers-that-make-your-website-go-faster) - [Alibaba Big Data and Machine Learning Technology](https://www.wearedevelopers.com/videos/37-alibaba-big-data-and-machine-learning-technology) - [On developing smartphones on wheels](https://www.wearedevelopers.com/videos/258-on-developing-smartphones-on-wheels) - [From Model to Metal: An Open Source Stack for Accelerating Intelligence](https://www.wearedevelopers.com/videos/1636-from-model-to-metal-an-open-source-stack-for-accelerating-intelligence) - [Event based cache invalidation in GraphQL](https://www.wearedevelopers.com/videos/433-event-based-cache-invalidation-in-graphql) ## Related Articles - [How We Built a Worry-Free System That Runs for 10+ Years – And What We’d Do Again](https://www.wearedevelopers.com/magazine/751-how-we-built-a-worry-free-system-that-runs-for-10-years-and-what-we-d-do-again) - [Making Data Warehouses Fast: A Developer’s Story](https://www.wearedevelopers.com/magazine/107-making-data-warehouses-fast-a-developer-s-story) - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [Top Big Data Technologies That You Need to Know](https://www.wearedevelopers.com/magazine/108-top-big-data-technologies-that-you-need-to-know) - [Top 10 Java Libraries](https://www.wearedevelopers.com/magazine/364-top-10-java-libraries) - [Dev Digest 121 - AI goes offline](https://www.wearedevelopers.com/magazine/456-dev-digest-121-ai-goes-offline)