Senior Software Engineer - Cloudera Context Search Team

Cloudera, Inc.
San Jose, CA, United States
3 months ago

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
0 years minimum
Compensation
$152,000.0 - $190,000.0
Working hours
Regular working hours
Job source

Tech stack

Query Performance Java (Programming Language) Artificial Intelligence Amazon Web Services Apache HTTP Server Audit Trail Microsoft Azure Big Data Software Bug Management Cloud Computing Cloud Engineering Apache Lucene
+25 more
Codecs Computer Programming Databases Data Sharing Shard (Database Architecture) Distributed Systems Apache Hadoop Apache Hive Java Virtual Machine (JVM) Open Source Technology Performance Tuning Role-Based Access Control Software Defined Everything Cloudera Systems Integration Transport Layer Security Google Cloud Apache Spark Kubernetes Helm Charts Generative AI Kubernetes Information Technology Apache Flink Deployment Automation Data Management

Job description

You will bridge the gap between big data storage and real-time retrieval, ensuring that OpenSearch operates seamlessly within our containerized (Kubernetes) and multi-cloud environments. As a Sr. Software Engineer you will:

  • Architect & Scale: Design and implement large-scale OpenSearch clusters capable of handling petabytes of data with low-latency indexing and query performance.
  • Platform Integration: Deeply integrate OpenSearch with CDP components (e.g., Apache Iceberg, SDX, and Ozone) to provide a unified search experience across the data lakehouse.
  • Performance Tuning: Optimize JVM settings, shard allocation strategies, and query DSL to ensure maximum throughput and stability.
  • Security & Governance: Implement enterprise-grade security including RBAC, TLS, and audit logging, ensuring compliance with Cloudera’s Shared Data Experience (SDX) standards.
  • Cloud Native Operations: Develop and maintain Kubernetes Operators and Helm charts for automated deployment, scaling, and self-healing of search services.
  • Community Contribution: Act as a liaison to the upstream OpenSearch community, contributing bug fixes, features, and performance improvements., You will tackle complex distributed systems challenges, crafting the foundational software for the control and data planes that powers CDP and keeps it running at massive scale. Working at the forefront of hybrid and multi-cloud technology, you will empower data scientists, engineers, and analysts with the tools and infrastructure they need for advanced analytics and modeling. Collaboration is key, you will work alongside brilliant minds across product, data science, and engineering to drive innovation, standardize best practices, and shape the future of enterprise AI and data platforms. This is your chance to build the future of data and see your work make a global impact. The expected base salary range for this role in

Requirements

  • Bachelor’s degree in Computer Science or equivalent and 5-6 years of related experience; OR Master’s degree and 3-5 years of related experience; OR PhD and 0-3 years of related experience
  • Search Expertise: 5+ years of experience working with OpenSearch or Elasticsearch in a production environment at scale.
  • Distributed Systems: Strong understanding of distributed system concepts (Consensus algorithms, CAP theorem, replication, and sharding).
  • Programming: Proficiency in Java (core OpenSearch development) and/or Go/Python for automation and tooling.
  • Infrastructure: Extensive experience with Kubernetes (K8s) and container orchestration.
  • Cloud Providers: Hands-on experience deploying search workloads on AWS (EKS/AOSS), Azure (AKS), or Google Cloud (GKE).
  • Big Data Ecosystem: Familiarity with the Hadoop ecosystem or modern equivalents like Spark, Flink, and Hive is a major plus.

You might also have:

  • Experience with Lucene internals (segment merging, bitsets, and codecs).
  • Knowledge of Vector Database capabilities within OpenSearch for Generative AI (RAG) use cases.
  • History of contributing to open-source projects (Apache Software Foundation or OpenSearch Project)

Benefits & conditions

  • Generous PTO Policy
  • Support work life balance with Unplugged Days

  • Flexible WFH Policy
  • Mental & Physical Wellness programs
  • Phone and Internet Reimbursement program
  • Access to Continued Career Development
  • Comprehensive Benefits and Competitive Packages
  • Paid Volunteer Time
  • Employee Resource Groups

About the company

At Cloudera, we empower people to transform complex data into clear and actionable insights. With as much data under management as the hyperscalers, we’re the preferred data partner for the top companies in almost every industry. Powered by the relentless innovation of the open source community, Cloudera advances digital transformation for the world’s largest enterprises. The Data Platform Pillar is the bedrock of Cloudera’s technology, where we design and build the core components that let our customers store, manage, and process data with unmatched scalability, security, and performance. As a Senior Engineer on the Cloudera Context Search Team, you will be a key architect and contributor to the search heartbeat of the Cloudera Data Platform. You won’t just be “managing clusters”-you will be designing the high-performance, scalable, and secure search infrastructure that powers data discovery, observability, and analytics for the world’s largest enterprises.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on indeed.com

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

2:19 min

Core OpenSearch cluster architecture and terminology

Olena Kutsenko ¡ WWC 2022

1:09 min

Integrating speech models using the Twilio real-time SDK

Marius Obert Marius Obert ¡ Coffee With Developers

3:28 min

Defining big data and machine learning fundamentals

Ayon Roy ¡ LIVE

6:16 min

Scaling topic partitions and configurations to maximize throughput

Kirill Kulikov ¡ LIVE

2:10 min

Why organizations combine big data and machine learning

Ayon Roy ¡ LIVE

1:02 min

Provisioning a cluster with managed OpenSearch

Olena Kutsenko ¡ WWC 2022

Videos

See all

Related articles

See all