Senior Software Engineer - Indexing

Dormont Manufacturing Co
Amsterdam, Netherlands
about 1 month ago

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
5 years minimum
Working hours
Regular working hours

Tech stack

Automated Storage and Retrieval Systems C++ (Programming Language) Software as a Service Cloud Computing Databases Data Systems Distributed Computing Environment Distributed Systems MapReduce Data Streaming Web Crawlers Data Processing
+8 more
Database Optimization Apache Spark Indexer Backend Apache Flink Free and Open-Source Software Search Engines Data Pipelines

Job description

We are looking for a Senior Software Engineer to work on the indexing and data processing layer of a novel search engine tailored for agentic AI consumption.

In this role, you will focus on building systems that ingest, process, and organize massive volumes of data into efficient, queryable structures. You will work primarily on offline and nearline pipelines, ensuring that data is fresh, complete, and efficiently accessible by downstream retrieval systems.

You will operate in an environment where throughput, scalability, and correctness are critical, designing systems that can handle tens of gigabytes per second and continuously evolving datasets.

In this position, your responsibility will be to

  • Design, implement, and operate large-scale indexing systems and data pipelines
  • Build ingestion workflows for internal and external data sources (including web-scale crawling and structured feeds)
  • Develop and optimize indexing strategies for performance, freshness, and efficiency
  • Work on storage formats, compaction strategies, and update mechanisms
  • Ensure reliability and predictability under high-throughput conditions
  • Build well-tested components with clear contracts while allowing architectural flexibility
  • Define observability primitives (logs, metrics, data quality signals) across pipelines
  • Monitor throughput, resource usage, and cost, and optimize when needed
  • Collaborate with runtime and ML teams to support retrieval and ranking requirements
  • Enable safe experimentation on indexing strategies and data processing logic

Requirements

  • Have 5+ years of experience building production backend or data systems
  • Are strong in Go, C++, or Rust
  • Have experience with high-load systems (e.g. 10k+ RPS or 10+ GiB/sec throughput)
  • Have worked on databases, data planes, or large-scale data pipelines
  • Understand distributed systems fundamentals and failure modes
  • Have operated systems in production and handled real-world tradeoffs
  • Think in terms of systems and data flows rather than isolated components
  • Can make pragmatic decisions without compromising long-term system health
  • Collaborate effectively across infrastructure, ML, and product teams

Strong candidates may also have experience with:

  • Distributed data processing frameworks (Spark, Flink, MapReduce, Beam)
  • Content systems (web crawling, scraping, proxying, anti-bot infrastructure)
  • Ad tech, social networks, or large-scale content platforms
  • DBMS (OSS or SaaS) and cloud infrastructure
  • Open-source contributions
  • Competitive programming or CTFs (ICPC, IOI, etc.)
  • SHAD or similar advanced technical programs
  • Conference talks or technical publications

Benefits & conditions

  • Competitive salary and comprehensive benefits package.
  • Opportunities for professional growth within Nebius.
  • Flexible working arrangements.
  • A dynamic and collaborative work environment that values initiative and innovation.

J-18808-Ljbffr

Salarisomschrijving

€70000 - €90000 monthly

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.jobbird.com

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

1:52 min

Structuring and scaling the backend engineering team

Stefan Lingler Stefan Lingler +1 · Coffee With Developers

2:36 min

Analyzing limitations with PostgreSQL bitmap heap scans

Dharin Shah Dharin Shah · World Congress 2025

3:04 min

Database evolution and the funding behind vector databases

Erik Bamberg · LIVE

3:02 min

Assembling the tech stack for scalable indexing workflows

Chris Heilmann +2 · LIVE

1:12 min

Choosing TypeScript for complex backend applications

Maximilian Otto Maximilian Otto · World Congress 2024

3:36 min

Critical infrastructure and performance skills for modern developers

Andrew Holway · LIVE

Videos

See all

Related articles

See all