RAG Engineer

Sumeru Global Technologies Private Limited
Jersey City, NJ, United States
2 months ago
Apply on indeed.com
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Experienced
Experience required
2 years minimum
Compensation
$120,000.0 - $200,000.0
Working hours
Regular working hours
Job source

Tech stack

Java (Programming Language) .NET Framework Application Programming Interfaces (APIs) Artificial Intelligence Audit Trail Microsoft Azure Cloud Storage Encodings Databases Continuous Integration Information Engineering Data Governance
+13 more
Data Security Elasticsearch Graph Database Python (Programming Language) Node.Js Parsing Role-Based Access Control Standard Sql Search Technologies Management of Software Versions Indexer Backend Search Engines

Job description

· Build document ingestion, parsing, chunking, metadata extraction, embedding, indexing, and versioning pipelines.

· Implement RAG services using vector search, keyword search, hybrid retrieval, metadata filters, reranking, and citations.

· Integrate with Azure AI Search, pgvector, Pinecone, Weaviate, Qdrant, Milvus, Elasticsearch, OpenSearch, or similar platforms.

· Build APIs for retrieval, source traceability, ingestion status, feedback capture, and citation support. Implement permission-aware retrieval, RBAC, tenant isolation, audit logging, PII handling, and secure data access.

· Support retrieval quality evaluation, grounding, hallucination reduction, client demos, UAT, and production readiness.

Requirements

Do you have experience in Search engines?, · Strong backend/data engineering using Python, Java, Node.js, or .NET. Strong SQL and database experience.

· Minimum 2 years hands-on AI/RAG implementation experience.

· Experience with embeddings, vector search, hybrid search, document ingestion, and RAG pipelines. Experience with Azure AI Search, pgvector, Pinecone, Weaviate, Qdrant, Milvus, Elasticsearch, OpenSearch, or equivalent.

· Understanding of chunking, retrieval tuning, reranking, context assembly, citations, and retrieval evaluation. BFSI, banking, insurance, capital markets, payments, or regulated enterprise experience. Client implementation experience.

· Preferred Skills

· Azure Blob Storage, Azure AI Search, Azure Document Intelligence, Azure OpenAI, or similar Azure services.

· Experience with data governance, document lineage, metadata models, knowledge graphs, or ontology-driven retrieval.

· Experience with CI/CD, containers, observability, and secure cloud deployment

Benefits & conditions

Pulled from the full job description

  • Parental leave
  • Retirement plan
  • 401(k) matching
  • Paid time off
  • Dental insurance
  • Life insurance
  • Employee assistance program, * 401(k) matching
  • Dental insurance
  • Employee assistance program
  • Flexible schedule
  • Life insurance
  • Paid time off
  • Parental leave
  • Retirement plan

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on indeed.com
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

1:50 min

Simplifying generative AI deployments using the RagStack opinionated framework

David Leconte David Leconte +1 · World Congress 2024

1:52 min

Structuring and scaling the backend engineering team

Stefan Lingler Stefan Lingler +1 · Coffee With Developers

2:36 min

Analyzing limitations with PostgreSQL bitmap heap scans

Dharin Shah Dharin Shah · World Congress 2025

2:56 min

Open-sourcing a complex parsing library for game data

Johan Hutting Johan Hutting · World Congress 2024

1:27 min

Implementing semantic RAG with fallback for complex queries

Piotr Menclewicz Piotr Menclewicz · Europe 2026 Virtual

1:12 min

Choosing TypeScript for complex backend applications

Maximilian Otto Maximilian Otto · World Congress 2024

Videos

See all

Related articles

See all