AI Engineer (LLM Infrastructure)

Yoursafe
Amsterdam, Netherlands
2 months ago
Apply on indeed.com
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Working hours
Regular working hours
Job source

Tech stack

Training Data Application Programming Interfaces (APIs) Artificial Intelligence Nvidia CUDA Databases Linux Memory Management Information Retrieval Python (Programming Language) Open Source Technology Search Technologies AI Infrastructure
+10 more
Data Ingestion Pytorch Large Language Models Bare Metal Data Analytics Hardware Acceleration Data Management Machine Learning Operations ZFS File System TensorRT

Job description

Behind our scalable Open Issuing platform and our push into 100 target countries, sits a highly advanced, bare-metal AI infrastructure. We are building a high-density, on-premise AI and automation environment consisting of cutting-edge heavy compute (Nvidia H200), high-performance enterprise storage, and a fleet of automated agents., We are seeking a senior AI Engineer to take ownership of our LLM inference infrastructure. You will own the hardware and software running our AI models, ensuring that our internal automation agents and our growth teams have zero-latency, highly optimized access to state-of-the-art open-source LLMs. You will collaborate closely with our CEO, COO, Product Manager and Head of Growth to engineer the AI-driven pipelines required for large-scale content creation, localization and personalised agentic support.

This is a senior, hands-on role. You are expected to manage hardware, deploy models, optimize inference speeds, and scale what works.

Your responsibilities:

LLM Deployment & Optimization

¡ Own and manage our heavy-compute AI hardware, specifically optimizing workloads for our Nvidia HGX H200 infrastructure.

¡ Deploy, fine-tune, and maintain open-source LLMs, ensuring maximum throughput and minimal latency.

¡ Manage inference engines (e.g., vLLM, TensorRT-LLM) and handle dynamic GPU memory allocation for hundreds of concurrent agent requests.

AI Automation & Agent Infrastructure

· Provide a flawless, millisecond-response API layer for our “OpenClaw” agent farm (a fleet of 200+ bare-metal Apple Silicon nodes).

¡ Monitor model performance, detect hallucinations, and build synthetic training data pipelines to continuously improve agent accuracy.

¡ Design, scale, and maintain a high-performance, on-premise RAG service and vector database (e.g., Qdrant, Milvus, Milvus/Chroma) to seamlessly serve internal documentation to our LLMs.

¡ Build robust data ingestion and embedding pipelines to ensure internal knowledge bases and documents are updated in real-time for the RAG service.

Cross-Functional AI Tooling

¡ Work tightly with Product, Operations, and the Growth team to ensure alignment.

· Provide the technical foundation for the Growth team’s AI-assisted content creation, verification, and contextual validation pipelines.

¡ Build repeatable AI engines that can guarantee linguistic, cultural, and regulatory correctness across dozens of markets simultaneously.

Requirements

Do you have experience in Python?, Do you have a Master’s degree?, · A senior AI/MLOps engineer with proven experience scaling self-hosted LLM infrastructure in a production environment.

¡ Experienced with semantic search, vector databases, and information retrieval techniques (RAG) at scale.

¡ Deeply experienced with Python, PyTorch, CUDA, and modern inference serving frameworks.

¡ Experienced in using AI as a production and verification tool, not a gimmick.

¡ Comfortable working closely with networking and storage architects to eliminate I/O bottlenecks in a ZFS/Linux ecosystem.

¡ Highly structured, data-driven, and execution-focused.

¡ Motivated by building systems that scale, not campaigns that win awards.

Benefits & conditions

¡ A senior tech role with direct ownership over a state-of-the-art AI hardware stack.

¡ A fast-scaling international fintech with real infrastructure, licenses, and products.

¡ Competitive salary and employment conditions.

· A modern office in Amsterdam Houthavens overlooking ‘Het IJ’.

¡ Daily healthy lunches prepared by our in-house chef.

¡ A culture that values execution, ownership, and long-term thinking.

Job Types: Full-time, Permanent

About the company

At Yoursafe, we believe that everyone deserves access to safe and easy financial services - wherever they are. Our mission is to build financial tools that empower people, especially those who are new to a country or outside the traditional banking system. We combine solid financial expertise with smart technology to make everyday money management simple, secure, and fair.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on indeed.com
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

2:08 min

Essential engineering roles in the generative AI space

Mary Grygleski Mary Grygleski ¡ LIVE

4:52 min

Essential phases in building and refining language models

Anshul Jindal Anshul Jindal +1 ¡ World Congress 2025

2:35 min

Preventing remote code execution in PyTorch models

Balåzs Kiss ¡ World Congress 2023

52 sec

Running persistent Linux environments directly on Windows

Ben Breard Ben Breard ¡ World Congress 2025

1:24 min

Comprehensive AI infrastructure stacks at the Linux Foundation

Matt White Matt White ¡ World Congress 2025

2:32 min

Core libraries driving inference engines and multi-GPU networking

Adolf Hohl Adolf Hohl ¡ World Congress 2024

Videos

See all

Related articles

See all