LLM Engineer (INTL - Europe)

Insight Global
Boston, MA, United States
about 1 month ago

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Experienced
Experience required
2 years minimum
Working hours
Regular working hours
Job source

Tech stack

Java (Programming Language) Application Programming Interfaces (APIs) Artificial Intelligence Automated Storage and Retrieval Systems Cloud Computing Software Debugging Distributed Systems Elasticsearch Python (Programming Language) Machine Learning Software Engineering Enterprise Software Applications
+6 more
Large Language Models Multi-Agent Systems Indexer Backend Build Management Restful APIs

Job description

Insight Global is hiring an LLM Engineer to support one of our large consulting clients. This role will be supporting cutting-edge projects within the AI space, and sitting remotely in Europe.

Day-to-Day

-Build and deploy LLM-powered features across backend systems and APIs within a production application

-Design and implement RAG pipelines including ingestion, indexing, embeddings, and retrieval logic

-Develop and optimize search and retrieval systems using Elasticsearch and vector-based architectures

-Leverage AI-assisted and agentic coding tools to accelerate engineering output and automate workflows

-Integrate LLM APIs and GenAI frameworks into scalable enterprise applications

-Collaborate cross-functionally with product and business teams to deliver AI-driven solutions

-Communicate feature progress, tradeoffs, and outcomes to both technical and business stakeholders

We are a company committed to creating diverse and inclusive environments where people can bring their full, authentic selves to work every day. We are an equal opportunity/affirmative action employer that believes everyone matters. Qualified candidates will receive consideration for employment regardless of their race, color, ethnicity, religion, sex (including pregnancy), sexual orientation, gender identity and expression, marital status, national origin, ancestry, genetic factors, age, disability, protected veteran status, military or uniformed service member status, or any other status or characteristic protected by applicable laws, regulations, and ordinances. If you need assistance and/or a reasonable accommodation due to a disability during the application or recruiting process, please send a request to HR@insightglobal.com.To learn more about how we collect, keep, and process your private information, please review Insight Global’s Workforce Privacy Policy: https://insightglobal.com/workforce-privacy-policy/.

Requirements

8+ years of software engineering experience with 2+ years working directly with LLM systems

-Hands-on experience building and deploying RAG systems end-to-end (retrieval, embeddings, evaluation)

-Strong experience with LLM APIs (OpenAI, Anthropic, or similar)

-Experience building scalable backend services (Python + REST APIs)

-Strong foundation in NLP/ML concepts as applied to LLM systems

-Experience working with cloud platforms and distributed architectures

-Experience with Elasticsearch and strong proficiency in Java OR ability to develop in Java using agentic coding tools

-Hands-on experience with agent-based or multi-agent systems and orchestration

-Strong problem-solving, systems thinking, and debugging capabilities -Experience with LLM observability, guardrails, and evaluation frameworks

-Familiarity with frameworks like LangChain or LlamaIndex

-Experience deploying GenAI solutions at scale beyond proof-of-concept

-Experience with vector databases (Pinecone, FAISS, etc.)

-Exposure to building AI-first engineering workflows and autonomous coding systems

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on dejobs.org

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

5:30 min

Building components of a real-world LLM lifecycle

Maxim Salnikov Maxim Salnikov · LIVE

1:52 min

Structuring and scaling the backend engineering team

Stefan Lingler Stefan Lingler +1 · Coffee With Developers

2:36 min

Analyzing limitations with PostgreSQL bitmap heap scans

Dharin Shah Dharin Shah · WWC 2025

2:59 min

The danger of adopting default REST APIs

Stefan Priebsch · WWC 2021

1:59 min

Building culturally aware LLMs for global audiences

Werner Vogels Werner Vogels +1 · WWC Europe 2026

1:12 min

Choosing TypeScript for complex backend applications

Maximilian Otto Maximilian Otto · WWC 2024

Videos

See all

Related articles

See all