AI Engineer

Amaris
United States
4 days ago
Apply on careers.mantu.com
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Starter
Experience required
1 year minimum
Working hours
Regular working hours

Tech stack

Application Programming Interfaces (APIs) Artificial Intelligence Code Review Encodings Computer Programming Continuous Integration Cursor (Graphical User Interface Elements) Python (Programming Language) Machine Learning Language Modeling Open Source Technology Azure Machine Learning
+14 more
Data Streaming Autoscaling Large Language Models Multi-Agent Systems Git AI Platforms Kubernetes Bug Reporting Information Technology Low Latency Machine Learning Operations Virtual Agents Software Version Control Docker

Job description

  • Architect and build advanced AI agentic systems end-to-end - including planning, memory, tool use, multi-agent delegation, evaluation loops, and guardrails - always choosing the right abstraction for the real problem, not just what’s trending
  • Design and implement LLM-powered applications in production, defining prompt and context strategies, tool interfaces, retrieval and reranking logic, structured outputs, streaming, and evaluation - across text and multimodal inputs (vision, documents, audio)
  • Own the evaluation strategy for AI/agent systems: build offline evaluation datasets, design online LLM-as-judge loops, and set up regression harnesses that focus on metrics that truly impact users, not just dashboards
  • Optimize and “squeeze” AI systems for performance and cost: implement prompt caching, batching, speculative decoding, model routing, token budget management, and latency targets; monitor and understand P50/P99 behaviour and continuously improve it
  • Lead productionization of AI workloads: design APIs and inference services, build RAG/embedding pipelines, containerize workloads, and handle CPU/GPU deployment while optimizing latency, throughput, reliability, and cost
  • Contribute upstream to the AI ecosystem: read SDK source code when documentation is limited, open PRs to open-source agent frameworks, and write clear bug reports for vendors when orchestration services misbehave
  • Operate AI systems in production with strong AI platform & MLOps practices - model/version management, CI/CD, evaluation gates, observability, autoscaling, rollback, and failure handling, using tools such as Docker, Kubernetes/AKS, Azure AI services, vLLM or NVIDIA Triton
  • Work hands-on with multimodal models (vision-language, document AI, audio) from ingest to grounded output, including OCR, layout understanding, tables, and speech processing
  • Collaborate closely with Product Owners, engineers, and stakeholders to understand business needs and translate them into robust agent architectures, LLM workflows, and technical solutions
  • Mentor other engineers and set the technical bar through design reviews, code reviews, and technical writing that shape how the team thinks about agents and AI systems, At Amaris, we strive to provide our candidates with the best possible recruitment experience. We like to get to know our candidates, challenge them, and be able to give them proper feedback as quickly as possible. Here’s what our recruitment process looks like:

Brief Call: Our process typically begins with a brief virtual/phone conversation to get to know you! The objective? Learn about you, understand your motivations, and make sure we have the right job for you!

Interviews (the average number of interviews is 3 - the number may vary depending on the level of seniority required for the position). During the interviews, you will meet people from our team: your line manager of course, but also other people related to your future role. We will talk in depth about you, your experience, and skills, but also about the position and what will be expected of you. Of course, you will also get to know Amaris: our culture, our roots, our teams, and your career opportunities!

Case study: Depending on the position, we may ask you to take a test. This could be a role play, a technical assessment, a problem-solving scenario, etc.

As you know, every person is different and so is every role in a company. That is why we have to adapt accordingly, and the process may differ slightly at times. However, please know that we always put ourselves in the candidate’s shoes to ensure they have the best possible experience. We look forward to meeting you!

Requirements

  • Bachelor’s degree in Artificial Intelligence, Computer Science, or a related field - or equivalent practical experience
  • 1+ years of experience as an AI / AI Agent Engineer, working on real-world AI or LLM-based applications
  • Solid machine learning fundamentals: able to clearly explain transformers (attention, positional encoding, KV cache, tokenisation, sampling) and familiar with core research papers beyond just the abstracts
  • Deep LLM application experience: you have built multiple production systems on top of frontier models (e.g., Anthropic, OpenAI, Gemini, open-weight models) and understand practical edge cases such as tool-use stability, structured-output failure modes, long-context degradation, prompt-injection defence, and cost control
  • Proven agent systems depth: experience building systems with real agent behaviour - planning, memory, tool orchestration, multi-step execution, and error recovery - beyond simple single-prompt loops; multi-agent coordination (delegation, sub-agent protocols, MCP-style tool servers) is a strong plus
  • Hands-on multimodal experience with vision-language models, document AI (OCR, layout, tables), or audio, including end-to-end pipelines from data ingest to grounded outputs
  • Strong engineering craft with Python at a senior level - async programming, typing, testing, packaging, observability - and the ability to navigate and contribute to large codebases with clean Git practices
  • Solid production AI engineering background: taking ML, LLM, embedding, vision, or multimodal models from prototype to production, designing APIs and inference services, building RAG/embedding pipelines, and deploying/optimizing workloads on CPU/GPU
  • Practical AI platform & MLOps experience: operating AI workloads in production with model/version management, CI/CD, evaluation gates, observability, autoscaling, rollback, and failure handling; experience with Docker, Kubernetes/AKS, Azure AI services, GPU inference, vLLM, or NVIDIA Triton is a strong plus
  • Fluent with modern coding agents (Claude Code, Cursor, Copilot, or equivalents) and able to design prompts, context windows, and tool boundaries to use them effectively while understanding their limitations
  • Fluent English communication skills (spoken and written) - able to explain complex technical concepts clearly and collaborate with international stakeholders
  • You are curious, rigorous, and proactive - comfortable working at the frontier of AI, collaborating with both technical and non-technical team members, and continuously raising the bar for agentic systems and AI engineering

Benefits & conditions

  • Competitive salary and 13th-month salary
  • 14+ annual leave days per year
  • Premium healthcare insurance starting from your probation period
  • Regular project reviews and yearly performance appraisal
  • Annual company trip
  • Team-building activities: team lunch/dinner, events and celebrations, sports clubs (football, basketball, badminton, pickleball)
  • International working environment
  • Tailor-made career path and clear growth opportunities
  • Technical workshops and training courses (internal & external)
  • Mobility opportunities to work on-site in our offices in 60+ countries

About the company

Amaris Consulting is an independent technology consulting firm providing guidance and solutions to businesses. With more than 1,000 clients across the globe, we have been rolling out solutions in major projects for over a decade - this is made possible by an international team of 7,600 people spread across 5 continents and more than 60 countries. Our solutions focus on four different Business Lines: Information System & Digital, Telecom, Life Sciences and Engineering. We’re focused on building and nurturing a top talent community where all our team members can achieve their full potential. Amaris is your steppingstone to cross rivers of change, meet challenges and achieve all your projects with success.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on careers.mantu.com
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

1:24 min

Building client-facing AI agents for engineering teams

Alfonso Graziano Alfonso Graziano · Coffee With Developers

6:21 min

Investigating push inefficiencies with upstream Git experts

Jonathan Creamer · Coffee With Developers

2:07 min

Inspecting default bridge architectures and custom Docker networks

Oliver Seitz Oliver Seitz · World Congress 2025

3:17 min

Optimizing character encoding with Kim variable byte encoding

Douglas Crockford Douglas Crockford · World Congress 2024

2:08 min

Essential engineering roles in the generative AI space

Mary Grygleski Mary Grygleski · LIVE

56 sec

Favorite git commands and the importance of patch commits

Eileen Uchitelle Eileen Uchitelle +1 · Coffee With Developers

Videos

See all

Related articles

See all