AIML Lead-Platform AI Acceleration

JPMorgan Chase & Co.
Glasgow, UK
15 days ago
Apply on www.themuse.com
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Working hours
Regular working hours

Tech stack

Java (Programming Language) Artificial Intelligence Systems Engineering Code Review Nvidia CUDA Software Debugging Distributed Computing Environment JSON Python (Programming Language) Performance Tuning Software Tools Systems Architecture
+7 more
Large Language Models Multi-Agent Systems Prompt Engineering Generative AI Web Filtering Data Analytics Natural Language Generation

Job description

As an Applied ML and Generative Lead within J.P.Morgan, you will operate as a hands-on engineering leader responsible for designing, building, and running production-grade ML and Generative AI services, while setting technical direction that scales across multiple workstreams. You will remain close to the code and architecture decisions, establish delivery and engineering standards, and ensure solutions meet enterprise expectations for security, stability, and operational rigor.

The ideal candidate brings a strong foundation in software engineering and AI/ML, along with proven experience leading the development and production operation of AI-enabled systems in secure, enterprise environments.

In this role, you will collaborate closely with Infrastructure Platforms AI teams to address priority use cases, design and build services, and promote best practices for scalable, resilient, and secure AI adoption. You will also mentor engineers, contribute to firmwide standards and thought leadership, and help ensure the organization stays at the forefront of AI engineering advancements., * Provide hands-on technical leadership by designing, developing, and deploying ML/LLM/GenAI solutions from concept through production, maintaining ownership for reliability and operability once deployed

  • Work closely with product managers, data scientists, ML engineers, and other stakeholders to understand requirements and prioritize use cases.
  • Develop secure, testable services and libraries that integrate LLMs, tool use, RAG, and agentic workflows.
  • Build end-to-end RAG/Agentic RAG pipelines: chunking and indexing, retrieval tuning, re-ranking, grounding checks.
  • Implement optimization strategies to fine-tune generative models for specific NLP use cases, ensuring high-quality outputs in summarization and text generation.
  • Mentor and uplift junior engineers through design reviews, code reviews, pairing, and coaching, raising engineering quality and delivery discipline across the team.
  • Implement monitoring mechanisms to track AI solution performance in real-time to ensure reliability and compliance.
  • Communicate AI/ML/LLM/GenAI capabilities and results to both technical and non-technical audiences.
  • Stay informed about the latest trends and advancements in the latest AI/ML/LLM/GenAI research, implement cutting-edge techniques, and leverage external APIs for enhanced functionality., Our professionals in our Corporate Functions cover a diverse range of areas from finance and risk to human resources and marketing. Our corporate teams are an essential part of our company, ensuring that we’re setting our businesses, clients, customers and employees up for success.

Requirements

  • Proven delivery of LLM-enabled applications using agentic patterns, including tool use, orchestration, guardrails, and structured outputs.
  • Hands-on experience building and operating MCP integrations reliably in production.
  • Hands-on experience on data-driven software/systems engineering experience delivering production services in secure, regulated environments.
  • Expertise in Python engineering skills, including production-grade design, testing, debugging, and performance tuning/optimization.
  • Advanced prompt engineering capabilities, including system prompts, few-shot prompting, tool/function calling, and schema-constrained outputs (e.g., JSON Schema).
  • Understanding of agentic AI system layers and concepts, such as context management, harness design, and loop engineering.
  • Experience building conversational AI solutions, including RAG, Agentic and Graph RAG techniques
  • Experience building and scaling AI/ML workloads using distributed training/serving frameworks (e.g., Ray) and GPU acceleration (e.g., CUDA) environments.
  • Proficiency with modern AI system architectures and patterns, including RAG, agentic RAG, and multi-agent systems.
  • Familiarity with LLM evaluation methodologies across quality, safety, and reliability, including guardrails, content filtering, and Responsible AI practices.
  • Proficiency in GenAI/agentic AI engineering practices, including data sensitivity, secure handling of inputs/outputs, and adherence to resiliency and security requirements
  • Demonstrated success driving adoption of enterprise-approved AI-assisted engineering tools (coding, review, testing, troubleshooting, * Financial Services industry experience
  • Understanding of Finops for LLMs
  • Good to have Java programming experience

About the company

We know our people are our strongest asset. You will never stop learning here, and we will learn with, invest in and support you along the way., J.P. Morgan is a global leader in financial services, providing strategic advice and products to the world’s most prominent corporations, governments, wealthy individuals and institutional investors. Our first-class business in a first-class way approach to serving clients drives everything we do. We strive to build trusted, long-term partnerships to help our clients achieve their business objectives.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.themuse.com
Prepare application

Good distractions

Talks and stories from around this role β€” technically off-topic, practically not.

3:47 min

Exploring JSON, CBOR, and JOSE for data serialization

Aaron Russell Β· LIVE

3:39 min

Addressing code review surrender and process exploitation

Laura Tacho Laura Tacho Β· World Congress 2026 Europe

6:21 min

Previewing upcoming hardware acceleration capabilities for Python environments

Chris Heilmann Chris Heilmann +2 Β· LIVE

4:04 min

Defining agentic AI and the tool execution architecture

Rijk van Zanten Rijk van Zanten Β· Europe 2026 Virtual

1:37 min

Accelerating compute with focused developer tools

Julia Koch Julia Koch +1 Β· World Congress 2026 Europe

56 sec

The hidden costs of delayed peer code reviews

Tim Gilboy Tim Gilboy

Videos

See all

Related articles

See all