Senior AI Engineer

Newfront Insurance Holdings, Inc.
United States
3 months ago

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
5 years minimum
Compensation
$160,000.0 - $250,000.0
Working hours
Regular working hours
Job source

Tech stack

Testing (Software) Application Programming Interfaces (APIs) Artificial Intelligence User Authentication Business Software Cloud Computing Code Review Computer Programming Continuous Integration Software Debugging DevOps Python (Programming Language)
+19 more
Machine Learning Node.Js Software Product Management Software Engineering TypeScript Large Language Models Multi-Agent Systems Indexer Backend Build Management Containerization AI Platforms Scikit Learn Information Technology Low Latency Integration Frameworks Machine Learning Operations TensorRT Api Design

Job description

As a Senior AI Engineer at Newfront, you will work across both our AI platform and the AI products built on top of it. You will design and build the core AI platform - the agent runtime, RAG and document understanding pipelines, connector framework, model routing layer, evaluation and observability backbone, and on-premise model hosting - and you will use that same platform to ship AI products that brokers, underwriters, and operations rely on every day. You will partner closely with product engineering, brokering and operations leaders, and security and compliance partners to make sure both the platform and the products reflect how the business actually works.

This position is a salary, exempt, and full-time position. This is a US-remote or hybrid role with the option to work from any of Newfront’s office locations. #LI-Remote

What You’ll Be Responsible For:

  • Build and scale the agent runtime at Newfront - tool use, planning, memory, multi-agent orchestration, human-in-the-loop handoff, and the SDKs product teams use to compose agents for brokering, underwriting, and client service workflows.
  • Design and own RAG and document understanding pipelines for complex insurance artifacts - submissions, policies, endorsements, loss runs, SOVs - including ingestion, chunking, indexing, retrieval, structured extraction, and grounding.
  • Build the connector framework that lets AI agents and pipelines reach into the systems brokers actually use (AMS, carrier portals, email, document stores, internal services), with first-class auth, rate-limiting, schema discovery, and auditability.
  • Establish the evaluation and observability backbone for AI at Newfront: offline eval harnesses, regression suites, hallucination and grounding checks, online quality, latency, and cost telemetry, and the dashboards every team uses to measure AI features against compliance and performance targets.
  • Own model routing and the model gateway - selecting between hosted frontier models and on-premise / self-hosted models per use case, balancing quality, latency, cost, and data-residency constraints.
  • Stand up and operate on-premise model hosting where regulatory, contractual, or data-sensitivity requirements make hosted APIs unsuitable: GPU capacity planning, inference serving, quantization and optimization, isolation, and lifecycle management.
  • Ship AI product features end-to-end on top of the platform - from problem discovery with brokering, underwriting, and operations partners through technical design, implementation, and the user-facing surfaces that make AI capabilities intuitive for non-technical users.
  • Mentor engineers on production AI systems, review designs, and help establish best practices for building reliable AI in financial services.
  • Partner with security, privacy, and compliance to bake controls - auth, audit, PII handling, retention, model risk management - into the platform rather than leaving them to each product team to rediscover.

Requirements

Do you have experience in Software testing?, Do you have a Bachelor’s degree?, * BS, MS or PhD in computer science, or related field, or equivalent work experience

  • 5+ years of professional software engineering experience, with a strong general software development background - building, shipping, and operating production services, not just notebooks or prototypes. We are looking for an engineer first, who has gone deep on AI, rather than someone whose experience is exclusively in ML research.
  • Solid fundamentals in API design, data modeling, testing, debugging production systems, code review, and collaborating in a team codebase.
  • Strong programming skills in TypeScript, including experience with Node.js or another TypeScript backend framework in production.
  • Experience with modern development and deployment practices (e.g., containerization, CI/CD, infrastructure-as-code, production observability).
  • A track record of leading AI/ML projects end-to-end, including API design, production operations, and long-term maintenance.
  • Experience designing systems for reliability, cost, and scale in production.
  • Passion for staying up-to-date with the latest advancements in AI/ML and applying them to real-world problems.
  • Strong problem-solving skills and the ability to take a pragmatic and efficient approach to tackling challenges.
  • Excellent collaboration and communication skills, with the ability to partner with product teams and effectively communicate complex technical concepts to non-technical stakeholders.

Preferred Knowledge, Skills, and Abilities:

  • Working knowledge of Python and/or Go - we use Python for ML/AI tooling and Go in parts of our backend, and you will read and occasionally write both.
  • Experience deploying or leveraging machine learning models and Large Language Models (LLMs) to power business applications at scale.
  • Hands-on experience building agent frameworks, tool-use runtimes, RAG systems, connector / integration frameworks, or evaluation harnesses for LLM-based applications.
  • Experience self-hosting or fine-tuning open-weight LLMs (e.g., GPU inference serving with vLLM/TGI/TensorRT-LLM, quantization, LoRA/PEFT, on-prem deployment).
  • Experience building model gateways or routing layers that span multiple model providers and self-hosted models.
  • Knowledge of state-of-the-art LLM techniques, models, and vendors, including trade-offs across providers.
  • Familiarity with LLM and related frameworks, including extracting structured data from unstructured text.
  • Experience with popular AI/ML libraries and frameworks.
  • Familiarity with DevOps practices, cloud infrastructure, authorization, authentication, and search infrastructure.
  • Experience with model risk management, AI governance, or building AI systems in regulated industries (financial services, healthcare, insurance).
  • Understanding of machine learning essentials and the ability to collaborate effectively with data scientists.

Benefits & conditions

3.53.5 out of 5 stars United States Remote $160,000 - $250,000 a year - Full-time, The pay range for this position in California, Washington, Colorado and New York at commencement of employment is expected to be between $160,000 and $250,000/yr; however, base pay offered may vary depending on multiple individualized factors, including market location, job-related knowledge, skills, and experience. The total compensation package for this position may also include other elements, including a bonus, restricted stock units, and discretionary awards in addition to a full range of medical, financial, and/or other benefits (including 401(k) eligibility and various paid time off benefits, such as vacation, sick time, and parental leave), dependent on the position offered. Details of participation in these benefit plans will be provided if an employee receives an offer of employment. If hired, the employee will be in an “at-will position” and the Company reserves the right to modify base salary (as well as any other discretionary payment or compensation program) at any time, including for reasons related to individual performance, Company or individual department/team performance, and market factors., If you require reasonable accommodations throughout the application or interview process, please contact us at careers@newfront.com. For information regarding how Newfront collects and uses personal information, please review our Privacy Policy.

Compensation Range: $160K - $250K

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on indeed.com

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

2:14 min

Exploring internal AI product initiatives and global engineering roles

Maria Apazoglou ¡ Coffee With Developers

2:17 min

Mapping the maturity roadmap for scaled devops adoption

Dominik Krichbaum Dominik Krichbaum ¡ WWC Europe 2026

1:52 min

Structuring and scaling the backend engineering team

Stefan Lingler Stefan Lingler +1 ¡ Coffee With Developers

45 sec

Working securely with Node.js path application programming interfaces

Sonya Moisset ¡ WWC 2023

2:08 min

Essential engineering roles in the generative AI space

Mary Grygleski Mary Grygleski ¡ LIVE

3:18 min

Scaling global network engineering through DevOps culture

Stuart Clark ¡ LIVE

Videos

See all

Related articles

See all