AI Platform Engineer
- Discuss this with your agent
- Open in Claude
- Open in ChatGPT
Prepare application
- Draft this with your agent
- Open in Claude
- Open in ChatGPT
Role details
Tech stack
+7 more
Job description
This role turns AI delivery from one-off, bespoke builds into reusable enterprise capability. You will build and mature the shared engineering foundations that every OneMain AI product depends on starting with OneAdvisor and extending to future assistants, agents, and knowledge systems. A key part of the mandate is transitioning enhancements and operations from vendor-led delivery to an internal OneMain team.
What You’ll Build
Retrieval-Augmented Generation (RAG) and retrieval pipelines as reusable services. A centralized embeddings service and vector index management, including chunking strategy and re-indexing/freshness jobs. A memory bank capability persistent, identity-scoped agent/assistant memory with a defined lifecycle for what is stored, summarized, decayed, and deleted. Context assembly logic that combines retrieval, memory, and system data into prompts in a consistent, reusable way. Content ingestion and processing pipelines for enterprise knowledge sources. Prompt and configuration management systems that support versioning and reuse. A shared PII detection and redaction library so sensitive data is scrubbed from model inputs, outputs, and logs. A platform SDK and golden-path templates that let teams stand up a new compliant AI product with telemetry, evaluation hooks, access, and deployment already wired in. Evaluation hooks that let quality checks plug into every AI product. Telemetry integration for usage, performance, and quality signals. Identity and access integration aligned to enterprise standards. Standardized deployment patterns and support practices for AI products.
Why This Role Matters
Requirements
- 7+ years of relevant experience in software or platform engineering, with recent hands-on work building AI/ML or data-intensive systems.
- Practical experience with RAG, vector search, embeddings, and retrieval pipeline design.
- Familiarity with agent memory, context management, and PII-handling patterns hands-on production experience is a plus but not required; these are areas the role will help build.
- Strong software engineering fundamentals: APIs, testing, CI/CD, and observability.
- Experience with prompt/configuration management and integrating LLMs into production systems.
- Proficiency in Python and/or another primary backend language.
- Fluent in AWS, with familiarity with other major cloud platforms (Azure, Google Cloud Platform).
- Experience operating and supporting production services, not just prototyping., * Experience building reusable, multi-tenant internal platform services.
- Familiarity with LLM orchestration frameworks.
- Experience transitioning vendor-built systems to internal ownership and operations.
- Background in regulated data environments
Apply for this position
This job is hosted externally. Click below to view the full posting and apply.
Prepare application
- Draft this with your agent
- Open in Claude
- Open in ChatGPT
Good distractions
Talks and stories from around this role — technically off-topic, practically not.
Moments
Explore playlistsVideos
See allRelated articles
See all
Navigating the AI Shift
MLOps And AI Driven Development
MLOps – What’s the deal behind it?
What Are Large Language Models?