Senior AI/ML Engineer

Everforth CyberCoders
Pittsburgh, PA, United States
2 months ago

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
3 years minimum
Working hours
Regular working hours
Job source

Tech stack

A/B Testing Artificial Intelligence Audit Trail Microsoft Azure Software as a Service Continuous Integration Python (Programming Language) Monte Carlo Methods Signal Processing Software Engineering TypeScript Large Language Models
+6 more
Multi-Agent Systems Prompt Engineering Kotlin AI Platforms Free and Open-Source Software Stream Processing

Job description

AI focus: LLMs, agent orchestration, RAG architectures, evaluation pipelines (Claude/Bedrock integration), You will design, build, and operate the AI components that make our platform intelligent and trustworthy:

  • Signal & anomaly detection - Build statistical and ML detectors that separate noise from real problems in CDC event streams and external integrations.
  • Insight synthesis engine - Ship an LLM-powered correlation engine that returns root causes, confidence scores, and evidence chains, not just alerts.
  • Planning rules compiler - Translate natural-language planning rules into structured parameters for a deterministic Monte Carlo scheduling engine.
  • Evaluation & testing frameworks - Create regression suites, A/B testing, and confidence-calibration pipelines so model changes are safe and measurable.
  • MCP tool definitions - Define LLM-ready tool specs (Item Store queries, capacity lookups, scenario simulations) for runtime tool use in a hub-and-spoke agent architecture.

Requirements

Do you have experience in Tooling?, * Proven track record shipping LLM-powered features or products (not prototypes) that real users rely on.

  • Hands-on experience orchestrating agents (multi-step reasoning, tool use, autonomous action with guardrails) - LangChain, LlamaIndex, AutoGen, CrewAI, or equivalent.
  • Deep LLM engineering fundamentals: prompt design, RAG architectures, function-calling/tool use, context management, and evaluation-driven development.
  • Production engineering discipline: tests, CI/CD, observability, and reliability for production AI systems.
  • Experience with event-driven or streaming systems (CDC, real-time pipelines).
  • 5+ years software engineering, with 3+ years focused on AI/ML in production.
  • Comfortable working embedded in a product team - collaborating daily with domain engineers, product managers, and designers.

Preferred experience

  • Experience with AWS Bedrock, Azure OpenAI, or GCP Vertex AI (we run on Bedrock with Claude today).
  • Familiarity with MCP (Model Context Protocol) or similar agentic frameworks.
  • Background in anomaly detection, time-series analysis, or statistical signal processing.
  • Experience building confidence scoring / calibration systems for AI outputs.
  • Proficiency in Kotlin or TypeScript in addition to Python (our product platform is Kotlin/TypeScript; AI platform is Python).
  • History of absorbing work from external partners and improving inherited architectures.

Nice to have

  • Monte Carlo simulation, optimization, or scheduling systems.
  • Domain experience in portfolio, project, or resource planning.
  • Enterprise SaaS experience (multi-tenancy, compliance, audit trails).
  • Open-source contributions to AI/ML tooling or frameworks.

Benefits & conditions

Pulled from the full job description

  • Health insurance
  • Vision insurance
  • Dental insurance

About the company

Founded in 2006, we’re one of the largest app vendors (300+ employees) supporting 30,000+ customers worldwide - including Amazon, Disney, Dell, PayPal, and Hulu. In fact, one in three Fortune 500 companies relies on our products!!

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on indeed.com

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

3:09 min

Understanding Kotlin Multiplatform and its compiler targets

Petar Marijanović · LIVE

1:00 min

Misconceptions about TypeScript safety capabilities

Simone Sanfratello · JS Congress

54 sec

Generating multiple hook options for outreach A/B testing

Leandro Gomes da Silva Leandro Gomes da Silva · WWC 2025

2:14 min

Exploring internal AI product initiatives and global engineering roles

Maria Apazoglou · Coffee With Developers

2:18 min

Recommended resources and frameworks for learning Kotlin natively

Iris Hunkeler · LIVE

3:22 min

Evaluating advanced artificial intelligence platforms for daily recruitment

Rudi Bauer Rudi Bauer +1 · Cappuccino with HR

Videos

See all

Related articles

See all