Machine Learning Engineer

Committee Of Interns And Residents (inc)
United States
1 day ago
Apply on startup.jobs
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
5 years minimum
Working hours
Regular working hours
Job source

Tech stack

Airflow Computer Vision Continuous Integration Github Python (Programming Language) Machine Learning Software Engineering Reliability of Systems Backend Containerization Kubernetes Apache Flink
+6 more
Deployment Automation Machine Learning Operations Api Design Terraform Docker Microservices

Job description

  • Collaborate closely with engineering, product, and data science teams to understand business challenges and the potential for machine learning and AI solutions.
  • Develop tools and automate manual processes to improve operational efficiency, accelerate experimentation velocity, and minimize human error.
  • Build, integrate, and monitor the end-to-end lifecycles of large-scale, distributed machine learning systems.
  • Investigate model performance and identify data quality and performance issues.
  • Enhance the ML pipeline for our forecasting platform, managing weekly automated model retraining and deployment across a range of production models
  • Elevate the team’s technical capabilities in MLOps best practices, automation, and production ML systems.

Requirements

We’re looking for a senior machine learning engineer who would thrive at the intersection of MLOps, infrastructure, and production forecasting systems. You excel at building robust, automated ML pipelines that enable data scientists to iterate quickly while maintaining the reliability that our healthcare customers depend on.

You are a systems thinker who flourishes in a startup environment, working across the stack to productionize ML models, automate deployment pipelines, and establish the infrastructure that makes our forecasting platform scalable and maintainable. As a critical member of our Forecasting team, you’ll work on production ML lifecycle - from training automation to serving infrastructure - and have the unique opportunity to shape the technical foundation of Apella’s ML platform while growing our team’s capabilities., * 5+ years of experience building and maintaining production ML systems, with deep expertise in MLOps, deployment automation, and model serving infrastructure

  • Strong software engineering skills with proficiency in Python, containerization (Docker/Kubernetes), CI/CD systems (GitHub Actions, ArgoCD), and infrastructure-as-code (Terraform, Helm)
  • Production ML deployment experience including model training orchestration (Dagster, Airflow, or similar), automated retraining pipelines, and A/B testing/variant management
  • Systems design expertise with experience building scalable microservices, API design, and managing complex service dependencies
  • Ownership mentality with a track record of driving projects from concept to production, maintaining them over time, and continuously improving system reliability
  • Excellent collaboration skills working with data scientists to productionize research, with backend teams on API integration, and with product org to meet customer needs
  • Passion for writing tested, maintainable, well-documented code that enables team velocity, * Experience working in healthcare or other regulated industries.
  • Experience with Forecasting / Time Series algorithms
  • Experience with Computer Vision
  • Experience with DAG frameworks, Flink

Benefits & conditions

  • Competitive salary and stock options
  • Flexible vacation policy and a culture that values time for rest and recharging
  • Remote-first work environment with unique virtual and in-person events to foster team connection
  • Comprehensive health, dental, and vision insurance-we’re a healthcare company that prioritizes your health
  • 16 weeks of parental leave for all parents

About the company

Apella is applying computer vision and machine learning to improve the standard of care in the most critical aspect of healthcare: surgery. We build applications to enable surgeons, nurses, and hospital administrators to deliver the highest quality care.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on startup.jobs
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

2:07 min

Inspecting default bridge architectures and custom Docker networks

Oliver Seitz Oliver Seitz · World Congress 2025

6:36 min

Funding open source through GitHub Accelerator and Sponsors

Stormy Peters · World Congress 2023

2:15 min

Empowering domain teams with an open data platform

Sandhya Menon Sandhya Menon · World Congress 2026 Europe

7:10 min

Exploring pathways into the machine learning engineering field

Jose Luis Latorre Millas · LIVE

3:05 min

Audience questions on AI agents and pipeline vectorization

Joy Joy · World Congress 2024

2:40 min

Using GitHub primitives for internal documentation and corporate operations

Kyle Daigle · Coffee With Developers

Videos

See all

Related articles

See all