> Markdown version of [/jobs/ext/2720841-ml-engineer](https://www.wearedevelopers.com/jobs/ext/2720841-ml-engineer). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # ML Engineer - **Company:** Torc Robotics, Inc. - **Location:** United States (Remote available) - **Experience:** Expert - **Salary:** $177,300.0 - $212,800.0 - **Contract:** Permanent contract - **Skills:** Artificial Intelligence, Amazon Web Services, Amazon S3, Computer Vision, Big Data, Cloud Computing, Cloud Engineering, Data Visualization, Programming Tools, Amazon DynamoDB, Github, Python (Programming Language), Language Modeling, Object Detection, OpenGL, Operational Databases, Search Technologies, Sensor Fusion, Parquet, Three.js, Data Processing, Pytorch, Large Language Models, Apache Spark, Deep Learning, Pandas, Information Technology, Integration Frameworks, Machine Learning Operations, Terraform, Software Version Control, Docker, Databricks - **Published:** September 4, 2026 - **Apply:** https://startup.jobs/senior-ml-engineer-vlm-torc-robotics-9639321 ## About the Role * Considered highly skilled and proficient in discipline; conducts complex, important work under minimal supervision and with wide latitude for independent judgment. * Scope of Influence: Expected to drive alignment across team interfaces to the rest of the organization. Designs, maintains, and owns team technical solutions and drives consensus. Mentors and guides engineers within the group. * Bachelor's Degree in Computer Science, Robotics, Electrical Engineering, or related technical field plus competences typically acquired through 6+ years of experience; OR Master's Degree in a related technical field plus competences typically acquired through 3+ years of experience. Required Qualifications (some combination of the following skills): * Computer Vision & Deep Learning - model training and at least two of: 2D/3D Object Detection, Tracking, Sensor Fusion, Semantic Segmentation, BEV, Depth Estimation. * Multimodal / VLM experience - hands-on work with vision-language models, open-vocabulary or zero-shot recognition, dense captioning, or semantic embeddings / search applied to perception data. * Model Data Curation - building targeted datasets that measurably improve downstream model performance; large-scale Parquet data processing (Databricks, Daft, Pandas, etc.). * Distributed ML & data frameworks - PyTorch, Lightning, Ray, Spark, or equivalent for training and large-scale data processing. * Scaled MLOps & Tooling - experiment tracking, model registry, MLflow / Weights & Biases, and ML metrics, evaluation, and quality. * Development Tools & Eco-System (at scale) - strong Python software development, VDI and cloud-based development environments, CI systems (GitHub Actions), and Docker. ## Description * Own the offline dataset pipeline - design, implement, test, and deploy Cloud-based pipelines that convert logged multi-sensor data into VLM/VLA training datasets, spanning geometric labels (3D/2D detection, tracking, segmentation, depth) through semantic, scenario-level, and action/trajectory-grounded annotations. * Build VLM-assisted auto-labeling - develop open-vocabulary detection, dense captioning, semantic enrichment, and scene/scenario description generation that move beyond closed-set bounding boxes, using foundation models to scale annotation and cut manual labeling cost. * Generate reasoning-grounded labels - produce language-grounded reasoning and chain-of-causation style annotations, temporally aligned to ego-motion and trajectories, to support VLA training and explainable driving behavior. * Mine and curate the long tail - surface rare, difficult, and high-uncertainty scenarios, and build curated datasets that measurably improve downstream VLM/VLA model metrics rather than simply adding volume. * Close the data flywheel - define dataset schemas, quality metrics, and validation; track auto-labeling quality against model requirements; route model failures back into re-labeling and retraining loops. * Partner with the end-to-end model team - co-define dataset specifications with VLM/VLA model developers, own the quality bar and delivery cadence, and operationalize a continuous dataset delivery loop into their training pipelines. * Scale on cloud infrastructure - build distributed, reproducible pipelines using columnar data formats and distributed compute, with disciplined software practices, version control, and documentation. * Lead and mentor - serve as project lead, guide less-experienced engineers, run design reviews, set coding and annotation standards, and drive alignment across team interfaces to the rest of the organization. * Stay current - track the latest advances in multimodal models, auto-labeling, and end-to-end autonomous driving, and translate relevant research into production data systems., * End-to-end / VLA driving - familiarity with VLM/VLA or end-to-end driving models, trajectory and action grounding, or chain-of-causation / reasoning-trace datasets. * Auto-labeling foundation models - experience with segmentation, open-vocabulary detectors, or VLM/LLM-driven data engines for annotation and verification. * High-throughput model serving - vLLM, SGLang, or similar for batch auto-labeling and inference at scale. * Semantic inference & retrieval - attribute mapping, semantic search, and vector databases (e.g., LanceDB) for automotive data. * AV data standards & tooling - scenario-description standards such as Pegasus layers; parsing robotics formats (ROS bags, MCAP) and optimizing columnar storage (Parquet, Arrow). * Cloud development & orchestration - Terraform and AWS managed services (S3, ECS, Lambda, DynamoDB, Step Functions, Athena); AWS HyperPod / Anyscale; inference orchestration. * Data visualization - Foxglove, FiftyOne (51), three.js, OpenGL, or similar for dataset inspection and accessibility. * Evaluation & research - closed-loop / open-loop evaluation frameworks (e.g., NavSim-style planning metrics); publications in top-tier CV/AI/Robotics venues (CVPR/ECCV/ICCV, NeurIPS/ICLR/ICML, CoRL). ## Related Videos - [Docker Compose: Rediscovered](https://www.wearedevelopers.com/videos/1978-docker-compose-rediscovered) - [Advanced Typing in TypeScript](https://www.wearedevelopers.com/videos/496-advanced-typing-in-typescript) - [MLOps - What’s the deal behind it?](https://www.wearedevelopers.com/videos/392-mlops-what-s-the-deal-behind-it) - [Docker build without Docker](https://www.wearedevelopers.com/videos/100114-docker-build-without-docker) - [DevOps for AI: running LLMs in production with Kubernetes and KubeFlow](https://www.wearedevelopers.com/videos/1222-devops-for-ai-running-llms-in-production-with-kubernetes-and-kubeflow) - [Data Science on Software Data](https://www.wearedevelopers.com/videos/162-data-science-on-software-data) ## Related Articles - [What Are Large Language Models?](https://www.wearedevelopers.com/magazine/304-what-are-large-language-models) - [MLops – Deploying, Maintaining And Evolving Machine Learning Models in Production](https://www.wearedevelopers.com/magazine/115-mlops-deploying-maintaining-and-evolving-machine-learning-models-in-production) - [How to Become an AI Engineer](https://www.wearedevelopers.com/magazine/331-how-to-become-an-ai-engineer) - [MLOps – What’s the deal behind it?](https://www.wearedevelopers.com/magazine/125-mlops-what-s-the-deal-behind-it) - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [MLOps And AI Driven Development](https://www.wearedevelopers.com/magazine/82-mlops-and-ai-driven-development)