Senior/Staff Machine Learning Engineer, Perception

MooCo Robotics, Inc.
South San Francisco, CA, United States
18 days ago

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
6 years minimum
Compensation
$200,000.0 - $280,000.0
Working hours
Regular working hours
Job source

Tech stack

Artificial Intelligence Computer Vision Big Data Nvidia CUDA Software Debugging Python (Programming Language) Machine Learning Language Modeling OpenCV Tensorflow Sensor Fusion Real Time Systems
+7 more
Pytorch Information Technology Production Code Machine Learning Operations TensorRT Lidar GNSS

Job description

We’re looking for a skilled ML engineer to build the perception systems that give our autonomous machines human-like awareness in rugged, unstructured environments. You’ll develop computer vision and machine learning systems that turns noisy camera and LiDAR data into robust 3D scene understanding - enabling heavy equipment to operate safely through dust, glare, occlusion, and whatever messy conditions a working site throws at it.

The field is moving past bounding-box detection and hand-tuned tracking toward learned, dense scene representations, foundation-model-driven data engines, and uncertainty-aware perception. You’ll be at the center of that shift. This role is hands-on: you’ll write production-grade software, distill and optimize models for embedded hardware, and validate your work on real machines at operating around the world., * Develop real-time perception models for open-world obstacle and terrain understanding.

  • Build multi-modal fusion that combines camera and LiDAR into a unified 3D/BEV representation, robust to occlusions, sensor degradation, and GNSS outages.
  • Optimize models for low-latency inference on resource-constrained hardware, balancing accuracy and performance.
  • Design auto-labeling pipelines that leverage foundation models and teacher-student distillation to scale labeling and close the loop from real-world field interventions.
  • Design data and evaluation pipelines that curate large multi-sensor datasets and surface failures fast, with strong visualization and debugging tooling.
  • Analyze performance metrics and iterate on algorithms to improve accuracy and efficiency of various perception subsystems.

Requirements

  • A MS/PhD in Computer Science, AI, or a related field, or 6+ years of industry experience building vision-based perception systems.
  • Deep expertise developing and deploying modern perception models: detection, segmentation, mono/stereo/metric depth, BEV/occupancy, sensor fusion, and 3D scene understanding.
  • Fluency adapting, fine-tuning, and distilling large pre-trained vision and vision-language models.
  • Strong grounding in multi-sensor integration (camera, LiDAR, radar): calibration, spatiotemporal sync, and cross-modal fusion.
  • Experience handling large datasets efficiently and organizing them for labeling, training and evaluation.
  • Fluency in Python with PyTorch/TensorFlow/OpenCV and the ability to write efficient, production-ready code for real-time systems.
  • Proven ability to design experiments, analyze metrics (mAP, IoU, latency/throughput, and calibration/ECE), and optimize to meet stringent real-world performance and safety requirements.
  • An eagerness to get your hands dirty and agility in a fast-moving, collaborative, small team environment with lots of ownership.

What Makes You a Strong Fit

  • Experience architecting multi-sensor ML systems from scratch.
  • Experience building auto-labeling / data-engine flywheels at scale.
  • Experience with compute-constrained pipelines including optimizing models to balance the accuracy vs. performance tradeoff, leveraging TensorRT, model quantization, etc.
  • Familiarity with emerging predictive world models for anticipation, anomaly detection, or closed-loop simulation, and adjacent policy paradigms such as Vision-Language-Action (VLA) and World-Action (WAM) models.
  • Experience with compute-constrained deployment: TensorRT, model quantization, and custom CUDA operations.
  • Publications at top-tier perception/robotics venues (CVPR, ICRA, CoRL, RSS, etc.).
  • Passion for how we feed, build, move, and maintain the world.

Benefits & conditions

Pulled from the full job description

  • 401(k)
  • Health insurance
  • Vision insurance
  • Health savings account
  • Dental insurance
  • Flexible spending account
  • Stock options, The US base salary range for this full-time position is $200,000 to $280,000 + equity + benefits + unlimited PTO, * 100% covered medical, dental, and vision for the employee (partner, children, or family is additional)

About the company

At Agtonomy, we’re not just building tech-we’re transforming how vital industries get work done. Our Physical AI and fleet services turn heavy machinery into intelligent, autonomous systems that tackle the toughest challenges in agriculture, turf, and beyond. Partnering with industry-leading equipment manufacturers, we’re creating a future where labor shortages, environmental strain, and inefficiencies are relics of the past. Our team is a tight-knit group of bold thinkers-engineers, innovators, and industry experts-who thrive on turning audacious ideas into reality. If you want to shape the future of industries that matter, this is your shot.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.indeed.com

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

3:37 min

Transitioning automated driving to neural networks and model-based perception

Katrin Lehmann Katrin Lehmann +1 · WWC 2025

2:17 min

Validating lidar sensor models against real noise

Ulrich Wurstbauer +1 · LIVE

4:53 min

Achieving real-time tracking performance with OpenCV and segmentation

Thomas Endres Thomas Endres +2 · WWC 2021

1:31 min

Core drivers fueling modern robotic capabilities

Thomas Tomow Thomas Tomow · WWC 2025

2:08 min

Processing physical environment data with lidar models

Oliver Zimmert · LIVE

2:03 min

Enhancing realistic mask blending through OpenCV image inpainting

Thomas Endres Thomas Endres +2 · WWC 2021

Videos

See all

Related articles

See all