> Markdown version of [/jobs/ext/1383969-software-engineer-perception-attributes-autolabeling-pipeline-in-united-states](https://www.wearedevelopers.com/jobs/ext/1383969-software-engineer-perception-attributes-autolabeling-pipeline-in-united-states). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Software Engineer, Perception Attributes Autolabeling Pipeline in , United States - **Company:** Energy Jobline - **Location:** Foster City, CA, United States - **Experience:** Experienced - **Salary:** $158,080.0 - $178,880.0 - **Contract:** Temporary contract - **Skills:** Application Programming Interfaces (APIs), Amazon Web Services, Amazon S3, C++ (Programming Language), Data Files, Data Infrastructure, Python (Programming Language), Runbook, Prompt Engineering, Backend, Pyspark, Machine Learning Operations, Data Pipelines - **Published:** July 22, 2026 - **Apply:** https://www.energyjobline.com/job/software-engineer-perception-attributes-autolabeling-pipeline-united-states-31212350 ## About the Role 3+ years of backend/data pipeline engineering experience; Strong Python; comfort with C++; Large-dataset experience with PySpark or equivalent. ML fundamentals - understanding of model inference, embeddings, structured output, and common eval metrics (precision, recall, calibration); able to reason about ML data shapes and integration patterns. Experience integrating foundation models (Gemini, OpenAI, Anthropic) at production scale; Excellent written communication for design docs and runbooks., 3+ years of backend/data pipeline engineering experience. Strong Python; comfort with C++. Large-dataset experience with PySpark or equivalent. ML fundamentals - understanding of model inference, embeddings, structured output, and common eval metrics (precision, recall, calibration); able to reason about ML data shapes and integration patterns. Experience integrating foundation models (Gemini, OpenAI, Anthropic) at production scale. Excellent written communication for design docs and runbooks. ## Description The Perception Attribute Flywheel team is looking for a Software Engineer to build and operate the autolabeling pipeline that accelerates human annotation throughput on vehicle attribute classification tasks. The company is building a future for Riders, not drivers. The accuracy of our perception attribute models - recognizing emergency vehicles, school buses, brake lights, hazard signals, and more - depends on a steady flow of high-quality labeled examples drawn from our fleet's drive data. Today, every label is produced by a human annotator from scratch. We are building a pipeline that uses off-the-shelf foundation models (Gemini, SigLIP, CLIP) to pre-label tasks, so human reviewers verify and correct rather than labeling from scratch. This role owns the pipeline engineering for that system: ingesting queued tasks from our annotator service, calling foundation-model APIs at fleet scale, writing structured predictions back into the labeling workflow, and operating the whole thing reliably. The team lead and supporting ML engineers own model selection, prompt design, and evaluation methodology; this role partners closely with them but is not expected to own those decisions. If you take pride in building reliable, observable, well-tested data pipelines and want to ship a system that visibly accelerates an autonomous vehicle program, you will excel in this role. Responsibilities: Build the autolabeling pipeline: ingest queued tasks from the annotator service, dispatch them to foundation-model APIs (Gemini and others), parse structured outputs, and write pre-labels back to the labeling workflow. Build the observability layer: per-task latency, per-model cost, per-attribute coverage, and error-mode dashboards. Run experiments designed by the team lead - set up the inputs, execute, and collect outputs in formats the ML engineers can analyze. Integrate the pipeline cleanly with existing systems, partnering with the data infrastructure team. Document the system, write runbooks, and ensure a clean handoff at end of the engagement., End-to-end ML pipeline stewardship - owned an ML system in production from data ingest through inference through monitoring. Annotation tooling or human-in-the-loop ML workflows. Autonomous-systems data pipelines. AWS, especially S3, ECS/EKS, Lambda. Working in a codebase shared with ML engineers (proto schemas, joint deploys). ## Related Videos - [PySpark - Combining Machine Learning & Big Data](https://www.wearedevelopers.com/videos/44-pyspark-combining-machine-learning-big-data) - [Machine Learning for Software Developers (and Knitters)](https://www.wearedevelopers.com/videos/154-machine-learning-for-software-developers-and-knitters) - [Developing the Backend with Stefan Lingler, CTO at Shpock](https://www.wearedevelopers.com/videos/100360-developing-the-backend-with-stefan-lingler-cto-at-shpock) - [Technical Documentation - How Can I Write Them Better and Why Should I Care?](https://www.wearedevelopers.com/videos/681-technical-documentation-how-can-i-write-them-better-and-why-should-i-care) - [Data Science in Retail](https://www.wearedevelopers.com/videos/586-data-science-in-retail) - [Empowering Retail Through Applied Machine Learning](https://www.wearedevelopers.com/videos/976-empowering-retail-through-applied-machine-learning) ## Related Articles - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [How to Become an AI Engineer](https://www.wearedevelopers.com/magazine/331-how-to-become-an-ai-engineer) - [Data Engineer Salary UK](https://www.wearedevelopers.com/magazine/253-data-engineer-salary-uk) - [Dev Digest 120 - Apple and peers](https://www.wearedevelopers.com/magazine/455-dev-digest-120-apple-and-peers) - [What Are Large Language Models?](https://www.wearedevelopers.com/magazine/304-what-are-large-language-models) - [Dev Digest 121 - AI goes offline](https://www.wearedevelopers.com/magazine/456-dev-digest-121-ai-goes-offline)