Founding Machine Learning Engineer job in San Francisco
- Discuss this with your agent
- Open in Claude
- Open in ChatGPT
Role details
Tech stack
+4 more
Job description
Weāre looking for founding Machine Learning Engineers (MLEs) to own and improve our core action models end-to-end - the intelligence that powers Compositeās proactive automation platform.
Youāll work at the intersection of LLM inference, browser understanding, and low-latency systems, shipping models that need to feel instant while reasoning over complex page state and user context.
Unlike hosted browser solutions that introduce latency and auth barriers, or consumer-focused āAI browsers,ā we run AI directly through professionalsā existing browsers via a Chrome extension, creating instant response times with zero migration or IT friction. This architecture creates unique ML challenges.
This is a high-ownership role on our small, exceptional team where your work ships directly to users and has the potential to tangibly improve the work lives of hundreds of millions of people.
About Composite
College-educated professionals spend 85% of their day as digital factory workers in Chrome, clicking through repetitive browser tasks.
Composite is building the proactive layer for productivity so professionals around the world can focus on meaningful, high-leverage work. Weāre training action prediction models that run in real time, anticipating what youāll do next based on page context and prior interactions.
Weāve raised $5.6M in seed funding led by Nat Friedman and Daniel Gross, with participation from Menlo Ventures, Anthropicās Anthology Fund, SVAngel, and other incredible investors.
What Youāll Work On
- Improve the accuracy and latency of our core models across diverse web applications to predict usersā intended next actions and execute them faster than manual input
- Design and optimize LLM inference pipelines, including token caching strategies, streaming architectures, and network-level optimizations between client and server
- Build evaluation frameworks and data pipelines to measure and improve model quality at scale
- Experiment with retrieval-augmented approaches using vector databases for contextual memory
- Develop synthetic data generation pipelines for browser interaction training data
- Work with DOM states, accessibility trees, and user interaction data to improve browser understanding
- Ship features end-to-end that go directly to users - this is not a research-only role, * Disagree and commit: Respectfully challenge decisions you disagree with, even when itās uncomfortable. Donāt censor yourself or your ideas. Once a decision is determined, everyone commits wholly.
- Clear and consistent standards: Decisions are made based on a shared framework that applies for everyone. We donāt leave room for ārules for you, not for meā or any perceived hypocrisy.
- Over-communicate: Nothing slows down a company more than confusion, mis-, or under-communication. Leave no room for ambiguity. Ask dumb questions. Write things down clearly.
- Health is #1: Stay hydrated. Eat a balanced diet. Sleep 8 hours a night. Exercise frequently. Maintain good social and mental health. Not doing so affects your mood and long-term productivity.
Requirements
- Strong ML fundamentals with hands-on experience training and deploying models in production
- Obsessive about latency - experience optimizing inference pipelines to feel instant to end users
- Deep care about data quality, with the instinct to build tooling that ensures it
- Experience with LLMs, transformer architectures, or sequence prediction problems
- Comfortable working across the stack - our system spans a Chrome extension, Electron app, Cloudflare Workers edge proxy, and inference providers
Core Qualities
- Character: Youāre someone weād want to work closely with for the next ten years. You approach challenges with curiosity rather than ego. Youāre a team player, a great communicator, and arenāt afraid to be wrong.
- Work Ethic: Youāre energized by hard problems and comfortable working intensely toward ambitious goals.
- Raw Intelligence: You can quickly understand complex systems and solve novel, ambiguous problems with self-guidance.
Bonus
- Experience with browser automation, Chrome extensions, or web scraping at scale
- Familiarity with accessibility tree / DOM parsing for page understanding
- Background in RL or online learning from user interaction data
- Experience with vector databases (e.g., Turbopuffer, Pinecone) and hybrid search
- Full-stack development experience (TypeScript, Node.js, React)
Apply for this position
This job is hosted externally. Click below to view the full posting and apply.
Apply on diversity.comGood distractions
Talks and stories from around this role ā technically off-topic, practically not.
Moments
Explore playlistsVideos
See allRelated articles
See all
Dev Digest 120 - Apple and peers
How to Become an AI Engineer
Dev Digest 121 - AI goes offline
Dev Digest 137 - AI'm not sure about this