Research Engineer, Synthetic Data
- Discuss this with your agent
- Open in Claude
- Open in ChatGPT
Prepare application
- Draft this with your agent
- Open in Claude
- Open in ChatGPT
Role details
Tech stack
Job description
Weâre looking for Research Engineers to build our synthetic data pipeline. Youâll turn domain-specific workflows into synthetic training tasks that are realistic enough to enough to teach useful behavior, structured enough to generate at scale, and difficult enough to expand model capabilities., * Work with subject-matter experts to create synthetic tasks for training AI agents across a range of professional and technical domains
- Design synthetic task generation methods that produce diverse, realistic, and learnable tasks
- Build systems and tooling to mutate, validate, and improve synthetic tasks
- Analyze model and agent performance on synthetic tasks to understand what the tasks are teaching and where they fail
- Develop metrics to quantify and understand synthetic task diversity, realism, learnability, etc.
Requirements
You may be a good fit if you have:
- Proficiency in Python, Docker, and Linux environments
- Have experience with synthetic data research methods - please elaborate in your application
- Strong understanding of what âgood synthetic dataâ means and its limitations
- Built synthetic data pipelines end-to-end without a fully prescribed roadmap
- Experience working on environments, evals, and benchmarks
Strong candidates may also:
- Be detail-oriented and able to spot subtle inconsistencies or edge cases in synthetic data
- Be able to reason from first principles about task design, scoring, and failure modes
- Thrive in unstructured problem spaces
- Early-stage startup experience with ability to work independently in fast-paced environments
- Strong communication skills for remote collaboration across time zones
We prioritize technical aptitude and learning potential over years of experience. Motivated candidates are encouraged to apply even if they donât meet all criteria.
Benefits & conditions
Pulled from the full job description
- 401(k)
- Health insurance
- Paid time off
- Vision insurance
- Dental insurance
- Commuter assistance
- Paid holidays, * Competitive compensation
- 100% covered top-of-the-line medical, dental, and vision from Blue Shield of CA (US employees)
- Lunch and dinner when youâre in the office
- Company-wide holiday break (Christmas Eve to New Yearâs Day) on top of PTO and paid holidays
- Other perks including an Equinox membership, 401k, and commuter benefits (US employees)
- Unlimited* access to tokens for ChatGPT, Claude Code, Cursor, etc. *By unlimited, we mean no one on our token usage leaderboard has ever hit a limit. So we have no idea what the limit is.
About the company
HUD is building infrastructure to create RL training data and evals for frontier AI agents, as well as a marketplace to sell these to frontier labs through the HUD marketplace. Our platform is used by frontier labs, Fortune 500 companies, and startups. Weâve raised $16M from top VCs and were YC W25.
Apply for this position
This job is hosted externally. Click below to view the full posting and apply.
Prepare application
- Draft this with your agent
- Open in Claude
- Open in ChatGPT
Good distractions
Talks and stories from around this role â technically off-topic, practically not.
Moments
Explore playlistsVideos
See allRelated articles
See all
Top 6 Hackathons for Developers in 2023
Is Software Engineering Over-Saturated?
Dev Digest 121 - AI goes offline
What Developers Are Building to Win $1 Million with Apify