Software Engineer - Data Platform

Spectraforce
Menlo Park, CA, United States
4 days ago
Apply on leoforce.us
Prepare application

Role details

Contract type
Internship / Graduate position
Employment type
Full-time (> 32 hours)
Experience level
Starter
Experience required
0 years minimum
Working hours
Regular working hours
Job source

Tech stack

Application Programming Interfaces (APIs) Artificial Intelligence Amazon Web Services Amazon S3 Big Data Cloud Computing Computer Programming Databases Data Infrastructure Data Structures Data Systems Software Debugging
+16 more
Programming Tools Distributed Systems Systems Theories Github Python (Programming Language) PostgreSQL Open Source Technology Azure Machine Learning SQL Databases Large Language Models Indexer Backend Information Technology Production Code Data Pipelines Databricks

Job description

Client Super Intelligence Labs - ESI Embodiment Team

  • Join the small team building Conductor, the data and operations platform for our robotics program. Conductor manages high-volume, multimodal robot data from ingestion through quality control, discovery, training, and evaluation.
  • You will own meaningful production systems, work directly with researchers and robotics engineers, and ship quickly with a high degree of autonomy.

Key Projects / Day-to-Day Responsibilities

  • Build reliable backend services, APIs, and data pipelines using Python and Go across PostgreSQL and AWS/S3.
  • Design systems for ingesting, indexing, transforming, validating, and serving large robotics datasets.
  • Develop agentic and LLM-powered workflows for data quality, enrichment, automation, and research productivity with evaluations and guardrails.
  • Improve performance, observability, correctness, and operational reliability across the platform.
  • Work end-to-end: clarify requirements, make pragmatic design choices, implement, test, deploy, and iterate with users.

Requirements

Full stack development, large scale data processing. Must be good with AI assisted coding. Need to have good engineering taste, exp with distributed systems, backend, large scale workloads, etc.

  • Contributions to open-sourced projects - shows initiative on their part (working on a project they are passionate about)
  • core contributor to well known or widely used software (show github profile link)
  • Competitive internships - working with big tech companies (quant, good engineering culture tech): jane street, hrt, google, stripe, databricks, cloudflare, meta, vercel, openai), companies known for having high engineering standards
  • Prioritize graduating students and candidates within two years of graduation from rigorous CS/engineering programs (for example MIT, Stanford, CMU, Berkeley, UIUC, Waterloo, Georgia Tech, UT Austin, Princeton, Cornell, or comparable programs). School is a sourcing signal, not a gate

Must-Have Skills:

  • Exceptional Software Engineering Fundamentals: Data structures, algorithms, systems, databases, and debugging.
  • Backend / Data Systems Experience: Proven track record building substantial backend, data-intensive, or distributed systems (via internships, research, open source, or ambitious projects).
  • Languages & Databases: Strong programming skills in at least one statically typed language (Go) alongside Python and SQL.

Nice-to-have Skills:

  • Open Source & Research:

Substantial contributions to respected open-source projects or systems/data research. Exceptional software-engineering fundamentals, including data structures, algorithms, systems, databases, and debugging. Evidence of building a substantial backend, data-intensive, distributed, or developer-infrastructure system through internships, research, open source, or ambitious independent/course projects.

  • AI / Workflows: Experience building useful agentic workflows with real evaluations.

Strong programming skills and experience with at least one statically typed language; working knowledge of SQL.

  • Ability to write clear, maintainable production code and reason carefully about reliability, performance, and data correctness.

High ownership, intellectual curiosity, sound judgment, and comfort moving quickly in an ambiguous environment. Cloud & Infrastructure: Hands-on experience with PostgreSQL, AWS/S3, ML platforms, or developer tools.

Especially strong signals

  • Maintaining or making substantial contributions to a respected open-source project; internships on data, infrastructure, databases, ML platforms, or developer tools; building useful agentic workflows with real evaluations; systems research; or exceptional technical projects with real users or meaningful scale. A BS/MS in computer science or a related field from a program with a strong systems culture is helpful, but demonstrated ability matters more than pedigree. Prior robotics experience is not required.

Years of Experience:

  • Early Career / New Graduate (0-2 years of experience, more is good too)

Degrees/Certifications Required:

  • BS/MS in Computer Science or a related field (preferably from top systems programs like MIT, Stanford, CMU, Berkeley, UIUC, Waterloo, Georgia Tech, UT Austin, Princeton, Cornell). Demonstrated ability matters more than pedigree.

Are there any types of candidate profiles or skills that may not be the right fit for this team? Possible Disqualifiers

  • Lack of strong software engineering fundamentals or programming ability in statically typed languages/SQL.
  • No evidence of building substantial backend or data-intensive systems projects/internships.
  • Inability to demonstrate at least two concrete proof points (selective internship, open source, research, shipped scale project, or competitive programming/systems coursework). (Note: Prior robotics experience is NOT required/disqualifying).

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on leoforce.us
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

6:36 min

Funding open source through GitHub Accelerator and Sponsors

Stormy Peters · World Congress 2023

1:52 min

Structuring and scaling the backend engineering team

Stefan Lingler Stefan Lingler +1 · Coffee With Developers

2:36 min

Analyzing limitations with PostgreSQL bitmap heap scans

Dharin Shah Dharin Shah · World Congress 2025

3:09 min

Balancing data science skillings alongside systems engineering rigor

Nico Schmidt · LIVE

2:40 min

Using GitHub primitives for internal documentation and corporate operations

Kyle Daigle · Coffee With Developers

2:37 min

Optimizing technical profiles for AI sourcing and recruitment

Mina Golesorkhi Mina Golesorkhi · World Congress 2026 Europe

Videos

See all

Related articles

See all