> Markdown version of [/jobs/ext/619286-research-data-engineer](https://www.wearedevelopers.com/jobs/ext/619286-research-data-engineer). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Research Data Engineer - **Company:** Hotspot Therapeutics, Inc. - **Location:** Boston, MA, United States - **Experience:** Expert - **Contract:** Permanent contract - **Skills:** Adaptable Database Systems, Artificial Intelligence, Amazon Web Services, Bioinformatics, Cloud Computing, Computational Biology, Data Cleansing, Extract Transform Load (ETL), Relational Databases, Python (Programming Language), PostgreSQL, Machine Learning, Oracle (Applications), Software Engineering, Technical Data Management Systems, Freeform SQL, Cloud Platform System, Data Ingestion, Information Technology, Data Delivery, Data Pipelines - **Published:** June 18, 2026 - **Apply:** https://www.hotspotthera.com/job/senior-research-data-engineer/ ## About the Role * Advanced Scientific Background: PhD degree in Bioinformatics, Computational Biology, Computer Science, or a related quantitative life sciences discipline and a minimum of 2+ years of relevant industry experience or a combination of post-doctoral and industry experience; MS degree in bioinformatics or related discipline with 5+ years of relevant industry experience may be considered. * Production-Grade Engineering: Proven commercial or advanced hands-on experience writing clean, maintainable Python code and advanced SQL queries. You know how to build software that stands up to production environments. * Cloud & Database Fluency: Demonstrated experience navigating cloud environments (AWS/GCP) and managing relational database schemas (Oracle, Postgres). * Execution-Driven Planner: An autonomous problem-solver who brings clarity to complex data dependencies, actively tracks technical risk, and handles firefighting or data cleaning with equal dedication. * Excellent Communicator: Ability to synthesize complex technical data constraints across disciplines and drive follow-through in a lean, fast-paced startup environment., At this time, HotSpot is not able to offer Visa sponsorship for this position. Candidates must be authorized to work in the United States without current or future sponsorship. ## Description We are seeking a hands-on, execution-driven Senior Research Data Engineer to scale and secure HotSpot's growing data operations infrastructure. This role serves as a vital operational bridge between our core software/cloud architecture and our computational biology and laboratory teams. You will partner closely with our Senior Data Engineer to drive alignment, eliminate single points of failure, and establish an automated, production-grade data paradigm that accelerates our drug discovery workflows., * Lead AI & LLM Data Readiness: Architect and execute the strategy to structure, clean, and index our historical and incoming research datasets, making them fully ready for machine learning and advanced ontology initiatives. * Automate External Data Pipelines: Own and maintain robust Python ETL pipelines, ensuring seamless automated ingestion of chemistry and biology experiments from external CROs. * System Stewardship & Data Ops: Act as the primary technical steward for our core research platforms (including Revvity Signals ELN, Certara D360, and our internal data ingestion tool, Ladle), handling Oracle schema updates, routine database maintenance, and direct user support for our scientists. * Bridge Science and Engineering: Partner across disciplines to help scientists map biology and chemistry workflows into production templates, serving as a data continuity bridge that translates laboratory progress into scalable data models. * Drive Operational Follow-Through: Identify dependencies and critical paths in our data delivery pipelines, actively manage infrastructure risks, and provide clear technical context to leadership. ## Related Videos - [Implementing continuous delivery in a data processing pipeline](https://www.wearedevelopers.com/videos/73-implementing-continuous-delivery-in-a-data-processing-pipeline) - [Optimizing Discovery: PostgreSQL's Role in Transforming GetYourGuide's Search](https://www.wearedevelopers.com/videos/1647-optimizing-discovery-postgresql-s-role-in-transforming-getyourguide-s-search) - [Geometric deep learning for drug discovery](https://www.wearedevelopers.com/videos/264-geometric-deep-learning-for-drug-discovery) - [Tracking vehicles at scale](https://www.wearedevelopers.com/videos/1999-tracking-vehicles-at-scale) - [Why and when should we consider Stream Processing frameworks in our solutions](https://www.wearedevelopers.com/videos/1085-why-and-when-should-we-consider-stream-processing-frameworks-in-our-solutions) - [Discover the open source trio you didn’t expect: .NET and PostgreSQL on Linux](https://www.wearedevelopers.com/videos/2042-discover-the-open-source-trio-you-didn-t-expect-net-and-postgresql-on-linux) ## Related Articles - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [Data Engineer Salary UK](https://www.wearedevelopers.com/magazine/253-data-engineer-salary-uk) - [Is Software Engineering Over-Saturated?](https://www.wearedevelopers.com/magazine/418-is-software-engineering-over-saturated) - [Where To Find Software Engineering Jobs](https://www.wearedevelopers.com/magazine/396-where-to-find-software-engineering-jobs) - [Fully Remote Software Engineer Jobs](https://www.wearedevelopers.com/magazine/447-fully-remote-software-engineer-jobs) - [The Most Popular IT Jobs on the Market](https://www.wearedevelopers.com/magazine/376-the-most-popular-it-jobs-on-the-market)