Data Engineer 1, Operational Technology - Operations #4941

GRAIL, Inc.
Durham, NC, United States
4 days ago
Apply on www.biospace.com
Prepare application

Role details

Contract type
Internship / Graduate position
Employment type
Full-time (> 32 hours)
Experience required
1 year minimum
Compensation
$86,000.0 - $106,000.0
Working hours
Regular working hours

Tech stack

Application Programming Interfaces (APIs) Artificial Intelligence Airflow Amazon S3 Data Analysis C++ (Programming Language) Databases Data Validation Information Engineering Extract Transform Load (ETL) Relational Databases File Transfer
+16 more
Python (Programming Language) Operational Databases Systems Development Life Cycle Cloud Services Software Engineering SQL Databases Systems Integration Rust (Programming Language) Snowflake Git Semi-structured Data Information Technology Data Lineage Software Version Control Data Pipelines Programming Languages

Job description

This role is based on-site in RTP, North Carolina, Monday through Friday. The position participates in an on-call rotation and may occasionally require weekend or holiday support for production incidents, maintenance, or critical deployments. Responsibilities:

  • Build and maintain data pipelines that ingest and integrate information from laboratory instruments, automation systems, sequencers, operational platforms, APIs, autonomous robotics platforms, databases and file based data sources.
  • Support downstream analytics, reporting, and AI systems by delivering clean, trustworthy datasets and timely data extracts for troubleshooting, root-cause investigations and platform improvements.
  • Develop and optimize SQL and transformation logic to cleanse, standardize, and model raw instrument and production data into reliable, well structured datasets.
  • Build and support datasets and data models used by operational dashboards, analytics, process monitoring, troubleshooting, and governed AI enabled workflows.
  • Implement orchestration, testing, monitoring and alerting so that data failures, freshness issues, schema changes, and incomplete processing are identified early.
  • Implement data validation and quality checks to ensure datasets are accurate, complete, and reliable.
  • Document pipelines, data models, and datasets to support reproducibility and compliance with ISO, CLIA, CAP, NYS, GMP, and FDA requirements.
  • Continuously improve your technical skills and the team’s engineering practices.

Requirements

  • Degree in Computer Science, Mathematics, Software Engineering, Data Science, Life Sciences, Physics or similar field.
  • 1+ years of relevant professional, internship, academic, or project experience in data engineering, analytics engineering, software development, or a related field, or equivalent practical experience.
  • Proficiency in SQL.
  • Working proficiency with one or more programming languages, such as Python, Rust, C++, or similar.
  • Basic understanding of ETL or ELT pipelines, relational databases, and structured or semi-structured data.
  • Strong attention to detail and a commitment to data quality, reliability and accuracy.
  • Ability to collaborate effectively in teams of technical and non-technical individuals, and comfortable working in a rapidly changing environment with dynamic objectives and fast iteration.
  • Ability to investigate technical problems methodically, continuously learn and communicate clearly.
  • A highly analytical mindset and eagerness to solve technical problems., * Familiarity with data pipeline orchestration and transformation tools such as Airflow, dbt, or comparable technologies.
  • Familiarity with cloud data platforms, object storage and warehouses such as AWS S3, Redshift, Glue, Snowflake or comparable technologies.
  • Familiarity integrating AI/agentic tooling into the data engineering SDLC.
  • Experience with semantic data modeling, data lineage, and automated data quality testing.
  • Familiarity with statistical methods or basic process analytics.
  • Exposure to manufacturing, clinical laboratory operations, diagnostics, or biotechnology.
  • Experience with version control systems such as Git and collaborative development practices.
  • Basic understanding of APIs, file transfers, networking and system integrations.

Benefits & conditions

The expected, full-time, annual base pay scale for this position is $86K - $106K. Actual base pay will consider skills, experience, and location.

This role may be eligible for other forms of compensation, including an annual bonus and/or incentives, subject to the terms of the applicable plans and Company discretion. This range reflects a good-faith estimate of the range that the Company reasonably expects to pay for the position upon hire; the actual compensation offered may vary depending on factors such as the candidate’s qualifications. Employees in this role are also eligible for GRAIL’s comprehensive and competitive benefits package, offered in accordance with our applicable plans and policies. This package currently includes flexible time-off or vacation; a 401(k) retirement plan with employer match; medical, dental, and vision coverage; and carefully selected mindfulness programs.

About the company

Our mission is to detect cancer early, when it can be cured. We are working to change the trajectory of cancer mortality and bring stakeholders together to adopt innovative, safe, and effective technologies that can transform cancer care.

We are a healthcare company, pioneering new technologies to advance early cancer detection. We have built a multi-disciplinary organization of scientists, engineers, and physicians and we are using the power of next-generation sequencing (NGS), population-scale clinical studies, and state-of-the-art computer science and data science to overcome one of medicine’s greatest challenges.

GRAIL is headquartered in the bay area of California, with locations in Washington, D.C., North Carolina, and the United Kingdom. It is supported by leading global investors and pharmaceutical, technology, and healthcare companies., GRAIL is a healthcare company whose mission is to detect cancer early, when it can be cured. GRAIL is focused on alleviating the global burden of cancer by developing pioneering technology to detect and identify multiple deadly cancer types early. The company is using the power of next-generation sequencing, population-scale clinical studies, and state-of-the-art computer science and data science to enhance the scientific understanding of cancer biology, and to develop its multi-cancer early detection blood test. GRAIL is headquartered in Menlo Park, CA with locations in Washington, D.C., North Carolina, and the United Kingdom. It is supported by leading global investors and pharmaceutical, technology, and healthcare companies. For more information, please visit www.grail.com.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.biospace.com
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

2:15 min

Empowering domain teams with an open data platform

Sandhya Menon Sandhya Menon · World Congress 2026 Europe

2:57 min

Core technical practices for robust data engineering

Sandhya Menon Sandhya Menon · World Congress 2026 Europe

3:43 min

The enduring legacy of the amazon S3 storage API

Chris Heilmann +3 · LIVE

6:21 min

Investigating push inefficiencies with upstream Git experts

Jonathan Creamer · Coffee With Developers

3:44 min

Automating storage savings with S3 intelligent tiering

Sébastien Stormacq · World Congress 2021

3:05 min

Audience questions on AI agents and pipeline vectorization

Joy Joy · World Congress 2024

Videos

See all

Related articles

See all