Data Engineer
We are a healthcare company, pioneering new technologies to advance early cancer detection.
We are a healthcare company, pioneering new technologies to advance early cancer detection. We have built a multi-disciplinary organization of scientists, engineers, and physicians and we are using the power of next-generation sequencing (NGS), population-scale clinical studies, and state-of-the-art computer science and data science to overcome one of medicine’s greatest challenges.
GRAIL is headquartered in the bay area of California, with locations in Washington, D.C., North Carolina, and the United Kingdom. It is supported by leading global investors and pharmaceutical, technology, and healthcare companies.
For more information, please visit grail.com
As a Senior Data Engineer on the Operational Technology team, you will own the data platform that connects GRAIL's lab instruments, automation systems, and operational platforms to a trusted, well modeled data foundation. You will architect and lead the development of complex, business-critical ingestion and transformation pipelines end to end, set the standards and patterns the broader team builds on, and act as a technical point of contact across systems engineering, lab operations, data science, and automation engineering. You will work independently on problems of diverse scope, devising solutions where precedent is limited, and you will mentor less experienced engineers while raising the bar on reliability, data quality, and engineering practice. This is a hands-on senior role for a fully qualified data engineer who is ready to take ownership of critical infrastructure in a fast paced, regulated environment. Expect to work alongside a talented and highly motivated team that moves quickly.
This role is based on-site in RTP, North Carolina, Monday through Friday. The position participates in an on-call rotation and may occasionally require weekend or holiday support for production incidents, maintenance, or critical deployments.
Responsibilities:
Architect, build, and maintain complex data pipelines that ingest and integrate information from laboratory instruments, automation systems, sequencers, operational platforms, APIs, autonomous robotics platforms, databases and file based data sources.
Own critical pipelines and data models end to end, from design through production operation, working independently on problems of diverse scope and adapting existing approaches where limited precedent exists.
Define and evolve the data architecture, modeling standards, and engineering patterns that the broader team builds on, and drive adoption across the group.
Support downstream analytics, reporting, and AI systems by delivering clean, trustworthy datasets and timely data extracts for troubleshooting, root-cause investigations and platform improvements.
Develop and optimize advanced SQL and transformation logic to cleanse, standardize, and model raw instrument and production data into reliable, well structured datasets.
Posted July 31, 2026