Data Engineer 1, Operational Technology - Operations #4941

Grailbio
Durham, NC
On-site

Who this role is best for

Geared toward data engineers comfortable with building and maintaining data pipelines for healthcare and biotechnology, with a focus on collaboration and compliance.

Best fit for

  • Candidates with a background in healthcare or biotechnology and experience in data pipeline development
    — “pioneering new technologies to advance early cancer detection
  • Professionals who value compliance and documentation in their data engineering practices
    — “Document pipelines, data models, and datasets to support reproducibility and compliance with ISO, CLIA, CAP, NYS, GMP, and FDA requirements.
  • Individuals with a strong analytical mindset and interest in integrating AI into data workflows
    — “Familiarity integrating AI/agentic tooling into the data engineering SDLC.

Things to consider

  • The role demands adherence to multiple regulatory compliance standards, including ISO, CLIA, CAP, NYS, GMP, and FDA
    — “compliance with ISO, CLIA, CAP, NYS, GMP, and FDA requirements.
  • Candidates must be prepared to work in a dynamic environment with fast iteration and shifting priorities
    — “comfortable working in a rapidly changing environment with dynamic objectives and fast iteration.

How to stand out

  • Highlight experience with data pipeline orchestration tools like Airflow or dbt in your resume and interview responses
    — “Familiarity with data pipeline orchestration and transformation tools such as Airflow, dbt, or comparable technologies.
  • Demonstrate your ability to implement data validation and quality checks in your past projects
    — “Implement data validation and quality checks to ensure datasets are accurate, complete, and reliable.
  • Showcase your SQL and transformation logic skills with concrete examples of data modeling and cleansing
    — “Develop and optimize SQL and transformation logic to cleanse, standardize, and model raw instrument and production data into reliable, well structured datasets.
  • Emphasize your experience with cloud data platforms like AWS S3, Redshift, or Snowflake
    — “Familiarity with cloud data platforms, object storage and warehouses such as AWS S3, Redshift, Glue, Snowflake or comparable technologies.
  • Demonstrate your ability to work in cross-functional teams with technical and non-technical members
    — “Ability to collaborate effectively in teams of technical and non-technical individuals.
Pace · Fast PacedCollaboration · HighAutonomy · MediumDecision Impact · Team

Derived from job-description analysis by Serendipath's career intelligence engine.

What success looks like

  • Built and maintained data pipelines
  • Supported downstream analytics
  • Developed and optimized SQL and transformation logic
  • Implemented data validation and quality checks
  • Documented pipelines and datasets
Typical background
Degree in Computer Science, Mathematics, Software Engineering, Data Science, Life Sciences, Physics1+ years of relevant professional experience in data engineering or related field

Skills & requirements

Required

SQLETL PipelinesRelational DatabasesStructured Or Semi-structured DataCloud Data PlatformsData ValidationData Quality ChecksData ModelingData PipelinesData IntegrationData OrchestrationData TestingData MonitoringData AlertingData GovernanceData Compliance

Preferred

AirflowdbtAWS S3Redshift

Stack & domain

SQLPythonRustC++ETLELTAirflowdbtAWS S3RedshiftAttention To DetailCommitment To Data QualityCollaborationProblem-solvingAnalytical MindsetHealthcareData ScienceNext-generation SequencingCloud Data Platforms

About the role

Original posting from Grailbio via Lever

Our mission is to detect cancer early, when it can be cured. We are working to change the trajectory of cancer mortality and bring stakeholders together to adopt innovative, safe, and effective technologies that can transform cancer care.

We are a healthcare company, pioneering new technologies to advance early cancer detection. We have built a multi-disciplinary organization of scientists, engineers, and physicians and we are using the power of next-generation sequencing (NGS), population-scale clinical studies, and state-of-the-art computer science and data science to overcome one of medicine’s greatest challenges.

GRAIL is headquartered in the bay area of California, with locations in Washington, D.C., North Carolina, and the United Kingdom. It is supported by leading global investors and pharmaceutical, technology, and healthcare companies.

For more information, please visit grail.com

Responsibilities::

-

Build and maintain data pipelines that ingest and integrate information from laboratory instruments, automation systems, sequencers, operational platforms, APIs, autonomous robotics platforms, databases and file based data sources.

-

Support downstream analytics, reporting, and AI systems by delivering clean, trustworthy datasets and timely data extracts for troubleshooting, root-cause investigations and platform improvements.

-

Develop and optimize SQL and transformation logic to cleanse, standardize, and model raw instrument and production data into reliable, well structured datasets.

-

Build and support datasets and data models used by operational dashboards, analytics, process monitoring, troubleshooting, and governed AI enabled workflows.

-

Implement orchestration, testing, monitoring and alerting so that data failures, freshness issues, schema changes, and incomplete processing are identified early.

-

Implement data validation and quality checks to ensure datasets are accurate, complete, and reliable.

-

Document pipelines, data models, and datasets to support reproducibility and compliance with ISO, CLIA, CAP, NYS, GMP, and FDA requirements.

-

Continuously improve your technical skills and the team's engineering practices.

Required Qualifications::

-

Degree in Computer Science, Mathematics, Software Engineering, Data Science, Life Sciences, Physics or similar field.

-

1+ years of relevant professional, internship, academic, or project experience in data engineering, analytics engineering, software development, or a related field, or equivalent practical experience.

-

Proficiency in SQL.

-

Working proficiency with one or more programming languages, such as Python, Rust, C++, or similar.

-

Basic understanding of ETL or ELT pipelines, relational databases, and structured or semi-structured data.

-

Strong attention to detail and a commitment to data quality, reliability and accuracy.

-

Ability to collaborate effectively in teams of technical and non-technical individuals, and comfortable working in a rapidly changing environment with dynamic objectives and fast iteration.

-

Ability to investigate technical problems methodically, continuously learn and communicate clearly.

-

A highly analytical mindset and eagerness to solve technical problems.

Preferred Qualifications::

-

Familiarity with data pipeline orchestration and transformation tools such as Airflow, dbt, or comparable technologies.

-

Familiarity with cloud data platforms, object storage and warehouses such as AWS S3, Redshift, Glue, Snowflake or comparable technologies.

-

Familiarity integrating AI/agentic tooling into the data engineering SDLC.

-

Experience with semantic data modeling, data lineage, and automated data quality testing. 

-

Familiarity with statistical methods or basic process analytics.

-

Exposure to manufacturing, clinical laboratory operations, diagnostics, or biotechnology.

-

Experience with version control systems such as Git and collaborative development practices.

-

Basic understanding of APIs, file transfers, networking and system integrations.

This role may be eligible for other forms of compensation, including an annual bonus and/or incentives, subject to the terms of the applicable plans and Company discretion. This range reflects a good-faith estimate of the range that the Company reasonably expects to pay for the position upon hire; the actual compensation offered may vary depending on factors such as the candidate’s qualifications. Employees in this role are also eligible for GRAIL’s comprehensive and competitive benefits package, offered in accordance with our applicable plans and policies. This package currently includes flexible time-off or vacation; a 401(k) retirement plan with employer match; medical, dental, and vision coverage; and carefully selected mindfulness programs.

GRAIL is an equal employment opportunity employer, and we are committed to building a workplace where every individual can thrive, contribute, and grow. All qualified applicants will receive consideration for employment without regard to race, color, religion, national origin, sex, gender, gender identity, sexual orientation, age, disability, status as a protected veteran, , or any other class or characteristic protected by applicable federal, state, and local laws. Additionally, GRAIL will consider for employment qualified applicants with arrest and conviction records in a manner consistent with applicable law and provide reasonable accommodations to qualified individuals with disabilities. Please contact us at rc@grailbio.com if you require an accommodation to apply for an open position.

GRAIL maintains a drug-free workplace. We welcome job-seekers from all backgrounds to join us!

Source: Grailbio careers (Lever)

Similar roles