Data Engineer II- Life Sciences

Added
37 minutes ago
Type
Full time
Salary
Upgrade to Premium to se...

Related skills

aws data engineering etl data pipeline sql

πŸ“‹ Description

  • Build and maintain the Python and PySpark pipelines behind the CTMS trial data pipeline intake
  • Develop the transformation logic that maps raw trial and customer data to H1's internal data
  • Write and tune SQL against large datasets to investigate data questions, validate pipeline output
  • Turn around customer-changes quickly, scoping requests as they arrive, ship changes that hold up
  • Partner with clinical SMEs to translate domain expertise into concrete data rules, then walk them
  • Build data quality checks, validation logic, and a reconciliation process that let non-engineers

🎯 Requirements

  • 3+ years of experience in software data engineering, with strong experience in Python
  • Experience building and maintaining production-grade data pipelines in Python
  • Hands-on experience with PySpark or a similar distributed data processing system
  • Strong SQL skills, including work with large, messy, multi-source data sets
  • Strong understanding of software quality practices: testing, code review, documentation, and CI/CD
  • Experience working with cross-functional and non-technical stakeholders

🎁 Benefits

  • Full suite of health insurance options with generous remuneration
  • Pre-planned company-wide wellness holidays
  • Retirement options
  • Health & charitable donation stipends
  • Impactful Business Resource Groups
  • Flexible work hours & opportunity to work from anywhere
Share job

Meet JobCopilot: Your Personal AI Job Hunter

Automatically Apply to Data Jobs. Just set your preferences and Job Copilot will do the rest β€” finding, filtering, and applying while you focus on what matters.

Related Data Jobs

See more Data jobs β†’