Senior Data Engineer, Data Lakehouse Infrastructure

Added
3 hours ago
Type
Full time
Salary
Upgrade to Premium to se...

Related skills

sql python airflow spark iceberg

📋 Description

  • Architect and scale a data lakehouse on GCP using Iceberg, GCS, BigQuery.
  • Design, build, and optimize distributed query engines such as Trino, Spark, Snowflake.
  • Implement metadata management in Iceberg and data discovery catalogs for governance.
  • Develop and orchestrate robust ETL/ELT pipelines with Airflow, Spark, and GCP-native tools.
  • Collaborate with data scientists, engineers, and product teams to implement data solutions.

🎯 Requirements

  • 5+ years in data or software engineering, focusing on distributed data systems.
  • Experience building and scaling data platforms on GCP (storage, compute, orchestration, monitoring).
  • Strong command of query engines such as Trino, Spark, or Snowflake.
  • Experience with modern table formats like Hudi, Iceberg, or Delta Lake.
  • Strong Python skills; proficient in SQL or SparkSQL.
  • Hands-on with Airflow and streaming/batch pipelines using GCP-native services.

🎁 Benefits

  • Remote-first and async-friendly environment.
  • High ownership culture with fast iteration.
  • Teams across US timezones with some meeting overlap.
  • Opportunity to work with petabyte-scale data and AI fluency.
  • Career growth through TRM’s Engineering Levels.
Share job

Meet JobCopilot: Your Personal AI Job Hunter

Automatically Apply to Data Jobs. Just set your preferences and Job Copilot will do the rest — finding, filtering, and applying while you focus on what matters.

Related Data Jobs

See more Data jobs →