- We are looking for a Trino Developer who can own query engine performance lead migration of existing
- pipelines and warehouses and help establish a scalable cost efficient lakehouse architecture
- You will work at the intersection of query engineering data modeling and platform operations translating
- legacy Snowflake Spark logic into performant Trino Iceberg workloads while ensuring correctness parity and
- reliability
Key Responsibilities:
- Hands on production experience with Trino or PrestoSQL Presto
- Solid deep SQL expertise complex analytical queries window functions CTEs query optimization
- Hands on experience with Apache Iceberg or comparable open table formats Delta Lake Hudi
- including schema partition evolution and table maintenance
- Practical experience with Snowflake and or Apache Spark enough to read understand and migrate
- existing workloads
- Understanding of distributed query execution MPP architecture join distribution memory spill behavior
- partition pruning and predicate pushdown
- Experience with cloud object storage and columnar file formats Parquet ORC
- Proficiency in at least one programming language Python Java or Scala for tooling UDFs and
- automation
- Version control Git and CI CD for data pipelines
Technical Requirements:
- Experience leading a Snowflake Trino or Spark Trino migration at scale
- Trino cluster administration and deployment on Kubernetes
- Workflow orchestration Airflow Dagster or similar and dbt
- Experience building data validation reconciliation frameworks for migration parity
- Knowledge of Trino internals or connector development contributing custom connectors UDFs
- Streaming CDC ingestion into Iceberg Kafka Flink Debezium
- Data governance lineage and cost optimization tooling
Additional Responsibilities:
- Strong analytical and problem solving mindset for debugging correctness and performance issues
- Clear communication able to document migration decisions and work with analytics platform and
- business teams
- Ownership mentality with attention to data correctness and reliability
Preferred Skills:
Technology->Big Data - Data Processing->Spark->SparkSQL,Technology->Data on Cloud-DataStore->Snowflake