29 Aug
|
Infosys
|
Bengaluru
- Experience leading a Snowflake ® Trino or Spark ® Trino migration at scale.
- Trino cluster administration and deployment on Kubernetes.
- Workflow orchestration (Airflow, Dagster, or similar) and dbt.
- Experience building data-validation / reconciliation frameworks for migration parity.
- Knowledge of Trino internals or connector development (contributing custom connectors/UDFs).
- Streaming/CDC ingestion into Iceberg (Kafka, Flink, Debezium).
- Data governance, lineage, and cost-optimization tooling.
Hands-on production experience with Trino (or PrestoSQL/Presto).
- Solid, deep SQL expertise — complex analytical queries, window functions, CTEs, query optimization.
- Hands-on experience with Apache Iceberg (or comparable open table formats — Delta Lake, Hudi) including schema/partition evolution and table maintenance.
- Practical experience with Snowflake and/or Apache Spark — enough to read, understand,
and migrate existing workloads.
- Understanding of distributed query execution: MPP architecture, join distribution, memory/spill behavior, partition pruning, and predicate pushdown.
- Experience with cloud object storage and columnar file formats (Parquet, ORC).
- Proficiency in at least one programming language (Python, Java, or Scala) for tooling, UDFs, and automation.
- Version control (Git) and CI/CD for data pipelines. Strong analytical and problem-solving mindset for debugging correctness and performance issues.
- Clear communication — able to document migration decisions and work with analytics, platform, and business teams.
- Ownership mentality with attention to data correctness and reliability.
📌 Trino Developer (Bengaluru)
🏢 Infosys
📍 Bengaluru