Job DescriptionRequirements
N
n
Solid understanding of AWS architecture best practices, scalability, security, and cost optimization strategies. N
Robust hands-on experience with AWS services including Lambda, Glue ETL, Athena, S3, DynamoDB, Step Functions, EventBridge, SNS, and SQS. N
Deep experience in Apache Spark (PySpark/Scala) development, unit testing, and performance optimization. N
Solid Python programming skills using libraries such as pandas, requests, json, and awswrangler. N
Experience on Apache Kafka and Confluent Kafka. N
Experience designing and optimizing data lakes using Apache Iceberg, including compaction and Iceberg optimization techniques. N
Hands-on experience integrating and optimizing Starburst/Trino,
including connecting Starburst from Lambda and Glue ETL jobs. N
Experience with NoSQL databases such as DynamoDB, MongoDB. N
Experience working withdata formats including Avro, Parquet, JSON, XML, and CSV. N
Comfortable challengingyour peers and leadership team. N
Can prove yourself quickly and decisively. N
Excellent communicationskills and Good Customer Centricity N
nDesign and optimize data lakes using Apache Iceberg on AWS, including table compaction and Iceberg performance tuning.
📌 Hiring: Data Engineer Pune
🏢 EXL
📍 Pune