Job DescriptionRequirements
N
- n
- Strong understanding of AWS architecture best practices, scalability, security, and cost optimization strategies. N
- Strong hands-on experience with AWS services including Lambda, Glue ETL, Athena, S3, DynamoDB, Step Functions, EventBridge, SNS, and SQS. N
- Deep experience in Apache Spark (PySpark/Scala) development, unit testing, and performance optimization. N
- Solid Python programming skills using libraries such as pandas, requests, json, and awswrangler. N
- Experience on Apache Kafka and Confluent Kafka. N
- Experience designing and optimizing data lakes using Apache Iceberg, including compaction and Iceberg optimization techniques. N
- Hands-on experience integrating and optimizing Starburst/Trino,
including connecting Starburst from Lambda and Glue ETL jobs. N
- Experience with NoSQL databases such as DynamoDB, MongoDB. N
- Experience working withdata formats including Avro, Parquet, JSON, XML, and CSV. N
- Comfortable challengingyour peers and leadership team. N
- Can prove yourself quickly and decisively. N
- Excellent communicationskills and Good Customer Centricity N
nDesign and optimize data lakes using Apache Iceberg on AWS, including table compaction and Iceberg performance tuning.
📌 Hiring: Data Engineer (Pune)
🏢 EXL
📍 Pune