Key Skill Requirements:
Solid experience in Python
Hands-on expertise with AWS services, particularly:
EMR (mandatory and most critical skill)
EC2
Lambda
Candidate should have experience with EMR performance and cost optimization
Other AWS services can be adaptable based on overall profile strength
Other notes:
Solid Real-World Data Engineering Experience
Not just theoretical knowledge.
Wants engineers who can demonstrate:
Practical implementation experience
Real production environments
Solving actual business problems
Experience working through technical challenges
Understanding why decisions were made, not just what was built
Looking for people with "practical knowledge and real-time experience."
Candidates Who:
Are naturally curious
Continuously learn
Adapt quickly to new technologies
Enjoy solving new problems
Demonstrate initiative
Show flexibility as technologies evolve
Experience Required
• Extracted: 6+ years of experience in data engineering, distributed systems, or backend platforms
Overview
Senior Data Engineer role focused on designing, developing, and optimizing large-scale data platforms and backend systems. The position emphasizes building API-driven data services, managing distributed batch data processing, and optimizing cloud workflows on AWS (with mandatory EMR). Work includes orchestration with Apache Airflow and building/maintaining pipelines using Spark, SQL, Hive, Python, and Scala.
Key Responsibilities
• Design and develop API-driven systems for managing large-scale batch data applications
• Build scalable backend services and data engineering solutions
• Develop and maintain data pipelines using Spark, SQL, Hive, Python, and Scala
• Design and optimize workflows using Apache Airflow (DAG design, scheduling,