25 Aug
|
Task Staffing
|
Chennai
25 Aug
Task Staffing
Chennai
Role: AWS Data Engineer/ Data Bricks/Pyspark
Location: Chennai, Bangalore
Experience: 5-10 years
Interview: Face-to-face
We are seeking an experienced AWS Data Engineer with strong Databricks and PySpark expertise to design, build, and maintain scalable data pipelines and analytical platforms. The ideal candidate will collaborate with data scientists, analysts, and engineering teams to enable reliable ETL/ELT processes, optimize performance, and ensure data quality across cloud-native architectures.
Responsibilities
- Design, develop, and maintain scalable ETL/ELT pipelines using Databricks and PySpark on AWS.
- Implement data ingestion, transformation, and enrichment workflows from diverse sources into cloud data stores.
- Develop and maintain data models and data warehousing solutions to support analytics and reporting.
- Optimize Spark jobs and cluster configurations for performance and cost efficiency.
- Manage data storage and lifecycle using Amazon S3, Delta Lake, and other AWS services.
- Integrate streaming and real-time data processing using Kafka or equivalent technologies where required.
- Collaborate with stakeholders to define data requirements, SLAs, and monitoring/alerting strategies.
- Implement CI/CD pipelines, automated testing, and deployment for data workflows.
- Document data pipelines, schemas,
and operational runbooks; participate in on-call rotation as needed.
Qualifications
- Bachelor's degree in Computer Science, Engineering, or a related technical discipline, or equivalent experience.
- 3+ years of professional experience building data pipelines and analytics solutions on AWS.
- Proven hands-on experience with Databricks, PySpark, and Apache Spark in production environments.
- Robust proficiency in Python and SQL, including query optimization and data modeling best practices.
- Experience with AWS data services such as Amazon S3, Amazon Redshift, Amazon EMR, and AWS Glue.
- Familiarity with streaming platforms (e.g., Apache Kafka) and orchestration tools (e.g., Apache Airflow).
- Experience with infrastructure-as-code tools such as Terraform and containerization technologies like Docker.
- Excellent problem-solving, communication, and collaboration skills; ability to work in cross-functional teams.
Skills
- AWS
- Databricks
- PySpark
- Apache Spark
- Python
- SQL
- ETL
- Data Modeling
- Data Warehousing
- Amazon S3
- AWS Glue
- Amazon Redshift
- Amazon EMR
- Delta Lake
- Apache Kafka
- Apache Airflow
- Terraform
- Docker
- Git
- Linux
- Performance Tuning
- CI/CD
Pay: ₹1,000,000.00 - ₹3,500,000.00 per year
Work Location: In person
📌 AWS Data Engineer (Chennai)
🏢 Task Staffing
📍 Chennai