We are looking for an experienced Python PySpark Developer with strong expertise in big data processing, ETL development, and data engineering. The candidate will be responsible for building scalable data pipelines and optimizing data processing frameworks.
Key Responsibilities
- Design and develop scalable data pipelines using Python and PySpark.
- Develop ETL processes for large-scale data processing.
- Work with structured and unstructured data sources.
- Optimize Spark jobs for performance and scalability.
- Collaborate with business and technical stakeholders.
- Ensure data quality, reliability, and performance.
Required Skills
- 5+ years of IT experience.
- Solid hands-on experience in Python and PySpark.
- Experience with Spark, Hadoop, Hive, or related Big Data technologies.
- Strong SQL skills.
- Experience in data transformation and ETL development.
- Good problem-solving and communication skills.