Job Description
Job Description
n6 to 8years of experience
nSkill sets: Python, Pyspark and GCP
nLocation: Hyderabad
nJob Requirements*
n
- n
- 6to 8 + Years of experience using Python and Pyspark.n
- Solid proficiency in Python programming.n
- Hands-on experience with PySpark and Apache Spark.n
- Knowledge of Big Data technologies (Hadoop, Hive, Kafka, etc.).n
- Experience with SQL and relational/non-relational databases.n
- Familiarity with distributed computing and parallel processing.n
- Understanding of data engineering best practices.n
- Experience with REST APIs, JSON/XML, and data serialization.n
- Exposure to GCP services and cloud computing environments.n
nKey Responsibilities*
n
- n
- Develop and maintain scalable data pipelines using Python and PySpark.n
- Design and implement ETL (Extract, Transform, Load) processes.n
- Optimize and troubleshoot existing PySpark applications for performance.n
- Collaborate with cross-functional teams to understand data requirements.n
- Write clean, efficient, and well-documented code.n
- Conduct code reviews and participate in design discussions.n
- Ensure data integrity and quality across the data lifecycle.n
- Integrate with cloud platforms like GCP, AWS or Azure.n
- Implement data storage solutions and manage large-scale datasets.n