16 Sep
|
Infosys
|
Bengaluru
Job Title: PySpark Developer
Experience: 23 Years
Location: PAN India
Employment Type: Full time
Job Summary
We are seeking a skilled PySpark Developer with 23 years of experience in designing, developing, and optimizing large-scale data processing solutions. The ideal candidate should have hands-on experience with PySpark, Python, Spark SQL, ETL development, and Big Data technologies. The role involves building scalable data pipelines, transforming large datasets, and supporting analytics and reporting requirements.
Key Responsibilities
- Design, develop, and maintain scalable data processing applications using PySpark.
- Build and optimize ETL/ELT pipelines for structured and unstructured data.
- Develop data transformation logic using Spark SQL and DataFrames.
- Work with large-scale datasets in distributed computing environments.
- Collaborate with Data Engineers, Data Analysts, and Business stakeholders to understand data requirements.
- Monitor, troubleshoot, and optimize Spark jobs for performance and reliability.
- Ensure data quality,
integrity, and consistency across data pipelines.
- Participate in code reviews and follow coding best practices.
- Support deployments, enhancements, and production issue resolution.
- Create technical documentation and maintain operational procedures.
Required Skillsa
- 23 years of experience in PySpark development.
- Strong programming skills in Python.
- Hands-on experience with Apache Spark, Spark SQL, DataFrames, and RDDs.
- Experience in ETL development and data integration projects.
- Good understanding of data warehousing concepts.
- Experience working with relational databases such as SQL Server, Oracle, MySQL, or PostgreSQL.
- Strong SQL querying and performance tuning skills.
- Knowledge of Linux/Unix environments and shell scripting.
- Understanding of version control systems such as Git.
- Strong analytical and problem-solving skills.
📌 PySpark Developer (2 To 3 Years) (Bengaluru)
🏢 Infosys
📍 Bengaluru