09 Sep
|
Infosys
|
Bengaluru
Service Line
Data Analytics Unit
Responsibilities
- Develop and maintain data pipelines using PySpark
- Process and analyze large-scale datasets in distributed environments
- Design and implement ETL/ELT workflows
- Optimize Spark jobs for performance and scalability
- Work with data stored in HDFS, Hive, or cloud storage (S3, ADLS)
- Collaborate with data engineers, analysts, and business teams
- Ensure data quality, integrity, and governance
- Debug and troubleshoot data processing issues
- Automate workflows using scheduling tools (Airflow, Oozie, etc.)
- Write clean, scalable, and productive code
Technical and Professional Requirements
- Technology - Big Data - Data Processing - PySpark
Preferred Skills Educational Requirements
- Bachelor of Engineering
- BTech
- BSc
- BCA
- MTech
- MSc
- MCA
📌 PySpark Developer (Bengaluru)
🏢 Infosys
📍 Bengaluru