Pyspark Data Engineer (India)

Pyspark Data Engineer (India)

31 Jul
|
Spacevinfotech
|
India

31 Jul

Spacevinfotech

India

Role : Pyspark Data Engineer

- Hands-on expertise in designing, building, and maintaining Apache Spark pipelines in production environments.
- Proven experience building and scaling data ingestion frameworks that integrate data from multiple source systems, with a focus on reliability, reusability, and scalability.
- Deep understanding of Spark architecture (driver/executors, DAG, partitioning, shuffles, caching, cluster resource management) and experience operating pipelines at scale, including data transformations on datasets ~500 GB+.
- Strong understanding of Oracle SQL and HDFS, including handling file formats and applying appropriate data cleansing, normalization, and formatting to produce curated output datasets.
- Ability to write Python, Pyspark, and shell scripts to process, transform,



and automate data workflows, and the candidate should be good in writing application programs and automation manual data processing steps using python.

Role : PySpark Developer / Senior Data Engineer

Skills:

- Robust hands-on experience in PySpark, Python, and SQL.
- Experience designing and optimizing Spark-based ETL/ELT pipelines and data processing jobs.
- Strong understanding of BigQuery.
- Strong understanding of data quality, governance, observability, and performance tuning.
- Good collaboration, debugging, and Agile delivery skills.

📌 Pyspark Data Engineer (India)
🏢 Spacevinfotech
📍 India

Reply to this offer

Impress this employer describing Your skills and abilities, fill out the form below and leave Your personal touch in the presentation letter.

Subscribe to this job alert:

Get the latest job offers by email for: pyspark data engineer (india) / india

Subscribe to this job alert:

Get the latest job offers by email for: pyspark data engineer (india) / india