- 4 to 8 + Years of experience using Python and Pyspark.
- Solid proficiency in Python programming.
- Hands-on experience with PySpark and Apache Spark.
- Knowledge of Big Data technologies (Hadoop, Hive, Kafka, etc.).
- Experience with SQL and relational/non-relational databases.
- Familiarity with distributed computing and parallel processing.
- Understanding of data engineering best practices.
- Experience with REST APIs, JSON/XML, and data serialization.
Exposure to GCP services and cloud computing environments