- Solid hands-on experience in Python Development.
- Extensive experience with PySpark / Apache Spark.
- Expertise in Data Engineering, ETL development, and data processing frameworks. (OR) Big Data & eco systems knowledge, distributed storage like MINIO, ADLS using S3, Iceberg tables
- Strong knowledge of SQL and database concepts.
- Experience working with large-scale datasets in distributed environments.
- Understanding of Data Warehousing concepts and dimensional data modeling.
- Exposure to cloud platforms such as Azure, AWS, or GCP.
- Experience with version control tools such as Git.