-
10+ years of experience in data engineering with strong proficiency in Python and PySpark.
- 3–5+ years of hands-on experience with Databricks platform on Azure, AWS, or GCP.
- Deep understanding of Apache Spark internals and distributed computing concepts.
- Solid knowledge of Delta Lake, Lakehouse architecture, and data lake management.
- Proven experience with large-scale Big Data platforms (Hadoop, Hive, HDFS, etc.).
- Proficiency in SQL and working with relational and non-relational databases.
- Solid understanding of CI/CD, version control (Git), and DevOps for data pipelines.
- Experience with data orchestration tools (e.g., Airflow, Azure Data Factory).
- Experience in data warehouse design and maintenance
- Experience in agile development processes using Jira and Confluence.
- Experience in cross-functional teams.
- Understanding on the SDLC.
- Understanding on the Agile methodologies.
- Communication with customer and producing the Daily status report.
- Should have good oral and written communication.
- Should be a good team player.
- Should be proactive and adaptive.