6+ years(6 year) of skilled software engineering mostly focused on the following:
- Developing ETL pipelines involving big data.
- Developing data processinganalytics applications primarily using PySpark.
- Experience of developing applications on cloud(AWS) mostly using services related to storage, compute, ETL, DWH, Analytics and streaming.
- Clear understanding and ability to implement distributed storage, processing and scalable applications.
- Experience of working with SQL and NoSQL database.
- Ability to write and analyze SQL, HQL and other query languages for NoSQL databases.
- Proficiency is writing disitributed & scalable data processing code using PySpark, Python and related libraries