- Python
- PySpark
- SQL
- Data Engineering
- GenAI, LLMs or RAG
- Azure/AWS (Any Cloud)
Job Responsibilities
- Build and optimize scalable data pipelines using Python and PySpark.
- Develop ETL/ELT workflows and process large-scale datasets.
- Work on GenAI applications using LLMs and/or RAG frameworks.
- Integrate AI solutions with enterprise data platforms.
- Collaborate with cross-functional teams to deliver data-driven solutions.
Preferred Candidate Profile
- 5+ years of experience in Data Engineering.
- Robust hands-on experience with Python, PySpark, and SQL.
- Experience with GenAI, LLMs, or Retrieval-Augmented Generation (RAG).
- Exposure to cloud platforms (Azure or AWS) is preferred.
- Good problem-solving and communication skills.