- Assist in developing and maintaining data pipelines using Databricks (PySpark/SQL).
- Support the design, training, and evaluation of AI/ML models , including Generative AI use cases.
- Contribute to automation scripts, data transformations, and workflow optimization.
- Create or support dashboards and visualizations using Power BI/Tableau (preferred).
- Collaborate with engineers to maintain high-quality data standards.
- (Nice to Have) Support CI/CD pipelines for data and ML workflows.
Requirements:
- Knowledge of Python , PySpark , or SQL.
- Curiosity or experience in Generative AI (LLMs, embeddings, vector stores).
- Awareness of Agentic AI concepts (AI agents, tool orchestration) is a plus.
- Familiarity with BI/reporting tools such as Power BI or Tableau (preferred).
- Bonus: Exposure to Darts, Stats Forecast, ML-Flow, GitHub , Azure DevOps , or CI/CD workflows.
- (Nice to have) Basic understanding of Machine Learning concepts and frameworks like scikit‑learn, TensorFlow, or PyTorch.
- Students currently pursuing or recently graduated with a degree in Computer Science, Data Science, Engineering , or related fields.
- Strong analytical and problem‑solving skills.
- Self‑starter with a passion for learning current technologies.
- Good communication, organization, and teamwork abilities.