Role Description
: Junior Data Scientist Experience 2–4 years Role Overview We are looking for a Junior Data Scientist with strong fundamentals in Python, statistics, data processing, data cleaning, analytics, and machine learning. The candidate will be responsible for preparing and analyzing datasets, building basic predictive models, and exposing analytical or machine learning capabilities through FastAPI-based REST services. The primary focus of this role is strong data science fundamentals rather than advanced Generative AI expertise.
Key Responsibilities
- Understand and analyze data from databases, CSV, Excel, JSON, and APIs.
- Clean and prepare data by handling missing values, duplicates, outliers, invalid values, and inconsistent formats.
- Perform data transformation, aggregation, filtering, joining, and feature creation using Python.
- Conduct exploratory data analysis to identify trends, patterns, correlations, and anomalies.
- Apply statistical methods to interpret data and support business decisions.
- Build basic machine learning models for regression, classification, clustering, or forecasting use cases.
- Evaluate models using appropriate metrics and clearly document the results.
- Develop reusable Python modules for data processing, analytics, and model inference.
- Build REST APIs using FastAPI to expose data-processing and model capabilities.
- Write basic unit tests, handle errors, and maintain clear technical documentation.
- Collaborate with AI engineers, data engineers, backend developers, and domain teams. Must-Have Skills
- Strong Python programming fundamentals.
- Hands-on experience with Pandas and NumPy.
- Positive understanding of descriptive statistics, probability, correlation, distributions, sampling, and hypothesis testing.
- Practical experience in data cleaning, preprocessing, transformation, and validation.
- Strong exploratory data analysis and data visualization skills.
- Basic machine learning experience using Scikit-learn.
- Understanding of train-test split, feature engineering, model evaluation, overfitting, and cross-validation.
- Working knowledge of SQL, including joins, filtering, grouping, and aggregation.
- Basic hands-on experience developing REST APIs using FastAPI.
- Understanding of HTTP methods, request and response models, status codes, and error handling.
- Familiarity with Git, debugging, virtual environments, and unit testing. Good-to-Have Skills
- Exposure to time-series forecasting or advanced statistical analysis.
- Basic knowledge of cloud platforms such as AWS, Azure, or GCP.
- Familiarity with Docker, CI/CD, and application deployment.
- Exposure to Generative AI concepts such as LLMs, embeddings, vector databases, or RAG.
- Familiarity with LangChain, LangGraph, or similar frameworks.
- Experience working in healthcare, retail, supply chain, inventory, agriculture, or logistics domains.
Educational Qualification
Bachelor’s or Master’s degree in Computer Science, Data Science, Statistics, Mathematics, Artificial Intelligence, Engineering, or a related field.
Core Skill Tags
Python, Data Science, Statistics, Data Processing, Exploratory Data Analysis, Machine Learning, SQL, FastAPI
Skills Python, FastAPI, Data Analysis, Machine Learning, Data Validation, Generative AI, LLMs, LangChain, LangGraph, SQL, Statistics
📌 Associate II - Data Science (Kochi)
🏢 UST
📍 Kochi