Optum is a global organization that delivers care, aided by technology to help millions of people live healthier lives. The work you do with our team will directly improve health outcomes by connecting people with the care, pharmacy benefits, data and resources they need to feel their best. Here, you will find a culture guided by inclusion, talented peers, comprehensive benefits and career development opportunities. Come make an impact on the communities we serve as you help us advance health optimization on a global scale. Join us to start Caring. Connecting. Growing together.
This role requires hands-on expertise in Spark, Scala, Python, SQL, and Azure Databricks, along with solid knowledge of data orchestration, version control, and CI/CD practices.
- Primary Responsibilities:
- Design, develop, and optimize large-scale data processing pipelines using Apache Spark (Scala/Python) and Azure Databricks
- Perform data analysis, feature engineering, and dataset preparation on structured and unstructured data
- Build and maintain ETL workflows using Apache Airflow for reliable data scheduling and orchestration
- Write effective and scalable code in Scala, Python, and Java for data transformations and analytics
- Work with relational and NoSQL databases such as SQL databases and MongoDB to support analytical use cases
- Leverage Hadoop ecosystem components for distributed data storage and processing
- Collaborate with cross-functional teams to understand business requirements and translate them into analytical solutions
- Use GitHub for source control and GitHub Actions for CI/CD automation and quality checks
- Follow Agile delivery practices using tools such as Rally for sprint planning and tracking
- Perform data validation, performance tuning, and troubleshooting in Linux-based environments
- Ensure data quality, reliability, security, and compliance with enterprise standards
- Comply with the terms and conditions of the employment contract, company policies and pro
📌 Data Scientist (Chennai)
🏢 Optum
📍 Chennai