We are looking for a Software Engineer to join our Enterprise Data Engine EDE Team in Bangalore The team consists of 30 colleagues and is reporting to the EDE Leadership If you love to code with Python Py-Spark we d love to hear from you We are looking for someone who has experience with Snowflake or any database like Oracle SQLServer Teradata etc and ideally has worked in a cloud environment It will be great if you have experience or familiarity with any ETL tool like Abinitio SSIS Informatica Talend or Pentaho would be added advantage as well Don t feel discouraged to apply if you don t - we are a big team of experienced Data Engineers so if you d like to tap into a new area we d be happy to provide you with all the support you need to become a confident Data Engineer About You - experience education skills and accomplishments Bachelor s degree in Computer Science Information Systems or related field with 2 years of experience in Data processing using Python Py-Spark At least 2 years of experience working on data processing by building data pipelines At least 1 years of experience with any relational databases like Oracle SQLServer Teradata etc Having exposure to any Bigdata platforms AWS etc It would be great if you also had Experience with Python Py-Spark Experience with any ETL tools relational databases Experience with Data pipelines development and testing Experience with writing SQL queries Stored Procedures Knowledge and experience in agile software development e g SCRUM What will you be doing in this role As a member of Data Engineering Team you ll Step into a key role on an expanding data engineering team to build our data platforms data pipelines and data transformation capabilities Define and implement our data platform strategy on Cloud have a meaningful impact on our customers and working in our high energy innovative fast-paced Agile culture in a leading tech company in the healthcare space Drive rapid prototyping and development with Product and Technical teams in building and scaling new high-value medical data capabilities to enable products and insights Interface with other technology teams to extract transform and load data from a wide variety of data sources using Apache suite airflow spark SQL Python ETL and AWS big data technologies Creation and support of batch and real-time data pipelines and ongoing data monitoring and validation built on AWS Snowflake Apache technologies for medical data from many different sources Continual research of the latest big data technologies to provide new capabilities and increase efficiency Help continually improve ongoing reporting and analysis processes through automation to enable self-service support for internal and external customers Product you will be developing We are building real world data platform includes more than 300 million longitudinal covered lives 2 million healthcare providers and 98 of payers in the United States with visibility into 100 million electronic health records EHR Our disease analysts market specialists and data scientists connect claims and EHR data delivering valuable assets to answer life science team s hardest questions We use Python with spark Airflow Snowflake as a Database and AWS cloud About the Team Enterprise Data Engine is a team of 20 people with Architects Bigdata Developers ETL Developers QA engineers based out of India US Canada Hours of Work 40-hours per week permanent full time position At Clarivate we are committed to providing equal employment opportunities for all qualified persons with respect to hiring compensation promotion training and other terms conditions and privileges of employment We comply with applicable laws and regulations governing non-discrimination in all locations