- Design, develop, and deploy data pipeline for clinical domain dataset
- As an infrastructure programmer, continuously develop and support Data Scientist R-Platform and integration with various technology (Kubernetes Container, HashiCorp Vault, SAS Storage, and Data Science Work Bench)
- Design and build various reusable program components using innovative technology (NLP, AI, Python, R, etc) to transform and harmonize clinical dataset for insight generation
- Collaborate with Data Architects, Business SMEs, and Data Scientists to capture the business requirement and translate into Agile product backlog
- Serve as primary data engineer to manage and support AWS, Databricks, RStudio platform, and cloud AI based system production DevOps
- Align to best practices for coding, testing, and designing reusable code/component
- Explore new tools and technologies that will help streamline data pipeline and add new durable capability for clinical development
- Participate in sprint planning meetings and provide estimations on technical implementation
- Collaborate and communicate effectively with the product teams
Must Have Skills:
- Advanced skills in SQL, Python, and R languages programing; AWS cloud technology and databricks data lake technology stacks
- Data modeling skills, and software development lifecycle knowledge and best practices
- Learning ability of recent technology in the information field
- Skill of using DevOps CI/CD tools, such Git, Jenkins and front UI Visualization technology
📌 Sr Mgr Data Engineer (Telangana)
🏢 Amgen
📍 Telangana
Reply to this offer
Impress this employer describing Your skills and abilities, fill out the form below and leave Your personal touch in the presentation letter.