13 Sep
|
Virtual Labs
|
India
13 Sep
Virtual Labs
India
Data Engineer – Databricks | Pharma DomainLocation: India (Remote)Experience: 4+ YearsEmployment Type: Contract / Full-TimeDomain: Pharma Domain – MandatoryJob SummaryWe are looking for an experienced Data Engineer with solid hands-on experience in Databricks, PySpark, SQL, CI/CD, and Version Control. The ideal candidate should have experience working on Pharma domain projects and be comfortable developing and maintaining scalable data pipelines.Key ResponsibilitiesDesign, develop, and maintain data pipelines using Databricks.Work with Databricks Notebooks and Databricks Pipelines for data processing and transformation.Develop data transformation and processing workflows using PySpark.Write optimized SQL queries for data extraction, transformation, and analysis.Build and maintain reliable ETL/ELT pipelines for large datasets.Implement data quality, validation,
and error-handling processes.Work with CI/CD pipelines for deployment and release management.Use Git/version control for source-code management and collaborative development.Collaborate with business and technical teams to understand data requirements.Follow security, governance, documentation, and data engineering best practices.Mandatory SkillsPharma Domain Experience – MUSTDatabricks – MUSTDatabricks NotebooksDatabricks PipelinesPySpark – Strong/basic hands-on experienceSQL – StrongCI/CDGit / Version ControlStrong understanding of ETL/ELT and data pipeline developmentPreferredExperience with cloud platforms such as Azure / AWSExperience working with large-scale datasetsUnderstanding of data quality and data governanceExperience working in Agile development environmentsImportantCandidates must have direct Pharma domain/project experience. Healthcare or non-Pharma experience will not be considered.
📌 Data Engineer – Databricks | Pharma Domain(Must) (India)
🏢 Virtual Labs
📍 India