Data Scientist / Bioinformatics Engineer (Hyderabad)

Data Scientist / Bioinformatics Engineer (Hyderabad)

06 Aug
|
Orbitnexa Technologies
|
Hyderabad

06 Aug

Orbitnexa Technologies

Hyderabad

Role Overview OrbitNexa is looking for an experienced Data Scientist / Bioinformatics Engineer with proven hands-on experience in building human genome bioinformatics pipelines and a track record of successfully delivering bioinformatics projects. This is an excellent chance for someone who is passionate about developing and optimizing human genome analysis pipelines using AWS HealthOmics and other AWS services. You'll work on cutting-edge human genomics and bioinformatics projects, collaborate with healthcare and research teams, and contribute to building scalable, efficient bioinformatics solutions for human genome analysis.

You'll have the opportunity to work with AWS HealthOmics, Nextflow, WDL, and modern cloud-based bioinformatics tools to process and analyze large-scale human genomic data. We are seeking candidates who have demonstrated ability to deliver end-to-end human genome bioinformatics solutions from design to production deployment. Candidates with strong human genome pipeline experience will be considered even if they have limited or no direct AWS HealthOmics experience, as we value hands-on bioinformatics expertise.

In This

Role, You Will Develop and optimize human genome bioinformatics pipelines using AWS HealthOmics (Omics Workflows)

Design and implement scalable human genomic data processing workflows on AWS

Build and maintain Nextflow and WDL (Workflow Description Language) pipeline definitions for human genome analysis

Integrate human genome bioinformatics tools and algorithms into cloud-native workflows

Process and analyze large-scale human genomic datasets (WGS, WES, RNA-seq, etc.)

Build pipelines for human genome variant calling, annotation, and interpretation

Optimize pipeline performance, cost, and resource utilization on AWS

Implement data quality control, validation,



and error handling in pipelines

Work with AWS HealthOmics Sequence Stores, Annotation Stores, and Reference Stores

Design and implement data processing workflows using AWS Batch, Step Functions, and Lambda

Build data transformation and ETL pipelines for genomic data

Implement data storage and retrieval strategies using S3, DynamoDB, and HealthOmics stores

Develop monitoring, logging, and alerting solutions using CloudWatch

Create data visualization dashboards and reports for bioinformatics results

Collaborate with research scientists and clinicians to understand requirements

Document pipeline architecture, workflows, and best practices

Implement version control and reproducibility for bioinformatics pipelines

Optimize costs and resource allocation for large-scale genomic processing

Integrate with laboratory information systems (LIS/LIMS) and healthcare systems

Ensure compliance with healthcare data regulations (HIPAA, GDPR) in data processing

Deliver end-to-end bioinformatics projects from requirements gathering to production deployment

Participate in code reviews, sprint planning, and team discussions

Stay updated with latest bioinformatics tools, AWS HealthOmics features, and genomics research To Be Successful You Will Bachelor's or Master's degree in Bioinformatics, Computational Biology, Data Science, Computer Science, or related field (PhD preferred)





3-5 years of professional hands-on experience in bioinformatics with proven track record of successfully delivering human genome bioinformatics projects from conception to production

2+ years of hands-on experience building and optimizing human genome bioinformatics pipelines with demonstrated project outcomes and strong portfolio demonstrating completed projects with measurable results

Deep expertise in human genome bioinformatics: variant calling (germline & somatic), annotation, interpretation, and analysis using tools like GATK, DeepVariant, Strelka2, Mutect2, VEP, ANNOVAR, BWA, and following GATK best practices

Strong proficiency in Python and/or R for bioinformatics analysis, with experience in Nextflow and/or WDL for pipeline development

Deep understanding of human genomic data formats (FASTQ, BAM, CRAM, VCF, gVCF) and reference assemblies (GRCh37, GRCh38) with knowledge of annotation databases (dbSNP, ClinVar, gnomAD)

Experience with AWS HealthOmics or similar cloud genomics platforms (preferred, but candidates with strong human genome pipeline experience will be considered even without direct HealthOmics experience)

Understanding of AWS services (Batch, Step Functions, Lambda, S3, DynamoDB, CloudWatch) or similar cloud platforms, with experience in containerization (Docker) and high-performance computing for genomics

Experience with version control (Git), data quality control, validation, QC metrics, and statistical analysis for genomic data

Valid cloud certification in AWS, Azure, or GCP (preferred)

Strong problem-solving, analytical thinking, and excellent communication skills with ability to work independently and in a team

Experience with healthcare/healthtech domain, HIPAA-compliant systems, clinical genomics, or precision medicine (preferred)

📌 Data Scientist / Bioinformatics Engineer (Hyderabad)
🏢 Orbitnexa Technologies
📍 Hyderabad

Reply to this offer

Impress this employer describing Your skills and abilities, fill out the form below and leave Your personal touch in the presentation letter.

Subscribe to this job alert:

Get the latest job offers by email for: data scientist / bioinformatics engineer (hyderabad) / hyderabad

Subscribe to this job alert:

Get the latest job offers by email for: data scientist / bioinformatics engineer (hyderabad) / hyderabad