09 Sep
|
Algoworks
|
India
Role: Lead Backend Engineer Location: Remote, India
Experience: 8+ Years
Algoworks is an award-winning artificial intelligence, engineering services and experience transformation firm with offices across the United States, Europe, South America and India. Our work combines human-centered design, engineering excellence and AI-powered capabilities to solve complex challenges with clarity and precision. Innovation, particularly in the responsible application of AI, is embedded in how teams approach problem-solving and continuous improvement.
Through collaboration, accountability and a focus on results, Algoworks operates at the intersection of technology and people, building not only advanced systems but strong global teams that elevate performance and create lasting impact.
Follow the video below to know about us!
We are seeking a Lead Backend Engineer with strong production-level Python experience and deep expertise in large-scale search platforms, preferably SolrCloud/Lucene , to help integrate and scale two large healthcare data platforms.
The role will focus on enabling enriched emergency medical records from one platform to be consumed effectively by another. The platform currently manages approximately 500 TB of data and processes around 50 GB of new data daily , with the initial implementation focused on one data source and plans to onboard approximately ten additional sources.
You will work closely with a small, senior engineering team and take ownership of complex technical challenges across data ingestion, enrichment, modeling, indexing, APIs and search architecture.
Data architecture and lifecycle
Own the lifecycle of data attributes across the technology stack, from definition and modeling through ingestion, indexing, APIs and presentation.
Define scalable data models and attribute structures to support evolving healthcare data requirements.
Develop and maintain data ingestion and enrichment workflows using Python.
Work closely with Java/Spring services to expose and serve enriched data through APIs.
Collaborate with front-end teams to ensure data is effectively presented and consumed.
Design and implement a metadata-driven framework for managing data attributes across the technology stack.
Create a single attribute definition that can generate the required configurations and artifacts currently maintained across multiple components.
Establish reusable patterns that simplify the introduction and management of new attributes.
Improve consistency, maintainability and operational efficiency across the data platform.
Data source onboarding
Design repeatable processes and frameworks for onboarding new data sources.
Reduce data-source onboarding time from weeks or days to hours wherever practical through automation and metadata-driven configuration.
Develop scalable ingestion, transformation, enrichment and indexing workflows.
Support the initial implementation for one data source and subsequent onboarding of approximately ten additional sources.
Solr architecture and search performance
Diagnose and resolve search performance, indexing and scalability challenges.
Establish best practices for Solr schema design, collections, replicas, shards and query optimization.
Ensure the search architecture can scale reliably as data volumes and sources increase.
Backend engineering
Develop scalable, maintainable and production-ready backend components using Python.
Contribute confidently to Java/Spring-based services and APIs.
Design robust data-processing pipelines for ingestion, enrichment and serving.
Take ownership of technical solutions from architecture and design through implementation, testing and production deployment.
Work closely with a small, experienced engineering team in an autonomous and fast-moving environment.
Strong production-level experience with Python .
Deep hands-on experience with Apache Solr/Lucene , preferably SolrCloud 9 or equivalent large-scale deployments.
Solid Elasticsearch experience may be considered for candidates who are willing and able to transition to Solr.
Working knowledge of Java and the ability to contribute confidently to Spring-based backend services.
Strong SQL and data-modeling skills.
Experience designing scalable data-ingestion and data-processing pipelines.
Experience with data enrichment, indexing and search-serving workflows.
Strong understanding of distributed systems and large-scale data platforms.
Demonstrated ability to take ownership of complex engineering problems from design through production deployment.
Strong Python development experience.
Hands-on experience with Solr/Lucene at scale.
Experience with Elasticsearch can be considered with willingness to transition to Solr.
Good working knowledge of Java and Spring.
Strong SQL and data-modeling skills.
Experience building scalable data ingestion and processing workflows.
Experience working with large-scale data platforms and production environments.
Experience working with healthcare data .
Knowledge of NEMSIS data and emergency medical records.
Experience with Databricks .
Front-end development experience.
Experience building metadata-driven platforms or configuration frameworks.
Experience working with very large-scale datasets and distributed search infrastructure.
Ability to balance long-term architecture with short-term delivery requirements.
Passion for building scalable, maintainable and production-ready systems.
Elasticsearch solr NEMSIS Databricks Python
📌 Hiring: Team Lead Backend Engineer (India)
🏢 Algoworks
📍 India