Who we are
Born in 2014, Yoti is a digital identity and biometric technology company that makes it safer for people to prove who they are. The Yoti app was designed with privacy at its core, giving people a secure way to prove their identity and share third-party credentials with organisations and other people.
Today, we have over twenty million app downloads around the world. We’ve expanded our offering to a suite of business solutions that span identity verification, age verification and estimation, e-signing, AI anti-spoofing technologies and we continue to think of innovative new offerings.
From day one, we’ve been working to fix an outdated identity system. This is not a journey we make on our own but with policy advisors, think tanks, researchers, academics, humanitarian bodies, our users and everyday people. We are committed to solving identity problems through grassroots research and social purpose initiatives.
About the Team and the Work
The BI Engineering team owns the acquisition, storage, modelling and provisioning of data that powers reporting, billing, product decisions and, increasingly, operational features consumed by systems outside BI. Our Redshift-centred warehouse handles hundreds of millions of events a day across a wide suite of products, and the roadmap for the next 18-24 months is the longest and most ambitious it has ever been. We are leveraging AI Workflows and agents across the team for ETL monitoring, query optimisation, code generation, documentation, and performance monitoring.
You'll join at a point where the team is moving beyond "keep the pipeline running" into four parallel streams of work:
- Redshift Optimisation: Optimising cluster and table configurations, including distribution, sort keys, and compression, to improve query performance. Analysing long-running queries and applying query-level optimisations, while leveraging data aggregation and retention strategies to improve performance, reduce storage, and manage data growth effectively.
- Data Aggregation & Retention: Designing and building roll-up layers that preserve analytical value while enabling safe deletion of raw-level data. This includes schema design, backfills, data reconciliation, and ensuring backward compatibility for existing reports and consumers.
- Data provision: Enabling teams to access data by creating curated, fit-for-purpose datasets and APIs for product, operations, and integrator-facing systems. These externally consumed solutions require high availability and strong data reliability.
- Infrastructure and Performance. Strengthening pipeline monitoring through automated setup and built-in notifications,
while proactively identifying query optimisation opportunities and cost-efficiency improvements. Defining solutions to manage data limitations and retention, with data volume remaining a key constraint shaping the overall roadmap.
What will you be doing at Yoti?
- Own Redshift performance and capacity, balancing reliability, query performance, scalability, and cost.
- Design and deliver data aggregation and retention solutions, including roll-ups, reconciliation, backfills, and safe removal of historical source data.
- Build data-provisioning solutions for product, operational, and integrator-facing consumers, ensuring reliable and accurate access to curated data.
- Build and maintain reliable ingestion and ETL/ELT pipelines, including third-party API integrations and failure handling.
- Establish data quality, reconciliation, and monitoring practices for critical data flows and reporting.
- Apply AI and automation to improve engineering workflows, monitoring, optimisation, and delivery.
- Make informed trade-offs between infrastructure improvements, feature delivery, and compliance priorities.
Requirements / What are we looking for? (5+ years of experience required)
Essential - Redshift depth. Deep, hands-on Redshift experience is the single most important requirement. You should be able to talk credibly about:
- Distribution and sort key design, query optimisation, WLM/Auto WLM, concurrency scaling, and Short Query Acceleration.
- Vacuum, Analyse, table maintenance, and managing bloat and unsorted data at scale.
- Redshift Spectrum, external schemas, S3 integration, and cross-cluster data movement using Data Sharing, UNLOAD/COPY, and snapshots.
- Capacity planning, RA3 sizing, workload isolation, and cost optimization.
Desirable
- Strong SQL across Redshift and PostgreSQL, including complex queries, window functions, and query tuning.
- Robust Python skills for building ingestion, ETL/ELT, and reconciliation frameworks. Node.js or Go is a plus.
- Hands-on AWS experience across S3, MWAA/Airflow, Lambda, Glue, Athena, IAM, Cloud Watch, RDS, and Open Search.
- Strong analytical data modelling skills, including dimensional modelling, SCDs, and scalable reporting structures.
- Experience integrating REST APIs, including authentication, pagination, retries,
idempotency, and rate-limit handling.
- Experience with containerisation, orchestration, and CI/CD for data workflows.
Interview Process
Stage 1 - Call with a Talent acquisition Team member (30 minutes)
Stage 2 - Call with Data Engineer & Chief of Product Operations (60 minutes)
Stage 3 - Technical Evaluation
What’s in it for you?
- Flexible working (Core working hours, Hybrid working)
- Performance based discretionary annual bonus
- LTIP (Long term incentive plan)
- Medical Insurance cover of INR 5 lakhs
- Life Insurance / Accidental cover
- Gratuity as per law
- PF as per law
- 18 days paid leave + 6 sick day leave (Annually)
- 10 declared holidays (Annually)
- 5 fully paid days of Selfie Time - for your own personal development, volunteering, charity events, etc
- Quarterly Team and company activities, Social clubs.
- Continuous learning opportunities (Annual Training budgets, conferences etc)
This is a great opportunity to join a company that is leading the way for innovative and responsible identity verification. We’re looking for people who can adapt to a fast-paced environment, as well as champion our brand and what we stand for. We value a positive attitude and people who have a collaborative, creative and transparent approach to solving problems.
AI Usage during the recruitment process
Please read our AI Usage in Recruitment policy to know more about how Yoti uses AI in the recruitment process and our stance on how candidates can use AI during the interview process.
We believe in equal opportunities
It takes a diverse community of passionate, talented and committed people to build a simpler, more secure way of proving identity. We’re an equal opportunity employer, so we welcome applications from people of all backgrounds, with different outlooks and experiences.
We are proud to be a Disability Confident employer and we’re committed to making our recruitment process as inclusive and accessible as possible. If you have a disability or long-term condition and need any adjustments or support during the application or interview process, please let us know — we’ll do everything we can to support you and to enable you to bring your best self to our hiring process.
Pre-employment checks
If your application is successful please be aware that as part of our pre-employment checks, we will check identity and proof of education. If you have previously been employed elsewhere, we will also carry out - Employment, Credit and Database Checks.
For more information about how we manage your data please read our Applicant Privacy Notice.
📌 BI Data Engineer (India)
🏢 Yoti
📍 India