Data Architect (India)

Data Architect (India)

21 Aug
|
Persistent
|
India

21 Aug

Persistent

India

Job Description

About Persistent

We are an AI-led, platform-driven Digital Engineering and Enterprise Modernization partner, combining deep technical expertise and industry experience to help our clients anticipate what’s next. Our offerings and proven solutions create a unique competitive advantage for our clients by giving them the power to see beyond and rise above. We work with many industry-leading organizations across the world, including 20 Fortune 50 companies and 4 of the 5 top banks in both the US and India, and numerous innovators across the healthcare ecosystem.

Our disruptor’s mindset, commitment to client success, and agility to thrive in the agile environment have enabled us to sustain our growth momentum. Persistent has been recognized across top industry platforms for innovation, leadership, and inclusion. We reported $1,654.4M FY26 revenue with 17.4% Y-o-Y growth. We have delivered 24 sequential quarters of growth with $436.0M in Q4 FY26 revenue, up 3.2% Q-o-Q and 16.2% Y-o-Y growth. Our 27,500+ global team members, located in 18 countries, have been instrumental in helping the market leaders transform their industries. We have been recognized as the Fastest Growing IT Services Brand Globally in the 2026 Brand Finance IT Services 25 Report. We named a Leader in the Everest Group Private Equity (PE) Services PEAK Matrix® Assessment 2026 and Software Product Engineering PEAK Matrix® Assessment 2026.

About Position:

Experience Data Modernization migration of an on premises SQL data warehouse to a target state Data Lake on Google Cloud (GCP), enabling metrics reporting, advanced analytics, and GenAI use cases (natural language querying, accelerated summarization, cross domain trend analysis) leveraging PySpark based processing, cloud native DevOps CI/CD pipelines, and containerized deployments on OpenShift (OCP) to deliver scalable, secure, and high performance data solutions. 14+ years of overall IT experience, with strong experience in Data Architecture, Big Data Engineering, Cloud Data Platforms, ETL/ELT Frameworks, Data Lake Architecture, and Enterprise Data Modernization programs.

-
Role: Data Architect
- Location: Bengaluru
- Experience: 8 to 12 Years
- Job Type: Full Time Employment

What You'll Do:

- Own the end-to-end data architecture for the IAM Data Modernization program across OCP, S3, PostgreSQL metadata repository, and GCP / BigQuery future state architecture.




- Define and guide implementation of a metadata-driven generic ETL framework supporting multiple source types including databases, files, APIs, MongoDB, Splunk, VDS, and Active Directory.
- Architect reusable metadata models using MD_ static metadata tables and OPS_ operational/runtime tables for source configuration, pipeline definition, run tracking, audit logging, checkpointing, schema drift, and notifications.
- Provide architecture direction for MVC-based orchestration, including repository layer, service layer, wrapper, controller, execution plan models, and standardized request/response contracts.
- Design scalable ingestion patterns for full load, incremental load, watermark-based processing, CDC-style patterns, batch ingestion, API ingestion, and file-based ingestion.
- Guide design of S3 data lake zones including Landing, Raw, Archive, and Error zones, with clear promotion, archival, replay, and error-handling patterns.
- Define data quality, schema validation, schema drift detection, CLOB validation, checkpoint restartability, auditability, and operational resilience standards.
- Provide technical leadership for GCP / BigQuery architecture, including S3 to BigQuery integration, schema mapping, dataset design, partitioning, clustering, performance optimization, and secure connectivity.
- Review and guide CI/CD automation using GitHub Actions, JFrog, Harness, Liquibase, Helm, OCP deployments, and environment-specific configuration management.
- Provide technical guidance to team on Python, PySpark, SQL, BigQuery, PostgreSQL, data modeling, performance tuning, and production-ready framework design.
- Support UAT and production readiness by defining validation scenarios, reconciliation strategy, performance testing approach, security validation, runbooks, and sign-off criteria.
- Partner with client stakeholders to finalize architecture decisions, design documents, Confluence documentation, implementation standards, and technical roadmaps.
- Ensure the solution is secure, scalable,



maintainable, metadata-driven, auditable, and extensible for future Splunk and Grafana observability integrations.

Expertise You'll Bring:

- Big Data Processing Hands on experience with PySpark for ETL/ELT, data transformation, and performance optimization Solid understanding of distributed data processing concepts Data Cloud Architecture Strong experience designing data platforms on Google Cloud Platform (GCP) Experience with Data Lakes, data warehousing, and large scale migration programs Data Lake Architecture Storage Proven experience designing and implementing data lake architectures (e.g., Bronze/Silver/Gold or layered models).
- Strong knowledge of Cloud Storage (GCS) design, including bucket layout, naming conventions, lifecycle polic
- Educational background relevant to the role and its technical domain.

Benefits:

- Competitive salary and benefits package
- Culture focused on talent development with quarterly growth opportunities and company-sponsored higher education and certifications
- Opportunity to work with cutting-edge technologies
- Employee engagement initiatives such as project parties, flexible work hours, and Long Service awards
- Annual health check-ups
- Insurance coverage: group term life, personal accident, and Mediclaim hospitalization for self, spouse, two children, and parents

Values-Driven, People-Centric & Inclusive Work Environment:

Persistent is dedicated to fostering diversity and inclusion in the workplace. We invite applications from all qualified individuals, including those with disabilities, and regardless of gender or gender preference. We welcome diverse candidates from all backgrounds.

- We support hybrid work and flexible hours to fit diverse lifestyles.
- Our office is accessibility-friendly, with ergonomic setups and assistive technologies to support employees with physical disabilities.
- If you are a person with disabilities and have specific requirements, please inform us during the application process or at any time during your employment

Let’s unleash your full potential at Persistent - persistent.com/careers

“Persistent is an Equal Opportunity Employer and prohibits discrimination and harassment of any kind.”

1

Open Positions

Java,Security

Skills Required

BENGALURU

Location

Java,Security

Desirable Skills

184066

Job Code

📌 Data Architect (India)
🏢 Persistent
📍 India

Reply to this offer

Impress this employer describing Your skills and abilities, fill out the form below and leave Your personal touch in the presentation letter.

Subscribe to this job alert:

Get the latest job offers by email for: data architect (india) / india

Subscribe to this job alert:

Get the latest job offers by email for: data architect (india) / india