Job DescriptionKey Responsibilities Nn Develop, maintain, and optimize SQL-based ETL processes, stored procedures, and data transformations across DB2 and SQL Server environments N Support and monitor ~225 scheduled jobs and ~40 job streams using IBM Workload Scheduler, ensuring timely execution and prompt failure resolution N Build and maintain datapipelines using IBM DataStage and UNIX scripting for enterprise data integration workflows N Support Oracle GoldenGate for real-time data replication and change data capture (CDC) across source and target systems N Develop and maintain data integration workflows from source systems to analytics platforms, including validation and reconciliation logic N Provide operational support for DB2 and SQL Server environments encompassing ~160 schemas, ~20TB active storage, ~4,000 tables, and ~3,300 views N Monitor pipeline healthproactively, detect anomalies, and resolve data quality and availability issues within defined SLAs N Support Dev/QA/Prod environment management including release coordination and production readiness validation N Assist with AWS stabilization activities for analytics data layers post migration from on-premises infrastructure N Track and manage all work through ServiceNow, ensuring accurate classification, status updates, and SLA compliance N Collaborate with Tableau and BusinessObjects developers to ensure data availability and pipeline reliability for reporting N Participate in L1/L2 triage for pipeline incidents, data quality failures,
and integration issues N Contribute to runbook documentation and standard operating procedures for supported pipelines and jobs NnRequired Qualifications Nn Minimum Degree Required: Bachelor’s Degree in Engineering, Statistics, Mathematics, Computer Science, Data Science, Economics, or a related quantitative field N 2-5 years of experiencein data engineering, ETL development, or data integration roles N Solid SQL proficiency — query optimization, stored procedures, and database objects in DB2 and/or SQL Server N Hands-on experience with at least one enterprise ETL or orchestration tool - IBM DataStage, Netezza N Experience with job scheduling tools (IBM Workload Scheduler, Control-M, or equivalent) N Familiarity with UNIX/Linux scripting for pipeline automation and file-based integrations N Understanding data warehouse concepts — schemas, star/snowflake models, dimensions, and facts N Ability to troubleshootdata pipeline failures end-to-end and communicate resolution steps clearly N Experience working in regulated, compliance-aware environments (healthcare preferred) NnPreferred Qualifications Nn Healthcare data experience — claims, clinical, EMR, pharmacy, or population health datasets N AWS experience — S3, Glue, RDS, Redshift, or Lambda in a data engineering context N Experience with Oracle GoldenGate or similar CDC/replication tools N Familiarity with Tableau or BusinessObjects as downstream consumers of engineered data N Knowledge of HIPAA datahandling requirements and PHI access controls N Exposure to ServiceNow for ITSM-based delivery tracking. N
📌 Data Engineer (Karnataka)
🏢 PwC
📍 Karnataka