Data Engineer Specialist
Position Summary
The Data Engineer is responsible for the development, maintenance, and operational support of enterprise data pipelines, ETL processes, and data platform components within Presbyterian Healthcare Services' Analytics Organization. Working across a complex environment of approximately 2,000 stored procedures, 225 scheduled jobs, 40 job streams, and a recently AWS-migrated data platform, this role ensures reliable, high-quality data flows that power enterprise reporting, analytics, and decision-making across clinical, operational, financial, and health plan domains.
Key Responsibilities
- Develop, maintain, and optimize SQL-based ETL processes, stored procedures, and data transformations across DB2 and SQL Server environments
- Support and monitor ~225 scheduled jobs and ~40 job streams using IBM Workload Scheduler, ensuring timely execution and prompt failure resolution
- Maintain data pipelines using IBM DataStage and UNIX scripting for enterprise data integration workflows
- Support Oracle GoldenGate for real-time data replication and change data capture (CDC) across source and target systems
- Provide operational support for DB2 and SQL Server environments encompassing ~160 schemas, ~20TB active storage, ~4,000 tables, and ~3,300 views
- Monitor pipeline health proactively, detect anomalies, and resolve data quality and availability issues within defined SLAs
- Support Dev/QA/Prod setting management including release coordination and production readiness validation
- Assist with AWS stabilization activities for analytics data layers post migration from on-premises infrastructure
- Track and manage all work through ServiceNow,
ensuring accurate classification, status updates, and SLA compliance
- Collaborate with Tableau and BusinessObjects developers to ensure data availability and pipeline reliability for reporting
- Participate in L1/L2 triage for pipeline incidents, data quality failures, and integration issues
- Contribute to runbook documentation and standard operating procedures for supported pipelines and jobs
Required Qualifications
- Minimum Degree Required: Bachelors Degree in Engineering, Statistics, Mathematics, Computer Science, Data Science, Economics, or a related quantitative field
- 1-2 years of experience in data engineering, ETL development, or data integration roles
- Good SQL proficiency query optimization, stored procedures, and database objects in DB2 and/or SQL Server
- Experience with job scheduling tools (IBM Workload Scheduler, Control-M, or equivalent)
- nderstanding data warehouse concepts schemas, star/snowflake models, dimensions, and facts
- Ability to troubleshoot data pipeline failures end-to-end and communicate resolution steps clearly
- Experience working in regulated, compliance-aware environments (healthcare preferred)
Preferred Qualifications
- Healthcare data experience claims, clinical, EMR, pharmacy, or population health datasets
- AWS experience — S3, Glue, RDS, Redshift, or Lambda in a data engineering context
- Experience with Oracle GoldenGate or similar CDC/replication tools
- Familiarity with Tableau or BusinessObjects as downstream consumers of engineered data
- Knowledge of HIPAA data handling requirements and PHI access controls
- Exposure to ServiceNow for ITSM-based delivery tracking.
📌 Data Engineer Specialist (Hyderabad)
🏢 PwC
📍 Hyderabad