Key Responsibilities
- Production Support Incident Management
- Provide L2/L3 support for Big Data and ETL applications
- Monitor batch jobs, workflows, and data pipelines (24/7 support model if required)
- Troubleshoot Informatica workflows, Hadoop/Hive jobs, and UNIX scripts
- Perform root cause analysis (RCA) and ensure timely resolution of incidents
- Track and resolve SLA breaches, P1/P2 incidents
- ETL Data Pipeline Support
- Support and maintain Informatica PowerCenter / IICS workflows
- Debug ETL failures, data mismatches, and performance issues
- Perform data validation and reconciliation across systems
- Handle data ingestion and transformation pipelines
- Big Data / Hadoop Ecosystem
- Troubleshoot issues in: Hadoop (HDFS, YARN)
- Hive queries and performance tuning
- Spark jobs (basic to intermediate exposure preferred)
- Analyze logs, optimize queries, and ensure smooth data processing
- Unix Scripting
- Develop and support Shell scripts for automation
- Analyze logs, manage file movements, job scheduling, and cron jobs
- Automate repetitive support activities
- Monitoring Tools
- Work with scheduling tools: Autosys / Control-M / Airflow
- Use monitoring and ticketing tools: ServiceNow / Jira
- Proactively monitor systems and prevent failures
- Performance Optimization
- Identify bottlenecks in ETL and Big Data workloads
- Optimize Hive queries, partitioning, and indexing strategies
- Improve job execution time and system stability
- Stakeholder Coordination
- Collaborate with: Development teams, Business stakeholders, Infrastructure teams
- Communicate status updates and incident reports clearly