13 Sep
|
Wipfli
|
Bengaluru
Position: Analyst - Data Center Operations
(2 to 5 years of experience in Data Analysis)
Type: Full Time Employee (FTE)
Wipfli is building a Data Operations Center (DOC) within its Enterprise Data & Analytics organization to provide continuous, enterprise-wide monitoring and incident resolution for our data pipeline ecosystem. This critical function ensures the integrity, availability, and reliability of data that thousands of professionals across the firm depend on every day.
Wipfli seeks a skilled offshore Data Operations Center Analyst to monitor, triage, and communicate the resolution of data pipeline incidents across the enterprise. This role is the data equivalent of a Security Operations Center (SOC) or Network Operations Center (NOC) analyst, owning day-to-day operation of our Pantomath observability platform to detect issues early, drive them to root cause, and keep the business informed throughout resolution.
Technology skills, Competencies and Experience:
Incident Monitoring & Response
- Continuously monitor the Pantomath Operation Center dashboard for pipeline failures, stale or missing data, and anomalies across the enterprise data stack
- Perform first-line triage using Pantomath’s automated root-cause analysis and cross-platform lineage tracing to identify affected pipelines and downstream impact
- Classify and prioritize incidents by business severity and follow defined escalation playbooks
- Drive incidents from detection through resolution, escalating to engineering or platform teams when root cause requires deeper remediation
Role and Responsibilities:
Pantomath Platform Administration
- Administer and maintain the Pantomath platform, including monitor configuration, alert thresholds, connectors, and lineage coverage
- Onboard recent pipelines and data sources into observability coverage as the data estate grows
- Tune monitors and reduce alert noise by refining thresholds based on historical incident patterns
- Maintain platform health, user access, and integration with upstream and downstream enterprise systems
Technical Skills
- Hands-on experience with Pantomath or a comparable data observability platform (Monte Carlo, Bigeye, or similar) strongly preferred
- Familiarity with cloud and lakehouse data platforms such as Databricks, Snowflake, Azure, or AWS
- Proficiency in SQL for incident investigation and root cause analysis
- Working knowledge of data pipeline orchestration tools (Airflow, dbt, Delta Live Tables, or similar)
- Experience with incident management and ticketing tools (ServiceNow, Jira, or similar)
- Familiarity with ITIL incident management principles and SLA-driven escalation models
Communication & Stakeholder Management
- Communicate outage status, root cause, and resolution timelines clearly to business stakeholders and data consumers across the enterprise
- Draft and distribute incident notifications and status updates through approved communication channels
- Maintain incident logs and root cause analysis (RCA) documentation, and contribute to post-incident reviews
- Hand off open incidents clearly across shift boundaries to ensure continuity of coverage
Qualifications:
- Bachelor’s degree in Computer Science, Information Systems, Data Science, or related field, or equivalent practical experience
- 2–5 years of experience in data engineering, data operations, IT operations, or a related technical support role
- Demonstrated experience monitoring and triaging incidents in a production data or IT environment
- Understanding of data pipeline concepts, including ETL/ELT processes, data lineage, and data quality
- Strong written and verbal communication skills, with the ability to translate technical incidents into clear business-facing updates
- Comfort working in a shift-based or extended-hours coverage model aligned to U.S. business hours
Nice to have:
- Experience in professional services, accounting, audit, tax, or advisory environments
- Prior experience in a Network Operations Center (NOC), Security Operations Center (SOC), or data operations center
- Familiarity with data governance and data quality frameworks
- Experience supporting offshore or globally distributed delivery teams
- Exposure to AI/ML pipeline monitoring or RAG system observability
Working Model and Expectation:
- Shift coverage aligned to U.S. business hours, with rotation to extend the monitoring window as the DOC matures
- Adherence to defined incident severity matrices and communication service-level agreements
- Active participation in shift handoff, daily DOC stand-up, and weekly RCA review
- Use of approved monitoring and ticketing tools only; no unauthorized integrations
- Participation in on-call rotation as DOC coverage expands
No. of positions: 02
Work location: Wipfli India, Bengaluru or Pune or Hyderabad.
📌 Analyst - Data Center Operations (Bengaluru)
🏢 Wipfli
📍 Bengaluru