Wipfli India
This role is sourced from an external job board and is not a ProSculpt-verified employer. Review the listing carefully, never pay any fee to apply, and verify the company before sharing personal details.
Technology skills, Competencies and Experience:Incident Monitoring & ResponseContinuously monitor the Pantomath Operation Center dashboard for pipeline failures, stale or missing data, and anomalies across the enterprise data stack Perform first-line triage using Pantomaths automated root-cause analysis and cross-platform lineage tracing to identify affected pipelines and downstream impact Classify and prioritize incidents by business severity and follow defined escalation playbooks Drive incidents from detection through resolution, escalating to engineering or platform teams when root cause requires deeper remediation Role and Responsibilities:Pantomath Platform AdministrationAdminister and maintain the Pantomath platform, including monitor configuration, alert thresholds, connectors, and lineage coverageOnboard new pipelines and data sources into observability coverage as the data estate growsTune monitors and reduce alert noise by refining thresholds based on historical incident patternsMaintain platform health, user access, and integration with upstream and downstream enterprise systems Technical SkillsHands-on experience with Pantomath or a comparable data observability platform (Monte Carlo, Bigeye, or similar) strongly preferredFamiliarity with cloud and lakehouse data platforms such as Databricks, Snowflake, Azure, or AWSProficiency in SQL for incident investigation and root cause analysis Working knowledge of data pipeline orchestration tools (Airflow, dbt, Delta Live Tables, or similar)Experience with incident management and ticketing tools (ServiceNow, Jira, or similar)Familiarity with ITIL incident management principles and SLA-driven escalation models Communication and Stakeholder Management Communicate outage status, root cause, and resolution timelines clearly to business stakeholders and data consumers across the enterpriseDraft and distribute incident notifications and status updates through approved communication channelsMaintain incident logs and root cause analysis (RCA) documentation, and contribute to post-incident reviewsHand off open incidents clearly across shift boundaries to ensure continuity of coverage .
Want a personalised feed?
Sign up to save jobs, track applications, and get AI-matched recommendations.
Create a free account