21 Aug
|
Talent500
|
Hyderabad
21 Aug
Talent500
Hyderabad
Talent500 is hiring for one of its clients.
Who are we:
Core Insurance Platforms (CIP) is Zurich’s global capability responsible for building, running, and evolving core insurance technology. We set a unified, scalable operating model—covering governance, standards, architecture, service delivery, and reuse—so our business units can deliver at speed and scale.
CIP is the strategic steward of Zurich’s Guidewire ecosystem, aligning platform roadmaps to business strategy while driving stability, modernization, reduced supplier dependency, and long term cost efficiency.
India delivery center is one of our global delivery and capability hub. We bring together experts in AI, engineering, analysis, quality, and architecture to deliver product & process solutions, application run services, change and transformation initiatives, and centralized platform services across both on prem and Guidewire Cloud environments. Our teams operate from multiple global delivery centers, supporting Zurich’s business units worldwide.
Key Responsibilities:
Automation & Tooling:
- Develop and maintain automation scripts using Python for:
- Monitoring onboarding
- Validation checks
- Service discovery
- API integrations
- Data processing
- Automate Infrastructure-as-Code using Ansible, including modular deployments and multi cloud provisioning
- Implement configuration automation and server orchestration using Ansible
- Automate Dynatrace OneAgent installation, tagging rules, dashboards, alerting profiles, and management zone creation
Cloud Engineering (Azure & AWS):
- Deploy and manage cloud resources across Azure and AWS using Ansible
- Enable cloud observability for compute, networking, serverless, container workloads, logs, and metrics
- Work with cloud teams to ensure visibility into critical services and hybrid environments
Observability & Monitoring:
- Configure Dynatrace capabilities, including:
- Automated service onboarding
- Tagging strategies and naming conventions
- Management zones and dashboards
- Custom Metrics API, log ingestion, and alerting pipelines
- Distributed tracing analysis
- Perform root cause analysis and performance troubleshooting using Dynatrace dashboards, extensions, service flows, and logs
- Collaborate with application owners to improve observability practices and drive adoption
- Contribute to observability standards, templates, and best practices
- Ensure monitoring consistency across regions, business units, and environments
- Evaluate observability gaps and propose automation based solutions
- Leverage AI capabilities (Dynatrace Davis AI, anomaly detection, log analytics, AI tools where approved) to improve monitoring, root cause analysis, troubleshooting efficiency, and decision-making, while adhering to security, compliance, and governance standards.
Technical Skills:
Automation & Programming:
- Solid expertise in Python for automation, API integration, file handling, retries, exception handling, and logging
- Hands on experience building Infrastructure as Code using Ansible in enterprise environments
- Proficiency with Ansible for configuration management, provisioning, and orchestration
Cloud:
- Understanding of Azure and AWS, including:
- Virtual machines, networking, and load balancers
- IAM concepts
- Monitoring tools such as CloudWatch and Azure Monitor
- Serverless and container workloads
Observability / APM:
- Deep experience with Dynatrace, including:
- OneAgent deployment on Windows and Linux
- Management Zones and tagging rules
- Problem notifications and integrations
- Custom dashboards and metrics
- Distributed tracing and log ingestion
- Robust understanding of observability pillars: metrics, logs, traces, and SLIs/SLOs
Nice to Have:
- Experience with Kubernetes, Docker, AKS, and EKS, ThousandEyes
- Experience with CI/CD tools such as Azure DevOps, GitHub Actions, and Jenkins
- Experience with ITSM tools including ServiceNow (Change, Incident, CMDB,)
- Scripting knowledge in Bash or PowerShell
Experience Required:
- Minimum 3–5 years of relevant experience
- Working knowledge of AI/AIOps concepts (e.g., anomaly detection, event correlation, AI-driven RCA) with experience using tools like Dynatrace Davis AI and log analytics to support incident analysis, decision-making, and continuous improvement, while following security and governance standards
📌 Senior Dynatrace Engineer-28794] (Hyderabad)
🏢 Talent500
📍 Hyderabad