06 Sep
|
TalentOla
|
Mumbai
We are looking for an experienced Infrastructure Observability Tools Solution Engineer with 6+ years of proven expertise in designing and implementing Observability solutions with a focus on Zabbix, Nagios, OpsRamp, SolarWinds, Prometheus, and similar tools. As a Solution Architect, you will play a key role in architecting end-to-end Observability solutions that meet the complex requirements of our organization and ensure the stability, performance, and availability of our infrastructure and applications.
Responsibilities:
- Lead the end-to-end design and architecture of infrastructure Observability solutions, leveraging Zabbix, Nagios, OpsRamp, SolarWinds, Prometheus, and other relevant tools.
- Collaborate with stakeholders to understand business objectives, technical requirements, and existing infrastructure environments.
- Assess current Observability capabilities and infrastructure landscape to identify gaps, inefficiencies, and areas for improvement.
- Develop comprehensive Observability strategies and roadmaps aligned with business goals, industry best practices, and emerging trends.
- Design and implement scalable, reliable, and cost-effective Observability architectures that address the dynamic needs of our organization.
6.
Define
Observability standards, policies, and procedures to ensure consistency, compliance, and effectiveness across the organization.
- Evaluate and recommend appropriate Observability tools, technologies, and integrations to meet specific use cases and requirements.
- Work closely with engineering and operations teams to implement Observability solutions, configure Observability tools, and integrate with existing systems.
- Provide technical leadership and guidance to cross-functional teams on Observability best practices, tool selection, implementation methodologies, and troubleshooting techniques.
- Conduct performance analysis, capacity planning, and trend analysis using Observability data to proactively identify and address potential issues.
- Stay updated on the latest advancements, features, and best practices in infrastructure Observability tools and technologies and evaluate their applicability to our environment.
- Collaborate with vendors, partners, and industry experts to assess new technologies, products, and services in the Observability space.
Requirements:
- 6+ years of hands-on experience in designing, architecting, and implementing infrastructure Observability solutions.
- Expertise in Zabbix, Nagios, OpsRamp, SolarWinds, Prometheus, or similar Observability tools, with a deep understanding of their capabilities, limitations, and configurations.
- Strong knowledge of IT infrastructure components and technologies, including servers, networks, storage, virtualization, cloud services, and microservices architectures.
- Proficiency in scripting and automation using languages such as Python, Bash, PowerShell, or similar.
- Excellent analytical, problem-solving, and decision-making skills with the ability to translate complex requirements into scalable and effective Observability solutions.
- Effective communication, presentation, and interpersonal skills with the ability to interact with stakeholders at all levels of the organization.
- Proven experience leading cross-functional teams and driving projects to successful completion.
- Relevant certifications such as Zabbix Certified Professional, Nagios Certified Administrator, SolarWinds Certified Professional, or Prometheus Certified Administrator are desirable
📌 Infrastructure Observability Engineer (Mumbai)
🏢 TalentOla
📍 Mumbai