02 Sep
|
Synthires
|
India
Incident Management, Reliability & SRE Evaluator
Position: Incident Management, Reliability & SRE Evaluator
Type: Freelance Hourly Contract
Compensation: $80–$120/hour
Location: Remote
About the Opportunity
This opportunity is for experienced Site Reliability Engineering (SRE), Incident Management, and Reliability professionals interested in contributing to advanced AI research and evaluation projects involving operational excellence, system reliability, infrastructure management, and incident response.
The role focuses on evaluating AI-generated work products, including incident reports, postmortems, operational documentation, reliability analyses, dashboards, presentations, and technical recommendations while applying professional judgment to assess accuracy, quality, and technical rigor.
You'll leverage your expertise in incident response, production operations, reliability engineering, service management, and system performance to help improve the capabilities of next-generation AI systems.
Key Responsibilities
- Evaluate AI-generated incident reports, postmortems, operational analyses, presentations, and reliability documentation for accuracy and quality.
- Assess outputs using domain-specific evaluation criteria and industry best practices.
- Identify technical inaccuracies, operational gaps, root-cause analysis weaknesses, and presentation issues.
- Review reliability strategies, incident response plans, monitoring frameworks, and service improvement recommendations.
- Provide detailed written feedback and scoring based on established evaluation guidelines.
- Assess the quality of technical communication, operational reasoning, and business impact analysis.
- Contribute to the development of high-quality reliability and infrastructure evaluation benchmarks and training datasets.
Required Qualifications
- 5+ years of experience in Site Reliability Engineering (SRE), Incident Management, Production Operations, Infrastructure Engineering, Platform Engineering, DevOps, or a related reliability-focused role.
- Skilled fluency in English with strong written communication skills.
- Advanced proficiency in Microsoft Office and Google Workspace , particularly PowerPoint/Google Slides , documentation tools, and operational reporting platforms.
- Strong understanding of incident response processes, reliability engineering principles, service monitoring, and operational excellence frameworks.
- Ability to identify technical, analytical, and presentation-related issues in complex work products.
Preferred Qualifications
- Master's degree or higher from a reputable institution.
- Experience managing large-scale production systems, critical incidents, or enterprise reliability programs.
- Familiarity with observability platforms, cloud infrastructure, distributed systems, and operational performance metrics.
- Experience creating executive incident reviews, postmortems, operational dashboards, or reliability reports.
- Familiarity with AI-generated technical content and evaluation methodologies.
Project Details
- Start Date: Immediate
- Fully remote and flexible schedule
- Hourly contract engagement
- Opportunity to contribute to frontier AI evaluation projects
- Additional project opportunities may be available based on performance
Compensation
- Competitive compensation of $80–$120/hour
- Weekly payments
- Independent contractor engagement
Application Process
1. Easy Apply on LinkedIn
2. Check Email for Next Steps
3. Participate in Resume Evaluation & Interview Stage
📌 SRE & Incident Management Specialist (Remote | ) (India)
🏢 Synthires
📍 India