We aren't automating scripts — we're deprecating the era of manual-heavy testing entirely. TestMu AI is building the world's first AI-native platform where Agentic Intelligence autonomously plans, authors, and self-heals the entire Quality Engineering lifecycle.
Drive reliability and scalability of TestMu AI's core infrastructure and service workflows.
THE PILLARS OF IMPACT
1. Platform Reliability (50%)
– Ensure robust cloud infrastructure and deployment pipelines.
– Improve observability tools and practices.
– Automate and streamline service workflows.
– Enhance platform systems for developer efficiency.
⚙️ 2. Infrastructure Engineering (30%)
– Build and maintain cloud-based systems.
– Optimize performance and scalability.
– Implement infrastructure improvements based on feedback.
3. Incident Management (20%)
– Lead incident response with strong SRE principles.
– Analyze and resolve production issues swiftly.
– Develop strategies to prevent future incidents.
MUST-HAVES — DO NOT APPLY UNLESS YOU HAVE THESE
– You must have strong cloud, scripting, and core DevOps fundamentals.
– You must have hands-on experience with observability tools like New Relic, Sumo Logic.
– You must demonstrate real coding ability.
– You must have a solid grasp of incident management and reliability/SRE thinking.
– You must explain service-layer architecture and reasoning effectively.
THE BAR — WHAT YOU MUST PROVE
Cloud Expertise: Designed and optimized cloud infrastructure for scalability and reliability.
Coding Proficiency: Developed scripts and tools to automate deployment processes.
Incident Management: Led a team to resolve critical production incidents effectively.
Architecture Understanding: Explained complex service architectures to non-technical stakeholders.
Skills:- DevOps
📌 Platform Engineer (Delhi)
🏢 TestMu AI
📍 Delhi
Reply to this offer
Impress this employer describing Your skills and abilities, fill out the form below and leave Your personal touch in the presentation letter.