Job Description:
Read this part first:
This is a build-and-automate role, not a babysit-the-servers one. Prodege is
moving fast and changing rapid, and we need an SRE who finds that energizing
rather than exhausting.
If you want a static environment where nothing changes and the runbook is
already written, this isn't your role, and that's okay. But if you're the kind
of engineer who sees manual toil and automates it away, who'd rather harden and
scale infrastructure than firefight it, and who's comfortable owning production
as the systems around you keep evolving, keep reading.
What you'll own:
You'll play a key role in our cloud and data infrastructure, making sure our
products run on a stable, scalable, secure foundation. You'll own AWS/Terraform
environments, CI/CD pipelines, and MySQL performance, work that directly affects
release velocity, site reliability, and customer experience.
Through thoughtful automation and scripting, you'll reduce operational toil,
increase consistency, and free engineering teams to focus on shipping features.
Paired with strong monitoring,
incident response, and documentation, you'll help
create predictable, repeatable operations as we grow, and by embedding security
best practices and continuously evaluating new tools, you'll help the org run
more efficiently while de-risking the infrastructure over time.
What you'll do:
* Infrastructure management: Use Terraform to define and provision AWS
infrastructure. Configure and maintain AWS services (EC2, S3, RDS, Lambda,
VPC).
* Automation and scripting: Build and manage automation scripts and tools in
Bash, Python, and PHP to streamline operations and improve efficiency.
* CI/CD integration: Implement and manage continuous integration and deployment
pipelines using Jenkins.
* Monitoring and optimization: Monitor system performance, availability, and
resource usage, and implement optimizations for efficiency and reliability.
* Incident management: Troubleshoot and resolve infras
📌 Site Reliability Engineer II (Pune)
🏢 Prodege
📍 Pune