01 Sep
|
Novo AI
|
Hyderabad
Role Overview
We're looking for a DevOps Engineer to own the infrastructure that keeps WatchMen running reliably — from AWS cloud services down to a distributed fleet of edge devices sitting on factory floors. This is a hands-on role across cloud infra, CI/CD, monitoring, and self-hosted services, not a ticket-queue ops job.
What You'll Work On
- Managing AWS infrastructure via Terraform, ECS on EC2, ALB, networking, and dedicated-tenant VPC architecture
- Operating our Docker Swarm–based backend deployment and improving its reliability and scaling
- Maintaining our data infrastructure: PostgreSQL, TimescaleDB, PgBouncer, and Redpanda/Kafka event pipelines
- Building and maintaining CI/CD pipelines on self-hosted GitLab CE
- Owning observability: Grafana Cloud dashboards, OpenTelemetry tracing, and alerting for both cloud services and production incidents
- Managing Grafana Alloy fleet monitoring across our Raspberry Pi/edge device fleet deployed at customer sites (remote config, log pipelines, health metrics)
- Maintaining self-hosted internal services (MQTT broker/EMQX, Vaultwarden, n8n) behind Caddy reverse proxies
- Diagnosing infrastructure-level production issues, network, database, message broker, and container orchestration problems
- Hardening system security, including tools like Wazuh
Requirements
- 3–6 years of professional DevOps/infrastructure experience
- Strong hands-on experience with Docker and container orchestration (Docker Swarm, ECS, or Kubernetes)
- Experience with Terraform or another infrastructure-as-code tool
- Solid Linux systems administration and networking fundamentals
- Experience with CI/CD pipelines (GitLab CI, GitHub Actions, or similar)
- Comfortable debugging production incidents independently
Nice to Have
- Experience with AWS specifically (ECS, ALB, VPC design)
- Experience with time-series databases (TimescaleDB) or event streaming (Kafka/Redpanda)
- Experience with MQTT or managing distributed IoT/edge device fleets
- Experience with Grafana, Prometheus, or OpenTelemetry-based observability stacks
- Exposure to self-hosted service management (GitLab, reverse proxies like Caddy/nginx)
Location: Hyderabad (in person) | Type: Full time
Why Join Novo AI
- Your code helps prevent costly machine failures on actual factory floors — not just another dashboard.
- Small team, real autonomy and your decisions shape the product, not a committee's.
- Work on live industrial data, AI-driven analytics, and systems challenges most companies only see at much larger scale.
- Early-stage teams mean broader scope and faster learning than a typical corporate role.
- Work closely with founders and leadership, with visibility into a global tech company.
Pay: ₹800,000.00 - ₹1,800,000.00 per year
Benefits
- Flexible schedule
- Work from home
Experience
- Flask: 2 years (Preferred)
- Back-end development: 3 years (Preferred)
Language
- English (Preferred)
Location
- Hyderabad, Telangana (Preferred)
Work Location: Hybrid remote in Hyderabad, Telangana
📌 DevOps / Infrastructure Engineer (Hyderabad)
🏢 Novo AI
📍 Hyderabad