09 Oct
|
Quarks Technosoft
|
Bengaluru
09 Oct
Quarks Technosoft
Bengaluru
">Platform Data Devops & SRE
4-6 Years Bengaluru
- AWS
- Automation
- Python
- Golang
- Terraform
Mandatory Skills
AWS,Automation , Python,Golang , Terraform,Database Management
Skill to Evaluate
AWS,Cloud , Automation,Python , Golang,Terraform ,Database Management
Job Description
We are seeking a Platform Data DevOps & SRE Engineer with 4-6 years of experience to build, automate, and operate scalable data platforms in cloud-native environments.
The position combines software engineering, DevOps, database reliability engineering, and Site Reliability Engineering practices. The engineer will help manage NoSQL, streaming, caching, and relational data services while improving their availability, performance, observability, security, and operational efficiency.
The ideal candidate has an automation-first mindset and hands-on experience with AWS or GCP, Infrastructure as Code, configuration management, containers, Kubernetes, observability, and production support. The candidate should be comfortable writing automation in Go or Python and troubleshooting complex distributed systems.
Required Skills and Experience
- 4-6 years of experience in DevOps, SRE, platform engineering, database reliability engineering, or a related field.
- Hands-on cloud experience, preferably with AWS; exposure to GCP or Azure is beneficial.
- Strong experience with Terraform and Infrastructure as Code practices.
- Experience with Ansible or a comparable configuration-management tool.
- Proficiency in Go or Python for scripting, APIs, infrastructure automation, or operational tooling.
- Experience supporting at least one NoSQL, streaming, or caching technology such as Cassandra, DynamoDB, MongoDB, Kafka, AWS MSK, Redis, or Aerospike.
- Working knowledge of SQL and relational database fundamentals; experience with PostgreSQL or Oracle is beneficial.
- Experience with monitoring and observability platforms such as Grafana, Datadog, Prometheus, or equivalent tools.
- Positive understanding of Linux, networking, storage, REST APIs, and distributed-system fundamentals.
- Experience with Docker and Kubernetes, preferably Amazon EKS or Google Kubernetes Engine.
- Understanding of backup and recovery, disaster recovery, database migration, performance tuning, high availability, and capacity planning.
- Familiarity with CI/CD pipelines, Git-based development workflows, and automated deployment practices.
- Experience participating in incident response, production support, and root-cause analysis.
- Strong analytical, troubleshooting, communication, and collaboration skills.
- Bachelor s degree in Computer Science , Engineering, or a related field, or equivalent practical experience.
- Certifications on AWS, Terraform and Kuberenets will be good to have.
Education Qualification
Bachelor s degree in Computer Science , Engineering, or a related field, or equivalent practical experience.
Roles & Responsibilities
- Build, automate, and operate reliable platform data services across cloud and hybrid environments.
- Develop Infrastructure as Code using Terraform and configuration-management solutions using Ansible.
- Automate the provisioning, configuration, scaling, monitoring, upgrades, backup, recovery, and lifecycle management of data platforms.
- Support NoSQL, streaming, caching,
and relational database technologies such as Cassandra, Kafka/AWS MSK, Redis, DynamoDB, PostgreSQL, MongoDB, or similar platforms.
- Improve platform availability, scalability, performance, security, and disaster-recovery readiness.
- Develop automation and operational tooling using Go or Python.
- Implement monitoring, logging, dashboards, and alerting using tools such as Grafana, Datadog, Prometheus, or equivalent solutions.
- Define and monitor SLIs, SLOs, and error budgets for critical platform services.
- Build self-healing capabilities, automated remediation, failover, and scaling mechanisms.
- Troubleshoot production issues involving databases, Linux systems, networking, storage, Kubernetes, and cloud infrastructure.
- Participate in incident response, conduct root-cause analysis, and implement permanent preventive fixes.
- Perform platform upgrades, migrations, performance tuning, capacity planning, and resilience testing.
- Implement database security best practices, including authentication, authorization, encryption, secrets management, and access controls.
- Develop and maintain CI/CD pipelines, operational runbooks, support procedures, and technical documentation.
- Collaborate with application, product, security, and platform engineering teams to deliver reliable self-service data capabilities.
- Contribute to technical reviews, knowledge-sharing sessions, and the development of junior engineers.
- Explore AI-assisted approaches for anomaly detection, predictive scaling, operational automation, and incident remediation where applicable.
Disclaimer: This job posting has been aggregated from external source. Role details, content, and availability are subject to change. Applicants are advised to confirm the latest information directly on the company website before applying.
📌 Platform Data Devops & SRE Professional (Bengaluru)
🏢 Quarks Technosoft
📍 Bengaluru