DevOps & Cloud Infrastructure Engineer (New Delhi)

DevOps & Cloud Infrastructure Engineer (New Delhi)

16 Aug
|
Menstar Services
|
New Delhi

16 Aug

Menstar Services

New Delhi

DevOps & Cloud Infrastructure Engineer

Company: Digitalis Technologies Pvt. Ltd.

Role: DevOps & Cloud Infrastructure Engineer

Experience: 36 Years

Employment Type: Full time

Location: Delhi

About Digitalis Technologies

Digitalis Technologies (DT) is an enterprise technology and consulting company focused on ERP implementation, custom business applications, integrations, automation, and technology support.

As our customer deployments and infrastructure footprint grow, we are looking for a hands-on DevOps & Cloud Infrastructure Engineer who can take ownership of production infrastructure, deployments, databases, backups, migrations, monitoring, security, and system reliability.

Role Overview The DevOps & Cloud Infrastructure Engineer will be responsible for managing Linux-based production environments and supporting development teams in deploying and maintaining business-critical applications.

The role requires robust practical knowledge across infrastructure and application operations, including Linux, cloud platforms, databases, web servers, APIs, monitoring, automation, deployments, and troubleshooting.

The candidate should be comfortable working directly on production systems, diagnosing complex issues, automating repetitive activities, and coordinating with developers when problems span application and infrastructure layers.

Experience with Python-based applications, Frappe Framework, or ERPNext is an advantage but is not mandatory.

Key Responsibilities

Cloud Infrastructure & Server Management

- Provision, configure, secure, and maintain Linux-based cloud servers and virtual machines.
- Manage infrastructure on AWS, DigitalOcean, Azure, Hetzner, or similar cloud platforms.
- Maintain Ubuntu/Linux environments for development, staging, and production.
- Configure DNS, domains, SSL certificates, firewalls, networking, Nginx, and reverse proxies.
- Monitor CPU, memory, disk, network usage, uptime, and overall infrastructure health.
- Perform capacity planning and recommend infrastructure upgrades where required.
- Optimize hosting and cloud infrastructure costs without compromising reliability.

Application Deployment & Production Operations

- Deploy applications and services across development, staging, and production environments.
- Establish standardized and repeatable deployment processes.
- Manage application dependencies, services, environment variables, and configuration.
- Investigate deployment failures and production incidents.
- Support version upgrades, workplace migrations, and rollback procedures.
- Coordinate with developers to diagnose issues across code, infrastructure, database, and third-party services.
- Maintain proper isolation and controls across multiple customer environments.

Backup, Recovery & Business Continuity

- Implement automated backups for databases and application files.
- Maintain off-site or cloud-based backup storage and retention policies.
- Monitor backup jobs and immediately address failures.
- Perform periodic restoration tests to ensure backups are usable.
- Maintain documented disaster recovery procedures.
- Restore applications and databases following failures or migration requirements.
- Support business continuity planning for critical systems.

A backup process is considered effective only when recovery has been successfully tested. Database Administration & Migration

- Manage relational databases including MariaDB, MySQL, PostgreSQL, and similar platforms.
- Perform database backups, restores, imports, exports, and migrations.
- Handle database transfers between servers and environments.
- Configure database access, permissions, and security.
- Monitor database size, resource utilization, and basic performance indicators.
- Troubleshoot connectivity, configuration, and performance issues.
- Understand indexing, query performance,



and database optimization fundamentals.
- Support application teams during schema changes, upgrades, and large data migrations.
- Experience with Oracle databases is an added advantage.

CI/CD & Source Code Management

- Work with Git-based development and release workflows.
- Manage or support GitHub, GitLab, or similar repositories.
- Build and maintain CI/CD pipelines where appropriate.
- Automate deployment and release activities.
- Maintain clear separation between development, staging, and production.
- Implement controlled release, rollback, and deployment procedures.
- Maintain deployment history, version visibility, and release documentation.

Monitoring, Logging & Incident Management

- Implement server, application, and service monitoring.
- Configure alerts for availability, CPU, memory, disk usage, and critical services.
- Monitor application processes, scheduled jobs, background workers, and integrations.
- Analyze operating system, application, and database logs.
- Diagnose production outages and service failures.
- Perform root-cause analysis for recurring incidents.
- Implement preventive monitoring and corrective actions.
- Maintain incident and troubleshooting documentation.

APIs & Integrations

- Troubleshoot application integrations with external systems.
- Work with REST APIs, HTTP/HTTPS, JSON, webhooks, and authentication mechanisms.
- Understand API keys, token-based authentication, and OAuth fundamentals.
- Use tools such as Postman and cURL for testing and diagnosis.
- Support SFTP/FTP-based integrations and scheduled data exchanges.
- Investigate API errors, connectivity issues, timeouts, authentication failures, and unexpected responses.
- Coordinate with application developers and external system teams to resolve integration issues.

Automation & Scripting

- Automate repetitive infrastructure and operational activities.
- Write and maintain Bash/Shell scripts.
- Use Python for infrastructure automation and operational utilities where appropriate.
- Automate backup, deployment, migration, monitoring, and housekeeping activities.
- Manage cron jobs and scheduled processes.
- Experience with Ansible, Terraform, or similar Infrastructure-as-Code tools is an advantage.

Security & Access Management

- Implement infrastructure security best practices.
- Configure SSH and key-based authentication.
- Manage server users, roles, and permissions.
- Maintain firewall rules and restrict unnecessary ports and services.
- Manage SSL/TLS certificates and secure database access.
- Maintain application secrets and environment variables securely.
- Perform routine operating system patching and security updates.
- Control production credentials, access keys, and privileged access.
- Maintain access records and remove obsolete accounts or credentials.

Required Technical Skills Candidates should have strong practical experience with most of the following:

- Linux / Ubuntu Administration
- Cloud Infrastructure
- AWS / DigitalOcean / Azure / Hetzner or similar
- Nginx and reverse proxy configuration
- MariaDB / MySQL / PostgreSQL
- Git and GitHub / GitLab
- Bash / Shell scripting
- Python scripting
- REST APIs
- DNS
- SSL/TLS
- SSH
- Cron and scheduled jobs
- Database backup, restoration, and migration
- Application deployment
- Monitoring and logging
- Networking fundamentals
- Firewall and security configuration





Preferred / Good-to-Have Skills Experience with the following will be considered an advantage:

- Docker and Docker Compose
- Kubernetes
- Redis
- Supervisor / systemd
- GitHub Actions or GitLab CI
- Terraform
- Ansible
- AWS services such as EC2, RDS, S3, Route 53, and CloudWatch
- Cloudflare
- Oracle databases
- Python web frameworks
- Node.js ecosystem
- Sentry or similar monitoring platforms
- Frappe Framework
- ERPNext
- Bench CLI

Frappe / ERPNext Added Advantage Candidates with prior Frappe or ERPNext exposure may also work on:

- Frappe/ERPNext installation and configuration.
- Bench and multi-site management.
- Version upgrades and migrations.
- Background workers, scheduler, Redis, and queue troubleshooting.
- Site backup, restoration, and environment cloning.
- Custom application deployment.
- Python-based Frappe development.
- Basic JavaScript/client-side development.
- REST API development and integrations.

Prior Frappe or ERPNext experience is not mandatory if the candidate has strong fundamentals in Linux, cloud infrastructure, databases, Python, deployment, and production operations.

Candidate Profile

We are looking for someone who:

- Has 3-6 years of relevant hands-on experience.
- Can independently manage Linux production environments.
- Has supported real production applications and not just development systems.
- Has strong troubleshooting and root-cause analysis skills.
- Is comfortable working from the command line.
- Can interpret system, application, and database logs.
- Understands the relationship between application architecture and infrastructure.
- Takes ownership of infrastructure reliability, backups, security, and production issues.
- Automates repetitive work wherever practical.
- Maintains proper documentation for environments, deployments, credentials, and procedures.
- Can coordinate effectively with developers, customers, and third-party service providers.

Typical Situations You May Handle

Examples include

- A production application becomes slow or unavailable.
- Disk utilization reaches critical levels.
- A database or complete application environment needs to be migrated to a new server.
- A large production workload needs to move between cloud providers with minimal downtime.
- An SSL certificate fails to renew.
- Scheduled jobs or background workers stop running.
- A backup or restore operation fails.
- A deployment or application upgrade does not complete successfully.
- A production environment needs to be recovered from backup.
- An API integration starts failing unexpectedly.
- An application works in development but fails in production.
- A server becomes inaccessible.
- Production needs to be cloned to staging safely.
- A new customer environment needs to be provisioned, secured, and monitored.

Key Performance Expectations Performance in this role will be evaluated based on:

- Infrastructure availability and stability.
- Reliability of backups and recovery procedures.
- Deployment quality and consistency.
- Resolution time for production issues.
- Reduction in recurring incidents.
- Security and access discipline.
- Successful server and database migrations.
- Quality of monitoring and alerting.
- Level of infrastructure automation.
- Accuracy and completeness of technical documentation.
- Responsiveness during critical incidents.

Career Growth The role offers exposure to enterprise applications, ERP systems, cloud infrastructure, databases, APIs, integrations, automation, and production engineering. For the right candidate, the position can evolve into a broader senior DevOps Engineer, Platform Engineer, or cloud & infrastructure lead role as Digitalis Technologies expands its managed infrastructure and enterprise application portfolio.

Interested candidate may share their CV at [email protected] or can connect me at (phone hidden)

📌 DevOps & Cloud Infrastructure Engineer (New Delhi)
🏢 Menstar Services
📍 New Delhi

Reply to this offer

Impress this employer describing Your skills and abilities, fill out the form below and leave Your personal touch in the presentation letter.

Subscribe to this job alert:

Get the latest job offers by email for: devops & cloud infrastructure engineer (new delhi) / new delhi

Subscribe to this job alert:

Get the latest job offers by email for: devops & cloud infrastructure engineer (new delhi) / new delhi