20 Sep
|
TeamPlus Staffing Solution
|
Bengaluru
20 Sep
TeamPlus Staffing Solution
Bengaluru
Senior AWS Cloud Support & Troubleshooting Engineer
- Experience: 10+ years overall; 5+ years hands-on AWS
- Strong expertise in AWS S3, IAM, KMS, CloudTrail & CloudWatch
Job time : 3 pm to 12 Night IST Role OverviewWe are seeking a highly experienced, hands-on AWS Cloud Support & Troubleshooting Engineer to support a large-scale cloud workplace used for media content ingestion, storage, processing and distribution workflows.This is not an application development role. The successful candidate will work alongside an existing engineering team responsible for developing the applications and ingestion pipelines and will serve as the deep technical support and troubleshooting resource for those solutions.The defining competency for this position is exceptional debugging and root-cause analysis.The engineer must be capable of taking complex or poorly defined production issues, working directly with affected users and technical partners, reproducing the problem within the environment, systematically isolating the failure point, identifying the underlying cause, and implementing or coordinating an appropriate resolution.The individual will directly represent the Cloud Operations team when working with Client Tech Ops, external labs and finishing partners, and internal users. This requires an unusual combination of deep technical expertise, disciplined problem solving and strong user-facing communication skills.Key Responsibilities
- Own the technical investigation and resolution of complex AWS production-support issues involving S3, security, networking, APIs, applications, ingestion pipelines and file movement.
- Reproduce reported bugs and failures within the environment to understand the precise conditions under which problems occur.
- Perform methodical root-cause analysis rather than relying on trial-and-error configuration changes or temporary fixes.
- Troubleshoot complex AWS S3 behavior, including bucket policies, IAM permissions, roles, object ownership, authentication, access controls, APIs, service limits and other S3-specific behaviors and edge cases.
- Diagnose cloud networking issues involving routing, connectivity, load balancing and communication between systems and services.
- Investigate problems involving application-to-S3 and system-to-system API interactions.
- Troubleshoot large-scale file and data movement, including ingestion workflows and potentially data-normalization issues.
- Work directly with Tech Ops, external labs and finishing partners such as PhotoKem, BitPress and Deluxe, and internal Lionsgate users to understand reported problems, gather evidence and drive issues through resolution.
- Determine whether failures originate within AWS infrastructure, security, networking, APIs, applications, data, external vendor environments, or user configuration.
- Establish and troubleshoot appropriate user and system access to cloud resources.
- Use scripting where appropriate to investigate, reproduce,
diagnose and validate technical issues.
- Document, track and manage technical issues and resolutions through Jira and established support channels.
- Collaborate closely with the application-development engineers when investigation identifies an issue requiring an application or engineering change.
- Validate fixes and ensure that resolutions address the underlying root cause without introducing unintended downstream impacts.
Required Technical ExperienceThe successful candidate should have extensive hands-on experience operating and troubleshooting large-scale AWS production environments, with particularly deep expertise in:
- AWS S3, including its architecture, APIs, permissions, object behavior, limitations, defaults, quirks and edge cases
- AWS IAM, security policies, roles and access management
- Cloud networking, including routing, connectivity and load balancing
- REST/API-based integrations and application-to-cloud interactions
- Scripting for investigation, automation and troubleshooting
- Large-scale file/object movement and ingestion
- Data processing and normalization concepts
- Troubleshooting distributed and heterogeneous technology environments
- Production incident investigation and root-cause analysis
- Jira or comparable incident/ticket-management platforms
Experience supporting environments with very large object counts and petabyte-scale storage is highly desirable.Experience within media and entertainment, digital asset management, content supply chains, media ingestion, post-production, or large-file workflows would be particularly valuable.
📌 Sr AWS Cloud Support Engineer -Media & Entertainment Remote (Bengaluru)
🏢 TeamPlus Staffing Solution
📍 Bengaluru