Description At Ameriprise we are passionate about building solutions that solves problem s We count on our site reliability engineers SREs to empower users with a rich feature set high availability and stellar performance level to pursue their missions As we expand customer deployments we re seeing an experienced SRE to deliver insights from massive-scale data in real time Specifically we re searching for someone who has fresh ideas and a unique viewpoint and who enjoys collaborating with a cross-functional team to develop real-world solutions and positive user experiences for every interaction Roles Responsibilities Ensure production system health availability and performance monitoring Build and maintain infrastructure and application platforms with automation for scalability Improve reliability quality and time-to-market through performance optimization and innovation Provide operational support for large-scale distributed systems and collaborate with development teams for service improvements Lead ITSM processes Incident Management with awareness of Change and Problem Management concepts Drive problem resolution with short-term fixes and long-term solutions manage Disaster Recovery with automation contributions Skillset and Qualifications Bachelor s degree in Computer Science or related field Solid programming scripting skills in languages such as Python Java C C PowerShell Ansible or JavaScript Experience with distributed systems and cloud platforms preferably AWS Knowledge of ITSM processes including Incident Management and awareness of Change and Problem Management Familiarity with monitoring scheduling and event management tools e g Dynatrace Tidal and dashboard creation using tools like Grafana and ServiceNow Understanding of Site Reliability Engineering SRE concepts and practices Proactive in identifying issues optimizing performance and driving improvements