Job Description:
At Bank of America, we are guided by a common purpose to help make financial lives better through the power of every connection. We do this by driving Responsible Growth and delivering for our clients, teammates, communities and shareholders every day.
Being a Great Place to Work and providing a culture of caring is core to how we drive Responsible Growth. We are intentional about fostering an inclusive workplace where every teammate has the opportunity to succeed, build a career and contribute to our shared success. This includes attracting and developing exceptional talent, recognizing and rewarding performance, and supporting our teammates’ physical, emotional, and financial wellness through affordable, competitive and flexible benefits.
We value the unique perspectives individuals bring from all backgrounds and career paths - whether shaped by military service, community college education, or a wide range of work and life experiences. These journeys foster resilience, leadership and innovation, strengthening our workforce and positively impact the communities we serve.
Bank of America is committed to an in-office culture that supports collaboration, engagement, and career development. Our approach includes clear in-office expectations, while providing an appropriate level of flexibility based on role-specific responsibilities and business needs.
At Bank of America, you can build a successful career with opportunities to learn, grow, and make an impact. Join us!
Job Description:
This job is responsible for providing front-line support to end users, responding to issues related to incidents and problem management governance for multiple applications, and leading triage activities on all business impacting incidents. Key responsibilities include ensuring compliance with incident management and problem management policies and procedures, serving as a focal point for the customer, client, and associate experience, restoring complex production incidents under tight Service Level Agreements, and pursuing root cause and problem resolution follow ups.
This role is critical to sustain Horizon CI/CD production stability, Ansible platform support, deployment reliability, and OpenShift/cloud transition readiness while reducing incident, SLA, vulnerability, DR, and support-coverage risk across enterprise application teams.
Responsibilities:
Leads production support triage efforts, manages bridge line troubleshooting, engages in technical research, and escalates issues to leadership as needed
Ensures all impacts are accurately recorded and documented in the system of record, oversees that documents and wikis are updated and available for use during triage, and supports the documentation of application flows, upstream/downstream impacts during outages, the customer experience, and contacts for support needs
Identifies and/or validates business impacts through interpretation of monitors, dashboards, and logs to communicate with leadership and vendors
Manages activities to identify incident root cause, resolution, preventative actions, and change requests, and reports on incident data quality
Promotes and enforces production governance during triage/testing and identifies production failure scenarios, vulnerabilities, and opportunities for improvement
Serves as a subject matter expert for applications within a portfolio, leveraging extensive knowledge of application functionalities and application flows
Assesses and prioritizes research requests, ad hoc reports, and offline incidents at the direction of senior team members and delegates work as needed to team members and peers
Required Qualifications:
Have 5+ years doing production experience
Experience with Horizon CI/CD production stability
Experience with Ansible platform support, deployment reliability
Experience with OpenShift/cloud transition readiness while reducing incident, SLA, vulnerability
Moderate expertise in writing Python, Powershell scripts
Some experience writing SQL query against RDBMS database.
Experience with Splunk, Dynatrace
Have hands on experience with debugging issues
Troubleshooting experience and determine root cause analysis
Must have good analytical skills, critical thinking, self-motivated and resourceful enough to act independent
Desired Qualifications:
Familiarity with database platforms such as Oracle or PostgreSQL
Knowledge of F5 Load balancer, GTM, high availability architectures and disaster recovery strategies.
Certification in Red Hat Linux, DevOps methodologies or related fields.
Familiar with Atlassian products like JIRA, Confluence and Ansible Tower
Any level of experience with Docker/Microsoft Azure is desirable
Skills:
Adaptability
Analytical Thinking
Influence
Production Support
Risk Management
Automation
Collaboration
Innovative Thinking
Result Orientation
Solution Design
Business Acumen
DevOps Practices
Project Management
Solution Delivery Process
Stakeholder Management
Shift:
1st shift (United States of America)Hours Per Week:
40