Job Description:
At Bank of America, we are guided by a common purpose to help make financial lives better through the power of every connection. We do this by driving Responsible Growth and delivering for our clients, teammates, communities and shareholders every day.
Being a Great Place to Work and providing a culture of caring is core to how we drive Responsible Growth. We are intentional about fostering an inclusive workplace where every teammate has the opportunity to succeed, build a career and contribute to our shared success. This includes attracting and developing exceptional talent, recognizing and rewarding performance, and supporting our teammates’ physical, emotional, and financial wellness through affordable, competitive and flexible benefits.
We value the unique perspectives individuals bring from all backgrounds and career paths - whether shaped by military service, community college education, or a wide range of work and life experiences. These journeys foster resilience, leadership and innovation, strengthening our workforce and positively impact the communities we serve.
Bank of America is committed to an in-office culture that supports collaboration, engagement, and career development. Our approach includes clear in-office expectations, while providing an appropriate level of flexibility based on role-specific responsibilities and business needs.
At Bank of America, you can build a successful career with opportunities to learn, grow, and make an impact. Join us!
Job Description:
This job is responsible for providing front-line support to end users, responding to issues related to incidents and problem management governance for multiple applications, and leading triage activities on all business impacting incidents. Key responsibilities include ensuring compliance with incident management and problem management policies and procedures, serving as a focal point for the customer, client, and associate experience, restoring complex production incidents under tight Service Level Agreements, and pursuing root cause and problem resolution follow ups.
This role is critical to sustain Horizon CI/CD production stability, Ansible platform support, deployment reliability, and OpenShift/cloud transition readiness while reducing incident, SLA, vulnerability, DR, and support-coverage risk across enterprise application teams.
Responsibilities:
Leads production support triage efforts, manages bridge line troubleshooting, engages in technical research, and escalates issues to leadership as needed
Ensures all impacts are accurately recorded and documented in the system of record, oversees that documents and wikis are updated and available for use during triage, and supports the documentation of application flows, upstream/downstream impacts during outages, the customer experience, and contacts for support needs
Identifies and/or validates business impacts through interpretation of monitors, dashboards, and logs to communicate with leadership and vendors
Manages activities to identify incident root cause, resolution, preventative actions, and change requests, and reports on incident data quality
Promotes and enforces production governance during triage/testing and identifies production failure scenarios, vulnerabilities, and opportunities for improvement
Serves as a subject matter expert for applications within a portfolio, leveraging extensive knowledge of application functionalities and application flows
Assesses and prioritizes research requests, ad hoc reports, and offline incidents at the direction of senior team members and delegates work as needed to team members and peers
Required Qualifications:
Strong knowledge in .NET, ASP.NET, IIS, Windows, SQL Server
Hands on experience in .NET, ASP.NET, IIS, Windows, SQL Server
Experience in working in Production Support environment, triage and RCA experience
Knowledge on ITSM Remedy, Splunk, Dynatrace
Have 4+ years doing production support experience.
Provide in-depth problem analysis of application, system errors, and performance issues
Proactive application stability analysis – investigate performance concerns, system errors, improve maintenance processes, automation, and resolution of open issues
Ownership of problem root cause analysis and remediation
Chronic issue investigation for critical application issues
Lead proactive monitoring review and implementation
On call support for production releases
Skills:
Adaptability
Analytical Thinking
Influence
Production Support
Risk Management
Automation
Collaboration
Innovative Thinking
Result Orientation
Solution Design
Business Acumen
DevOps Practices
Project Management
Solution Delivery Process
Stakeholder Management
Shift:
1st shift (United States of America)Hours Per Week:
40