About the Role
Senior Storage Engineer with a strong SRE mindset to own the reliability, availability, and operational excellence of large-scale SAN, NAS, and object storage platforms powering mission-critical production systems. This is a hands-on, production-focused role for engineers who thrive in high-impact environments where uptime, data durability, and recovery speed matter. You will act as a final-tier escalation engineer, lead major incident response, and manage storage systems, namely NetApp, EMC, Cisco, Brocade & Hitachi, that are designed to fail safely, recover predictably, and improve continuously.
Responsibilities
- Engineers reliability in the infrastructure.
- Responsible for the stability, availability, and reliability of the Bank’s globally deployed storage infrastructure across AMRS, APAC, EMEA.
- Production-first role operating massive SAN, NAS, and object storage platforms that support business-critical applications at scale.
- Own availability, performance, and failure recovery.
- Operate at the front line of incidents, lead deep technical investigations, and continuously push the platform toward being boring, predictable, and self-healing.
- Own uptime, latency, durability, and capacity as measurable outcomes.
- Define and drive SLIs, SLOs, and error thresholds for critical storage services.
- Work closely with vendors, engineering teams, data center operations, application owners, and incident management teams to resolve complex issues, reduce repeat incidents, and improve overall platform reliability.
- Key resource of the 24/7 Global Storage Operations Team.
Requirements
- Drive reliability and availability of storage services using SRE principles, including monitoring, SLIs/SLOs, and proactive issue detection.
- Strive towards a Never Down environment.
- Automate routine storage operations and reduce manual toil through scripting and process improvements.
- Deep hands-on operational expertise with enterprise storage platforms: EMC PowerMax, NetApp (ONTAP – SAN & NAS), Hitachi Storage platforms / HCP, Cisco SAN switches (MDS).
- Proven experience providing Level 3 storage support in large, complex, mission critical environments.
- Strong experience managing Major incidents (P1/P2), Bridge calls and global outage coordination, Root cause analysis (RCA) and problem management.
- Solid understanding of SAN, NAS, and object storage concepts, including performance, replication, snapshots, zoning, multipathing, and capacity management.
- Hands-on operational exposure to Linux and Windows environments from a storage integration and support perspective.
- Working knowledge of networking fundamentals (FC, IP, zoning, VLANs) and Active Directory interactions with storage platforms.
- Excellent verbal and written communication skills, with the ability to clearly engage: Incident management teams, Application owners, Vendors, Senior stakeholders.
- Demonstrated ability to work independently, make sound technical decisions, and handle high pressure operational scenarios.
- 10+ years of experience in enterprise storage technologies and production operations.
Skills
- SRE principles
- Monitoring
- SLIs/SLOs
- Proactive issue detection
- Automation
- Scripting
- Process improvements
- Enterprise storage platforms
- EMC PowerMax
- NetApp (ONTAP – SAN & NAS)
- Hitachi Storage platforms / HCP
- Cisco SAN switches (MDS)
- Level 3 storage support
- Major incident management (P1/P2)
- Bridge calls
- Global outage coordination
- Root cause analysis (RCA)
- Problem management
- SAN concepts
- NAS concepts
- Object storage concepts
- Performance
- Replication
- Snapshots
- Zoning
- Multipathing
- Capacity management
- Linux
- Windows
- Networking fundamentals (FC, IP, zoning, VLANs)
- Active Directory
- Communication skills
- Independent work
- Technical decision making
- High pressure operational scenarios
- Perl
- Python
- Shell scripting
- Operational tooling
- Health checks
- Reporting
- Automation frameworks
- Orchestration tools
- Runbooks
- Auto remediation
- Operational workflows
- Incident reduction
- Error budgets
- Improving MTTR
- Operational resilience
- Eliminating toil
- Operational documentation
- Standard operating procedures
- ITIL / Change Management processes
- Change risk assessment
- Technical approvals
- Post change validation
- MS Office tools
Work Type
- in-office culture
Experience Level
- Senior
- 10+ years of experience
Benefits
- competitive benefits to support their physical, emotional, and financial well-being
About the Company
- At Bank of America, we are guided by a common purpose to help make financial lives better through the power of every connection.
- Responsible Growth is how we run our company and how we deliver for our clients, teammates, communities and shareholders every day.
- One of the keys to driving Responsible Growth is being a great place to work for our teammates around the world.
- We’re devoted to being a diverse and inclusive workplace for everyone.
- We hire individuals with a broad range of backgrounds and experiences and invest heavily in our teammates and their families by offering competitive benefits to support their physical, emotional, and financial well-being.
- Bank of America believes both in the importance of working together and offering flexibility to our employees.
- We use a multi-faceted approach for flexibility, depending on the various roles in our organization.
- Working at Bank of America will give you a great career with opportunities to learn, grow and make an impact, along with the power to make a difference.
- Core Technology Infrastructure: Believes diversity makes us stronger so we can reflect, connect and meet the diverse needs of our clients and employees around the world.
- Is committed to building a workplace where every employee is welcomed and given the support and resources to perform their jobs successfully.
- Wants to be a great place for people to work and strives to create an environment where all employees have the opportunity to achieve their goals.
- Provides continuous training and development opportunities to help employees achieve their career goals, whatever their background or experience.
- Is committed to advancing our tools, technology, and ways of working to better serve our clients and their evolving business needs.
- Believes in responsible growth and is dedicated to supporting our communities by connecting them to the lending, investing and giving them what they need to remain vibrant and vital.
- Partnering Locally Learn about some of the ways Bank of America is making a difference in the communities we serve.
- Global Impact Learn about the six areas that guide Bank of America’s efforts to help make financial lives better for customers, clients, communities and our teammates.
- Opportunity and Inclusion Each employee brings unique skills, background and opinions. We see opportunity and inclusion as our platform for innovation and a key component in our success.
- Our Values Learn about our four values that represent what we believe.
