Prepin
Log in
Bank of America

engineering opportunity

Site Reliability Engineer Lead

This role is responsible for building and leading a team to deliver technology products while defining SRE frameworks and governance models. The lead will drive automation, observability, and resilient architecture across a federated technology ecosystem.

Plano, Texas, United StateshybridFULL_TIME

Posted

About the role

What will you do at Bank of America?

Job Description

At Bank of America, we are guided by a common purpose to help make financial lives better through the power of every connection. We do this by driving Responsible Growth and delivering for our clients, teammates, communities and shareholders every day. Being a Great Place to Work is core to how we drive Responsible Growth.

This includes our commitment to being an inclusive workplace, attracting and developing exceptional talent, supporting our teammates’ physical, emotional, and financial wellness, recognizing and rewarding performance, and how we make an impact in the communities we serve. Bank of America is committed to an in-office culture with specific requirements for office-based attendance and which allows for an appropriate level of flexibility for our teammates and businesses based on role-specific considerations.

At Bank of America, you can build a successful career with opportunities to learn, grow, and make an impact. Join us!

Job Description

This job is responsible for building and leading a team to deliver technology products and services that meet business outcomes. Key responsibilities include developing a technology strategy, ensuring technology solutions comply with applicable standards, promoting design, engineering, and organizational practices, and advocating and advancing modern, Agile solution delivery practices.

Job expectations may include coaching, mentoring, providing feedback and hands on career development, identifying emerging talent, fostering leadership skills, and managing stakeholders. Overview: Seeking a seasoned Site Reliability Engineering (SRE) Leader to drive the reliability, scalability, and performance of critical Infrastructure Automation platforms.

This role will lead the design and implementation of SRE practices across a federated technology ecosystem, ensuring operational excellence through automation, observability, and resilient architecture.

The ideal candidate will bring deep expertise in distributed systems, cloud-native infrastructure, SaaS application support and DevOps/SRE principles, along with strong leadership and collaboration skills to influence cross-functional engineering and Production management teams and drive continuous improvement in service reliability.

Responsibilities

  • SRE Strategy & Governance: Define and implement SRE frameworks, including SLIs/SLOs/SLAs, error budgets, and incident response protocols.
  • Establish governance models for reliability engineering across distributed teams.
  • Champion a culture of observability, proactive monitoring, and continuous feedback loops.
  • Reactive & Proactive Problem Management: Lead root cause analysis (RCA) and post-incident reviews to identify systemic issues and prevent recurrence.
  • Implement proactive problem detection using telemetry, anomaly detection, and trend analysis.
  • Collaborate with engineering and operations teams to eliminate toil and reduce incident frequency and impact.
  • Capacity & Performance Management: Develop and maintain capacity models to ensure systems scale efficiently with business demand.
  • Monitor performance trends and lead optimization efforts across infrastructure and applications.
  • Partner with finance and engineering teams to align capacity planning with cost and growth objectives.
  • Platform Reliability & Automation: Drive automation of operational tasks including deployments, scaling, and recovery.
  • Integrate reliability tooling with CI/CD pipelines, ITSM platforms (e.g., ServiceNow), and observability systems.
  • Incident Management & Operational Excellence: Oversee major incident response, escalation, and communication processes.
  • Develop and maintain runbooks, playbooks, and escalation protocols.
  • Drive continuous improvement through blameless retrospectives and operational reviews.
  • Technical Leadership: Serve as a senior technical advisor and thought leader in SRE and platform engineering.
  • Mentor and guide SRE teams and partner with engineering leaders across the enterprise.
  • Provide input on staffing, tooling strategy, and budget planning for reliability initiatives.
  • Managerial

Responsibilities

  • This position may also have responsibilities for managing associates.
  • At Bank of America, all managers at this level demonstrate the following responsibilities, in addition to those specific to the role, listed above.
  • Opportunity & Inclusion Champion: Models an inclusive environment for employees and clients, aligned to company Great Place to Work goals.
  • Manager of Process & Data: Demonstrates deep process knowledge, operational excellence and innovation through a focus on simplicity, data based decision making and continuous improvement.
  • Enterprise Advocate & Communicator: Communicates enterprise decisions, purpose, and results, and connects to team strategy, priorities and contributions.
  • Risk Manager: Ensures proper risk discipline, controls and culture are in place to identify, escalate and debate issues.
  • People Manager & Coach: Provides inspection, coaching and feedback to motivate, differentiate and improve performance.
  • Financial Steward: Actively manages expenses and budgets in alignment with objectives, making sound financial decisions.
  • Enterprise Talent Leader: Assesses talent and builds bench strength for roles across the organization.
  • Driver of Business Outcomes: Delivers results by effectively prioritizing, inspecting and appropriately delegating team work.
  • Required

Qualifications

  • 10+ years of experience in systems engineering, DevOps, or SRE roles in large-scale environments.
  • Deep understanding of Linux/Unix & Windows systems, networking, and distributed computing.
  • Proven experience with observability stacks (e.g., Dynatrace, Grafana, Splunk, OpenTelemetry).
  • Expertise in infrastructure-as-code and automation tools (e.g., Terraform, Ansible, Python).
  • Strong knowledge of cloud platforms and container orchestration (Kubernetes).
  • Demonstrated success in leading incident response and driving systemic improvements.
  • Experience with capacity planning, performance tuning, and cost optimization.
  • Excellent communication and stakeholder management skills, including executive engagement.
  • Desired

Qualifications

  • Experience with ITIL/ITSM processes and integration with platforms like ServiceNow.
  • Familiarity with security and compliance in regulated industries (e.g., financial services).
  • Background in performance engineering and infrastructure analytics.
  • Experience developing dashboards and metrics for operational health and reliability.
  • Skills: Influence Risk Management Solution Design Stakeholder Management Technical Strategy Development Analytical Thinking Application Development Collaboration Result Orientation Solution Delivery Process Agile Practices Architecture Automation Data Management DevOps Practices Shift: 1st shift (United States of America) Hours Per Week: 40 Bank of America is committed to help employees through the transition period when they’re displaced as a result of a workforce reduction, realignment or similar measure.
  • Please review the resume writing and interviewing tips provided below to help prepare you for your next career opportunity.
  • Getting started Regardless of the position you are interested in, the starting points to building your resume are the same: 1.
  • Determine the job or types of jobs you want to do and research their responsibilities and qualifications.
  • 2.
  • Think about why you can do the job and make a list of your skills that are relative to the job.
  • 3.
  • Identify experiences or accomplishments that show your proficiency in the skills required for the job.
  • 4.
  • Summarize your abilities, accomplishments and skills into a brief, concise document.
  • Considerations when writing a resume • Do be brief.
  • Resumes should be 1-2 pages in length.
  • • Do be upbeat and active in your wording.
  • • Do emphasize what you have done clearly and concretely.
  • • Do be neat and well organized.
  • • Do have others proofread and critique your resume.
  • Spell check.
  • Make it error free.
  • • Do use high quality, white or light colored 8½ x 11 paper.
  • Use a laser printer if possible.
  • • Don't be dishonest, always tell the truth about yourself in the most flattering light.
  • • Don't include salary history or requirements.
  • • Don't include references.
  • • Don't include accomplishments that do not support your professional goals.
  • • Don't include anything that isn't relevant.
  • (For example, don't mention your fondness for swimming unless you want to work on the water.)
  • • Don't use italics, underlining, shadows or other fancy treatments.
  • Seven steps to a successful interview 1.
  • Anticipate –Put yourself in the interviewer's position.
  • What do you believe the interviewer is most interested in?
  • Why do you think you have been invited to interview?
  • 2.
  • Research –What are the primary functions of the line of business?
  • What are the success factors for the job?
  • Is there a job description available?
  • 3.
  • Assess –Think about your skills, abilities, knowledge, interests, traits, values and accomplishments.
  • Match them to what you know about the job.
  • Consider which ones you should highlight.
  • 4.
  • Prepare Answers –Think about what the interviewer may ask, determine what the best answer is and write it down.
  • 5.
  • Prepare Questions – Interviewing is a two-way street.
  • By asking thoughtful questions, you communicate your interest and learn a lot about the job.
  • Choose two or three questions to ask your interviewer.
  • Avoid asking a lot of questions about vacation time or breaks.
  • 6.
  • Practice – It may seem awkward, but it is the best way to come across well in an interview.
  • Practice your own "great responses" with others or in front of a mirror until you appear relaxed and at ease.
  • 7.
  • Follow-up – Send a brief follow-up letter to the interviewer.
  • Keep in mind that the many job searchers will not send a follow-up letter.
  • Sending one can become a competitive advantage.
  • Pay Transparency - https://careers.bankofamerica.com/en-us/pay-transparency Privacy Statement - https://careers.bankofamerica.com/en-us/privacy-notice

Which skills does this role require?

Site Reliability EngineeringLinuxUnixWindowsNetworkingDistributed systemsObservabilityInfrastructure as codeTerraformAnsiblePythonCloud platformsKubernetesCapacity planningPerformance tuningStakeholder managementDevOpsInfrastructure AutomationDistributed SystemsCloud-nativeSaaSDynatraceGrafanaSplunkOpenTelemetryInfrastructure-as-codeITILITSMServiceNowCapacity PlanningPerformance EngineeringAgileRoot Cause AnalysisGovernanceSystem ArchitectureFinancial ServicesLeadership

Make your next move

Build a shortlist and prepare

Identify the requirements you can demonstrate, then choose examples from your work to discuss with the hiring team.

Review the responsibilities and requirements before adding an opening to your shortlist.

Role information can change. Confirm current details on the original application page.

Product

AI Candidate AgentCompaniesBrowse JobsDeep ProfileSkill AssessmentOpportunity Matching
Prepin.ai

© 2026 Prepin | All rights reserved.