Job Description:
About Us
At Bank of America, we are guided by a common purpose to help make financial lives better through the power of every connection. Responsible Growth is how we run our company and how we deliver for our clients, teammates, communities, and shareholders every day.
One of the keys to driving Responsible Growth is being a great place to work for our teammates around the world. We’re devoted to being a diverse and inclusive workplace for everyone. We hire individuals with a broad range of backgrounds and experiences and invest heavily in our teammates and their families by offering competitive benefits to support their physical, emotional, and financial well-being.
Bank of America believes both in the importance of working together and offering flexibility to our employees. We use a multi-faceted approach for flexibility, depending on the various roles in our organization.
Working at Bank of America will give you a great career with opportunities to learn, grow and make an impact, along with the power to make a difference. Join us!
Position Summary
- Responsible for reliability and support of Cloud Platform including Public Cloud (Azure /AWS /Google) services.
- Monitor and troubleshoot Azure/AWS /Google environment performance issues, connectivity issues, security issues, etc.
- Perform deep dives into systemic and latent reliability issues, incident management, problem management
- Identifying, analyzing, and resolving infrastructure vulnerabilities and application deployment issues.
- Perform blameless RCA, partner with engineering and operation teams across the organization to roll out fixes.
- Identify and drive opportunities to improve automation for the cloud services; scope and create automation for deployment, management, and visibility of our services.
- Evaluating and automating the scaling and capacity requirements within Azure environments
- Engage with engineering teams throughout the full lifecycle from design, engineering, deployment, & operations.
- Partner with risk and compliance teams to bring visibility and implement right controls and policies in the Cloud Platform
- Ensure resiliency during implementation and identify/fix resiliency problems by collaborating with engineering teams
- Be a key stakeholder in the design of cloud services and work with Architecture, engineering, product teams
- Participate in 24x7 on-call coverage follow the sun model
Required Skills:
- BS /MS degree in Computer Science or related technical field involving systems or equivalent practical experience.
- Minimum 8+ years of hands-on experience maintaining cloud platforms on a major cloud service provider.
- Experience working on Azure operations and Administration.
- Azure /Terraform /AWS /Google certifications are a plus
- Strong experience in implementing, monitoring, and maintaining Microsoft Azure solutions, including major services related to Compute, Storage, Network and Security
- Experience with monitoring tools such as Prometheus or Dynatrace, as well as cloud native tools like Azure Monitor and Log Analytics
- Understanding of cost management, inventory management, FinOps model
- Strong understanding and background of working with a complex IAM infrastructure, including Active Directory, Azure AD Connect, Azure AD, and PingIdentity, Okta, or other SSO solutions.
- Advanced knowledge of DNS, DHCP, Kerberos and Windows Authentication
- Experience with IaC with Terraform
- Python, Ansible and shell scripting
- Experience with CI/CD tools such as git and Jenkins, familiarity with using a GitOps model
- Excellent understanding of Linux /Windows operating systems administration
- Systematic problem-solving approach, sense of ownership and drive
- Excellent interpersonal, organizational and communication (written, verbal, and presentation) skills are a must
Desired Skills
- Experience in Terraform, Ansible
- Experience working in a highly available multi-datacenter environment
- Proven ability to work independently with minimal supervision and as part of a team with direct responsibilities.
- Ability to juggle competing priorities and adapt to changes in project scope
Shift:
1st shift (United States of America)
Hours Per Week:
40