Job Title
Senior Manager, Site Reliability Engineering (SRE)
Intro:
At FIS, our technology and our people are moving forward. We advance the way the world pays, banks and invests. We believe in building inclusive, diverse teams. Together, we innovate to help our colleagues, clients and communities succeed. If you’re ready to grow your career and make an impact in fintech, we have one question: Are you FIS?
About the Role
We are seeking a Senior Manager, Site Reliability Engineering to lead the strategy, execution, and continuous evolution of reliability practices for mission-critical payment and financial technology platforms. In this role, you will combine deep technical expertise with people leadership, guiding high-performing SRE teams while driving operational excellence, platform resilience, and engineering effectiveness. You will partner with senior technology leaders to establish reliability goals, influence architecture decisions, and ensure highly available, secure, and scalable services. Success in this role is measured through improved platform reliability, reduced operational risk, accelerated delivery, and the development of strong engineering talent.
What You Will Be Doing
- Lead and develop a team of Site Reliability Engineers, fostering engineering excellence, accountability, and continuous learning.
- Define and execute the reliability vision, roadmap, and operating model across critical payment and transaction platforms.
- Own service reliability outcomes, ensuring compliance with SLAs, SLOs, availability targets, and regulatory requirements.
- Drive reliability architecture standards across cloud infrastructure, applications, platforms, and distributed systems.
- Establish scalable observability capabilities utilizing metrics, logging, tracing, alerting, and service health monitoring.
- Lead major incident management activities and executive communications during critical production events.
- Drive root-cause analysis, corrective action planning, and systemic reliability improvements across engineering organizations.
- Partner with engineering, product, platform, security, and operations leaders to influence technology strategy and platform modernization.
- Champion automation, self-healing capabilities, and platform engineering initiatives that reduce operational toil and risk.
- Oversee capacity planning, resilience testing, disaster recovery readiness, and operational risk management programs.
- Establish performance metrics, engineering KPIs, and reporting mechanisms that measure reliability and operational effectiveness.
- Mentor senior engineers, technical leads, and managers while building leadership succession and organizational capability.
Required Qualifications
- Demonstrated experience leading Site Reliability Engineering, Production Engineering, Platform Engineering, or Infrastructure Engineering teams.
- Deep expertise designing, building, and operating large-scale distributed systems in production environments.
- Strong leadership experience managing technical teams, setting priorities, driving execution, and developing engineering talent.
- Expertise in observability and reliability engineering practices including SLIs, SLOs, error budgets, alerting, and incident management.
- Strong experience with cloud platforms such as AWS, Azure, or GCP and cloud-native architecture patterns.
- Experience implementing infrastructure-as-code, automation, and platform engineering solutions at enterprise scale.
- Proven success operating mission-critical systems within Payments, FinTech, Banking, or highly regulated industries.
- Strong knowledge of Linux, Windows, enterprise platforms, databases, networking, and complex systems troubleshooting.
- Experience influencing senior stakeholders and driving cross-functional alignment across multiple teams and business units.
- Bachelor’s degree in computer science, Engineering, Information Technology, or equivalent practical experience.
Preferred Qualifications
- Experience leading global or geographically distributed SRE and platform engineering teams.
- Strong automation and development skills using Python, Go, Bash, Ansible, Terraform, or similar technologies.
- Experience implementing CI/CD platforms, release automation, and DevSecOps practices.
- Knowledge of Oracle, distributed databases, messaging platforms, and high-volume transaction processing systems.
- Experience modernizing legacy financial systems into cloud-native or hybrid architectures.
- Industry certifications related to cloud platforms, reliability engineering, security, or infrastructure management.
What We Offer you
At FIS, you can grow your career as far as you want to take it. Here’s what else we offer:
- Opportunities to make an impact in fintech
- Personal and professional learning
- Inclusive, diverse work environment
- Resources to give back to your community
- Competitive salary and benefits
Privacy Statement
FIS is committed to protecting the privacy and security of all personal information that we process in order to provide services to our clients. For specific information on how FIS protects personal information online, please see the Online Privacy Notice.
Sourcing Model
Recruitment at FIS works primarily on a direct sourcing model; a relatively small portion of our hiring is through recruitment agencies. FIS does not accept resumes from recruitment agencies which are not on the preferred supplier list and is not responsible for any related fees for resumes submitted to job postings, our employees, or any other part of our company.
#pridepass