Overview As Director of Site Reliability Engineering, you will lead a global SRE organization focused on reliability, scalability, and operational excellence for mission-critical applications within Service Management. You'll partner with senior technology, product, and business leaders to advance observability, incident management, automation, and AI-driven operations. You will shape the SRE strategy across the Corporate Tax and Trade product portfolio and mentor a worldwide team towards high performance and service excellence. This role offers visibility, strategic influence, and the chance to significantly reduce incident impact at scale.
Compensation / Benefits- hybrid work model
- flexible vacation and mental health days
- Grow My Way and skills-first development
- retirement savings with company match
- tuition reimbursement
- Employee Stock Purchase Plan
Responsibilities- Lead 24x7 reliability and operational excellence for a global portfolio of mission-critical apps and services
- Serve as executive escalation point during major incidents with clear communication to stakeholders
- Embed reliability into product design, development, and release processes
- Define and continuously improve observability, incident management, disaster recovery, and resilience practices
- Champion automation and AI-driven operations to enhance efficiency and MTTR
- Shape and evolve SRE strategy, operating model, and engineering behaviors across the portfolio
- Lead and develop a global team of SRE managers and engineers with a culture of innovation and inclusion
Key requirements- 10+ years in site reliability engineering, software/infrastructure/platform leadership
- Proven experience leading global teams for large-scale distributed systems and cloud-native architectures
- Strong expertise in reliability engineering, observability, automation, incident management, DR, and operational readiness
- Experience collaborating with Service Management functions (incident, problem, change, availability, continuity)
- Ability to influence product and technology roadmaps for reliability and performance improvements
- Proven ability to lead transformation with automation, AI operations, and modern SRE practices
- Strong people leadership, talent development, succession planning, and budget management
- Excellent communication and stakeholder management with senior leaders and cross-functional partners
- strong communication
- stakeholder management
- leadership and people development
- observability tools
- automation frameworks
- incident management processes