Overview In this role, you will lead a Core SRE team within Business Services, ensuring reliable, scalable cloud environments and efficient incident response. You'll shape reliability culture, drive automation, and collaborate with Dev, Security, and Product teams to improve performance and cost. You'll steward post-mortems, RCAs, and cross-team initiatives to reduce toil and boost service resilience. This is a hands-on management role at LexisNexis Risk Solutions, offering scope to influence cloud-native platforms across the organization.
Compensation / Benefits- annual incentive bonus
- country-specific benefits
- inclusive hiring process accommodations
Responsibilities- Lead and grow a team of SREs, conduct 1:1s, performance reviews, and career development
- Drive hiring, onboarding, and team capacity planning
- Set team goals, prioritize backlog, and guide sprint planning
- Foster blameless post-incident culture and cross-team collaboration (Dev, Security, Product)
- Lead reliability initiatives across infrastructure and services
- Drive incident response and continuous service improvement
- Champion automation and operational excellence across the platform
- Support scalable, secure, and resilient cloud-native environments
Key requirements- Expert Kubernetes knowledge (cluster architecture, upgrades, autoscaling, security hardening, scale troubleshooting)
- Terraform expertise (modular IaC design, state management, multi-environment provisioning, policy-as-code)
- Deep Azure Cloud knowledge (compute, networking, identity/AAD, storage, cost optimization)
- Experience designing/scaling CI/CD with GitHub Actions, release strategies, rollback automation
- Observability platforms experience (Prometheus, Grafana, OpenTelemetry, SLO/SLA, error budgeting)
- Strong automation skills to eliminate toil and enable self-healing systems
- Proficiency in Python, Bash, and/or PowerShell
- Networking expertise (TCP/IP, DNS, load balancing, VPN, cloud networking)
- SRE/DevOps/Infrastructure leadership experience; proven incident response leadership
- leadership
- cross-team collaboration
- problem-solving
- Kubernetes
- Terraform
- Azure