RELX
Site Reliability Engineer
About this role
Are you passionate about building reliable, scalable systems that power critical business solutions? Do you enjoy automating complex operational processes and collaborating across teams to improve performance, resilience, and customer outcomes? About the Business LexisNexis Risk Solutions is the essential partner in the assessment of risk. Within our Business Services vertical, we offer a multitude of solutions focused on helping businesses of all sizes drive higher revenue growth, maximize operational efficiencies, and improve customer experience. Our solutions help our customers solve difficult problems in the areas of Anti-Money Laundering/Counter Terrorist Financing, Identity Authentication & Verification, Fraud and Credit Risk mitigation and Customer Data Management. You can learn more about LexisNexis Risk at About the Role As a Site Reliability Engineer (SRE), you will bridge software development and IT operations by applying software engineering principles to infrastructure and operational challenges. You will focus on building scalable, reliable, and automated systems, helping to improve service performance, reduce manual effort, and ensure high availability across critical platforms.
Responsibilities
- Develop software and scripts to automate operational tasks such as provisioning, configuration, and incident response
- Monitor system performance and establish alerting to support rapid incident detection and resolution
- Reduce Mean Time to Respond (MTTR) through effective monitoring, troubleshooting, and operational improvements
- Scale infrastructure and support capacity planning to ensure high availability and reliability
- Analyze logs, traces, and metrics to identify bottlenecks and optimize system performance
- Define, monitor, and maintain Service Level Indicators (SLIs) and Service Level Objectives (SLOs)
- Manage error budgets to balance innovation, delivery speed, and system reliability
- Collaborate with development teams to enhance service performance, software delivery, and release processes Requirements
- Proficiency in one or more programming languages such as Python, Go, Java, or Bash
- Strong understanding of Linux/Unix operating systems
- Knowledge of networking concepts including TCP/IP
- Experience working with cloud platforms such as AWS, GCP, or Azure
- Familiarity with monitoring and observability tools such as Prometheus and Grafana
- Experience with CI/CD pipelines and automation practices
- Knowledge of configuration management and infrastructure-as-code tools such as Ansible and Terraform
- Strong troubleshooting and problem-solving skills within distributed systems environments Risk benefit statement Learn more about the LexisNexis Risk team and how we work here Primary Location Base Pay Range: Home Based