Oxylabs
Senior Site Reliability Engineer (SRE) – Infrastructure & Systems (B2B)
About this role
Description: Own and evolve Webshare's production infrastructure. Lead the migration from Docker Swarm to Kubernetes on premises. Maintain high availability across hundreds of servers and ~50 services. Drive observability in cooperation with the development team. Establish and enforce IaC practices, CI/CD pipeline reliability. Participate in the on-call rotation alongside backend engineers. Respond to and lead infrastructure incident resolution, run post-mortems, and drive systematic remediation. Contribute to platform tooling that improves developer experience and reduces infrastructure toil. Collaborate with a strong team of backend engineers on infrastructure ownership. Requirements: Experience building and operating highly available infrastructure at scale (hundreds of servers, dozens of services). Hands-on experience with Kubernetes in self-hosted/bare-metal environments. Proficient in Infrastructure as Code. Proactive problem-solver who keeps the team informed. Scripting and development skills. Benefits: Gross salary starting from 7000 EUR/month + quarterly KPI-based performance bonus. 40+ internal learning options, external conferences, mentorship, and year-round knowledge-sharing. Private health insurance, gym allowance, and a wellness app. Team events, overseas workation, and opportunities to celebrate milestones.