Confluent
Staff Software Engineer I - SRE
About this role
Description: Join Confluent to enhance data streaming technology and improve reliability in a multi-cloud environment. Role involves proactive reliability engineering and incident management, focusing on preventing incidents and improving response practices. Requirements: 10+ years in SRE, incident management, or reliability engineering with cloud experience (AWS, GCP, or Azure). Expertise in incident management tools (Rootly, PagerDuty) and strong understanding of distributed systems and failure modes. Experience with observability, Kubernetes, CI/CD pipelines, and familiarity with SLO/SLA frameworks. Proven track record as a trusted advisor and experience in large organizations (500+ engineers). Benefits: Opportunity to work in a diverse and inclusive environment that values different perspectives. Freedom to innovate with modern CI/CD tools and AI-assisted workflows.