Invisible Technologies
Site Reliability Engineer (Contract, Rotation-Based)
About this role
Description: Join a 24/7 incident response rotation as a Site Reliability Engineer supporting a production platform for key clients. Act as a first responder for critical incidents, triaging and stabilizing issues at the infrastructure level. Requirements: Hands-on experience with Kubernetes, RabbitMQ, and PostgreSQL in an enterprise setting, preferably in Financial Services. Strong knowledge of Azure; familiarity with GCP or AWS is a plus. Ability to diagnose issues from system logs and troubleshoot production systems under time pressure. Clear communication skills during live incidents. Benefits: Flexible contractor role with a commitment of 10+ hours per week and rotational weekend coverage. Opportunity to work in a dynamic environment at the intersection of AI and human ingenuity.