Mirantis
Senior Software Engineer, AI Infrastructure
About this role
Description: Design and build LLM serving infrastructure on Kubernetes, including deployment, GPU scheduling, scaling, and model lifecycle management. Package the platform for enterprise environments with Helm-based installs and upgrades. Integrate the serving layer with the platform's API gateway, identity, and metering services. Build observability for GPU inference in production, including serving metrics and GPU telemetry. Contribute to a multi-service codebase and influence engineering direction through design documentation and reviews. Requirements: 5+ years of software engineering experience in infrastructure, platform, or distributed systems. Deep hands-on Kubernetes experience, including building and operating production workloads and Helm charts. Experience with GPU workloads or LLM inference, or strong adjacent systems experience with a quick learning ability. Strong Go programming skills and solid CI/CD and infrastructure-as-code skills. Fluency with AI-assisted development tools as part of daily engineering workflow. Comfortable with high autonomy in a small, remote-first team. Benefits: Work with a leader in the cloud infrastructure industry and passionate colleagues. Engage in cutting-edge, open-source innovation. Thrive in a collaborative and growth-oriented environment. Opportunities for professional development and training, including conferences. Participate in company outings, happy hours, hackathons, and tech talks. Competitive compensation package with a strong benefits plan.