Parallel Wireless
Senior/Principal Local LLM & Generative AI Platform Engineer
About this role
Description: Lead the development and operation of a secure local large-language-model (LLM) platform for Parallel Wireless. Collaborate with various teams to prioritize use cases and translate them into product requirements. Build and optimize a modular inference and model-gateway layer with stable APIs and model routing. Ensure security and access control throughout the platform, integrating with company identity systems. Establish evaluation datasets and automated testing for quality and performance metrics. Design and implement observability for workflows and integrate the platform into existing tools. Requirements: BSc or MSc in Computer Science, Engineering, Data Science, or equivalent experience. 7+ years of experience in software, ML platform, or infrastructure engineering, with recent LLM experience. Strong Python skills; experience with Go, Java, or C/C++ is a plus. Understanding of transformer-based models and production inference techniques. Experience with RAG or enterprise-search systems and defining LLM evaluations. Proficient in deploying containerized services using Docker and Kubernetes. Knowledge of distributed systems, API security, and data lifecycle controls. Familiarity with Git, CI/CD, and observability practices. Benefits: Opportunity to work on cutting-edge AI technology in a pioneering company. Collaborative environment with cross-functional teams. Potential for professional growth and development in AI and telecommunications.