Your missionYou will expand the capabilities of Lyceum's AI inference platform, the first EU-sovereign inference cloud.
You'll own the features that customers interact with directly: model serving configurations, API surface, framework integrations, and developer experience.
This means understanding what customers need, building it fast, and making sure it works reliably at scale.
Your focusFeature development: Design and ship new platform capabilities - from supporting new model architectures and serving frameworks to building out API features that customers are asking for.
Customer-facing engineering: Work closely with customers and the commercial team to understand real-world usage patterns, translate feature requests into technical designs, and iterate based on feedback.
Your KPIsNumber of platform features shippedTime from customer request to feature availabilityBreadth of supported models, frameworks, and deployment configurationsYour profileWe consider candidates from diverse backgrounds, with a deep love for technical challenges and the desire to take on ownership beyond what's reasonably expected.
You're someone who stays close to the rapidly evolving open-source AI ecosystem and gets energy from turning emerging tools into production-grade platform capabilities.
Requirements3+ years of experience in software engineering, with a focus on backend or infrastructure systemsStrong proficiency in Go and PythonHands-on experience with at least one ML inference serving framework (vLLM, TGI, etc)Solid understanding of how large language models and other AI models are deployed and served in productionExperience working with REST/gRPC APIs and designing developer-facing interfacesNice to haveFamiliarity with GPU scheduling, batching strategies, or inference optimisation (quantisation, speculative dec.