Volver

Senior ML Engineer (Token Factory)

CompraTica Empleos

EMP:Technology
Czech Republic; Remote - Europe; United Kingdom
Tiempo Completo
Remoto
0 vistas

Descripción

<div class="content-intro"><p><strong>.

Acerca de

Nebius:

Nebius is leading a new era in cloud infrastructure for the global AI economy

We are building a full-stack AI cloud platform that supports developers and enterprises from data and model training through to production deployment, without the cost and complexity of building large in-house AI/ML infrastructure.

</p> <p>Built by engineers, for engineers.

From large-scale GPU orchestration to inference optimization, we own the hard problems across compute, storage, networking and applied AI.

</p> <p>Listed on Nasdaq (NBIS) and headquartered in Amsterdam, we have a global footprint with R&D hubs across Europe, the UK, North America and Israel.

Our team of 1,500+ includes hundreds of engineers with deep expertise across hardware, software and AI R&D.

</p></div><h2 id="The-role" data-renderer-start-pos="1"><strong data-renderer-mark="true">The role</strong></h2> <p data-renderer-start-pos="11">Token Factory is a part of Nebius Cloud, one of the world’s largest GPU clouds, running tens of thousands of GPUs.

We are building an inference & fine-tuning platform that makes every kind of foundation model — text, vision, audio, and emerging multimodal architectures — fast, reliable, and effortless to train & deploy at massive scale.

</p> <div class="ewa-rteLine"><strong>Some directions we currently working on and which you can be a part of:</strong></div> <ul class="ak-ul" data-indent-level="1"> <li> <div class="ewa-rteLine"> <p><strong>Advanced Fine-Tuning:</strong> Enhancing fine-tuning methodologies - both LoRA-based and full-parameter - for cutting-edge LLMs (e.

, GPT-OSS, Kimi K2.

5, DeepSeek V3.

7), focusing on both model quality and training efficiency.

</p> </div> </li> <li> <div class="ewa-rteLine"><strong>Inference Optimization:</strong> Identifying LLM inference bottlenecks to drive production speedups.

This involves building model training and evaluation pipelines in JAX for.

¿Te interesa? Aplicá ahora