Inference Systems Engineer: Scale
Enregistrez cette offre et organisez votre recherche
Créez un compte gratuit pour enregistrer des offres d'emploi, créer des alertes et revenir à cette liste depuis votre tableau de bord.
En continuant, vous acceptez nos Conditions d’utilisation & Politique de confidentialité.
Mistral is seeking an experienced engineer to own and evolve the inference foundation that powers production LLM serving and frontier model training. You will optimize the core stack, manage releases, and push upstream improvements to high-performance serving systems.
The role blends engine, platform development, and capacity engineering in a hybrid setup. You will work across CUDA, NCCL, and distributed architectures to deliver low-latency, scalable inference.