ML Inference Engineer | Accelerate AI on Custom Accelerator

Il y a 5 jours

Paris, Île-de-France Arago Temps plein

Arago is seeking an experienced engineer to optimize AI model execution on its custom accelerator. You will work across kernels, model execution, multi-device distribution, and inference serving, shaping the software stack around the hardware capabilities.

You will contribute to high-performance ML inference, kernel optimization, and distributed execution while collaborating with hardware, compiler, and runtime teams to drive performance improvements.

#J-18808-Ljbffr