AI Infrastructure Engineer

Il y a 2 mois

Ville de Paris, France XpertDirect Temps plein

AI Infrastructure Engineer

Paris, France (Hybrid)

Enterprise AI Infrastructure | Deep Tech | Cloud Platforms


For one of our clients, a rapidly growing Enterprise AI Infrastructure scale-up based in Paris, we are looking for an AI Infrastructure Engineer to help build the high-performance cloud platform powering large-scale AI training, inference, and next-generation enterprise AI applications.


You'll work alongside machine learning engineers, platform engineers, and software developers to create the infrastructure that enables production AI systems to operate reliably, efficiently, and at scale.


What You'll Be Working On

• Designing and operating cloud infrastructure supporting AI model training and large-scale inference

• Building Kubernetes platforms optimised for GPU-intensive workloads

• Developing automation for AI deployment, scaling, and infrastructure management

• Optimising GPU scheduling, resource utilisation, and distributed computing performance

• Supporting machine learning teams with scalable development and production environments

• Improving observability, reliability, and operational performance across AI platforms

• Collaborating with software, platform, and ML engineers to accelerate AI product delivery


Experience Required

• 5 years of experience in Infrastructure Engineering, Platform Engineering, MLOps, DevOps, or Cloud Engineering

• Strong Python development and automation skills

• Hands-on experience managing Kubernetes in production environments

• Experience supporting GPU infrastructure or machine learning platforms

• Good understanding of distributed systems and cloud-native architecture

• Experience building Infrastructure as Code using Terraform or similar tools


Nice to Have

• NVIDIA CUDA optimisation and GPU orchestration experience

• Ray, Kubeflow, KServe, or distributed AI frameworks

• MLflow, Weights & Biases, or AI experiment management platforms

• Experience with Large Language Models or Generative AI infrastructure

• Vector databases and Retrieval-Augmented Generation (RAG) platforms

• Prometheus, Grafana, or OpenTelemetry observability tooling


Why Join?

🚀 Join one of Europe's fastest-growing Enterprise AI Infrastructure companies

🚀 Build the cloud platform powering production AI systems used by enterprise customers worldwide

🚀 Work with cutting-edge GPU computing, distributed AI, and cloud-native technologies

🚀 High level of technical ownership with opportunities to influence platform architecture and engineering strategy

🚀 Collaborate with exceptional AI, platform, and cloud engineers tackling complex technical challenges

🚀 Excellent long-term career progression within a company investing heavily in AI infrastructure and engineering excellence