Head of Infrastructure

Il y a 1 jour

Paris, Île-de-France Diabolocom Temps plein 120 000 € - 180 000 € Contrat

Build the infrastructure behind the next generation of AI-powered customer communications

Diabolocom is building an AI-first communication platform where AI agents, human agents, business workflows and communication channels work together seamlessly.

As we enter our next phase of growth, we are looking for a Head of Infrastructure to own and scale the infrastructure foundation behind this platform.

This is a unique opportunity to lead an infrastructure organization combining carrier-grade voice systems, on-premise data centers, distributed systems and AI workloads, while building the operational maturity required to support 2x growth.

You will report directly to Sergei, our CTO, and lead the infrastructure organization — owning operational excellence, team development and execution of infrastructure initiatives, while actively contributing to technical strategy and architecture decisions.

Why this role is different

Most infrastructure leadership roles today focus on managing cloud platforms and abstracted services. Here, you will own the foundations of a real-time, AI-first communication platform — from hardware and networks to telecom systems, distributed services, and operational processes.

Our philosophy is simple: we build and operate our infrastructure ourselves. 100% of our production infrastructure is managed by our teams, giving us full technical ownership, deep system knowledge, and the ability to solve complex challenges without relying on external vendors. This allows us to move faster, make better architectural decisions, and build systems tailored to our needs.

You will work across:

  • Telecom infrastructure — we are a global voice carrier powering real-time communications across multiple markets. We manage carrier relationships, routing decisions, and voice traffic flows end-to-end, with the expertise and tooling built internally.
  • On-premise infrastructure — around 95% of our production infrastructure runs on our own hardware across several data centers around the Paris and Frankfurt areas, with servers, networking, and connectivity fully managed by our teams.
  • Cloud infrastructure — while the majority of our production runs on-premise, we strategically leverage cloud services for remote locations and specific workloads where they provide faster deployment and operational efficiency, particularly for smaller-scale deployments.
  • Stateful systems — PostgreSQL, ClickHouse, Kafka, RabbitMQ, Redis, and others powering high-volume workloads.
  • Distributed systems — the platform is deployed in a fully redundant way across multiple data centers, with the ability to sustain the loss of a data center without impacting workloads.
  • AI infrastructure — GPU workloads and compute infrastructure supporting our AI capabilities.
  • Storage system — Ceph provides our main object and volume storage layer, supporting all other services across the infrastructure.

This is infrastructure engineering where reliability, latency, scalability, and operational excellence directly impact the product.

You will not simply maintain existing systems. You will help shape the architecture, processes, and team structure required for the next stage of our growth — taking ownership of critical infrastructure decisions from the ground up.

Your mission

Your goal is to make infrastructure ready for the next stage of scaling up by building a strong operational foundation across all the infrastructure domains.

You will:

  • Own infrastructure reliability and operations across the whole stack: incident management, monitoring, alerting, post-mortems and prevention of recurring failures.
  • Improve operational visibility across infrastructure teams: identify bottlenecks, improve processes and drive continuous improvements.
  • Improve capacity planning processes, connecting business growth and infrastructure needs.
  • Lead the evolution of our network and telecom infrastructure: topology, connectivity, carrier relationships, routing, redundancy and voice platform performance and reliability.
  • Own our on-premise infrastructure operations: data centers, hardware, storage, connectivity and deployment practices.
  • Drive scalability and reliability of our stateful infrastructure: PostgreSQL, ClickHouse, Kafka, RabbitMQ and other critical systems.
  • Support the evolution of our AI infrastructure, including GPU workloads and compute clusters.
  • Drive infrastructure improvements with automation, observability and tooling.
  • Partner with engineering teams to ensure applications run reliably on our platforms.
  • Lead and grow the Infrastructure team across Networking, SRE, Stateful Systems, Voice, and Data C