ob / ouahabi-benhenni FR

Expertise · Backend

Backend Architect & Backend Lead

Sober backends that hold the load, deploy without drama and leave room for AI.

Discuss your backend

The problem I solve

A product slows down, deployments are scary, every new feature touches everything else, and adding an AI component threatens the balance of the whole. Most of the time the technology is not to blame; the boundaries are: mixed responsibilities, data without clear ownership, synchronous calls where asynchronous work belongs.

I design readable backend architectures where each layer has a role, each decision is written down with its reason, and an AI service can be added as a component — not as a rewrite.

What I design

APIs & services

APIs and microservices

REST APIs, JWT authentication, RBAC, service boundaries, asynchronous jobs (Celery), integration of AI inference services.

Real-time

Real-time systems

WebSockets (Socket.io, Django Channels), voice and video with WebRTC and a mediasoup SFU, room and role management.

SaaS

Multi-tenant platforms

Tenant isolation, role management, analytics dashboards, monorepos (Turborepo) and shared packages.

How I work

  • Start from real flows — who calls what, how often, with what latency tolerance.
  • Separate responsibilities — signalling and media, reads and writes, sync and async.
  • Write decisions down — every architecture choice documented with its reason and alternatives.
  • Measure under load — an architecture is judged by what it survives in real life.

Examples

  • Real-time voice platform — 218,000 requests in 3 h on 2 vCPU / 2 GB. See the diagram
  • Multi-tenant B2B SaaS — Django 5, DRF, Channels, Celery, JWT, RBAC. See the diagram
  • Inference microservices — Python and ONNX services integrated into an existing platform. See the diagram
  • Video streaming — adaptive FFmpeg encoding, 1,500 concurrent users.
  • Applications for a public administration — mail management (Express API, React, RBAC) and risk mapping.

Read: Handling 218,000 requests on 2 vCPUs.

Frequently asked questions

What does a backend architect do?

They design the server side of an application: service boundaries, data model, APIs, authentication and permissions, asynchronous processing, real-time, deployment. The goal is a system that holds the load, stays secure and can evolve without a rewrite.

Microservices or monolith?

It depends on the team and the product. A well-structured monolith is often the right starting point; microservices are justified when parts of the system have different deployment rhythms, load profiles or technologies — for example an AI inference service next to a web application.

Can you build a high-performing system on small infrastructure?

Yes. A real-time voice platform I designed handled 218,000 requests in 3 hours on a 2 vCPU / 2 GB RAM server. The key: separate responsibilities and size each layer for what it does.

Which backend technologies do you work with?

Python (FastAPI, Flask, Django with DRF, Channels and Celery), Node.js (Express, Socket.io), WebRTC with mediasoup, PostgreSQL, MySQL, MongoDB, Docker, Nginx, Caddy, and the GCP, Azure and OVH clouds.

A transformation project, an architecture to design?

Companies, startups, institutions: describe your context in a few lines. I answer personally, with a first opinion on feasibility and approach.

Propose an engagement LinkedIn ↗ GitHub ↗

Engagements in Algeria, France, Europe, North Africa and remote · Arabic, French, English · contact@ouahabi-benhenni.com