Skip to content
View fpaupier's full-sized avatar
🎯
Focus
🎯
Focus

Sponsoring

@cyrilou242

Block or report fpaupier

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
fpaupier/README.md

Twitter Personal Website Site

Header

Building AlexandrIA · Fractional Head of AI

LLM products · agents · RAG · hybrid retrieval · vLLM · LiteLLM · Langfuse · EU-sovereign cloud

Python TypeScript Next.js FastAPI LiteLLM vLLM Langfuse PostgreSQL Docker Kubernetes

I ship AI products that people actually use. I founded and operate AlexandrIA, a research co-pilot that turns messy questions into visual knowledge maps and client-ready deliverables. On the side I take a very small number of Fractional Head of AI mandates with post-seed startups that need someone to own the AI vision and make it real in production.

AlexandrIA

Live product: tryalexandria.fr · status

Knowledge workers use it to frame a problem, curate sources, and package insights. It is in production, with billing, EU-sovereign hosting, and recurring paying users. Built and operated end-to-end: multi-agent orchestration, RAG, LLM-as-judge evals, product, and infra.

Fractional Head of AI

For post-seed teams that already have a product and need an AI lead who has shipped, not a slide deck. I help set the roadmap, choose the stack, and stay close to delivery. Capacity is intentionally limited.

What I actually work in

Production LLM products, agents, RAG, knowledge graphs, and the infra that keeps them up.

  • Product / AI: TypeScript, Next.js, Python, FastAPI, multi-agent orchestration, tool-calling, RAG, hybrid retrieval, LiteLLM, vLLM
  • Eval & observe: LLM-as-judge, Langfuse, OpenTelemetry
  • Data / backend: Postgres, Supabase, embeddings
  • Ship & operate: Docker, Kubernetes, Ansible, GitHub Actions, EU cloud (OVHcloud, Scaleway), Cloudflare
  • Also in the field: on-prem / air-gapped copilots (self-hosted vLLM, Keycloak), knowledge graphs

I do not train computer-vision models in PyTorch day to day. I design, ship, and run LLM systems.

Background

Centrale Nantes, MSc Signal Processing & Statistics (top 10%). Engineering manager and ML engineer at the French DoD, then Schneider Electric and Kpler. Now founder of AlexandrIA, and AI lead on client production systems.

Guest Lecturer, LLM Pretraining & Post-training, SCAI (Sorbonne Université) & Centrale Nantes. CKA (Certified Kubernetes Administrator).

Teaching & speaking

Selected talks: TensorFlow World (Santa Clara, 2019) · AGIR 2024 for the Gendarmerie Nationale (LLM industrialization panel) · GenAI delivery at Qonto HQ (2025) · LLM sovereignty & supply chains, Centrale Nantes (2025). Catalog: fpaupier.fr/speaking

Post Mortem

I created, produced, and hosted Post Mortem for four years as producer and interviewer. Engineers walk through real incidents: outages, cyber, ML in production. 24 episodes, 2020–2024. On hold while I focus on shipping.

#24 The New Face of Conflict: AI in Warfare with COL ANTAL
#24 AI in warfare
COL John Antal · 60 min
#23 D'entrepreneur à investisseur — Philippe Laval
#23 From founder to investor
Philippe Laval · Sinequa, Jolt Capital
#19 DevSecOps at the US Air Force with Nicolas Chaillan
#19 DevSecOps at the US Air Force
Nicolas Chaillan · 31 min

Catalog: Spotify · notes · Apple Podcasts

Writing

I wrote a run of how-to pieces (Medium, fpaupier.fr/writings). Not currently writing. Focused on execution.

How Apache NiFi works — freeCodeCamp
How Apache NiFi works
freeCodeCamp · 1.3k claps
Unleash multiprocessing with Python and gRPC
Multiprocessing with Python and gRPC
Medium
How to communicate effectively to a non-technical stakeholder
Talking to non-technical stakeholders
Medium

Also: Taxonomy of leading Generative AI architectures (2024) · Practical insights for LLM fine-tuning and evaluation (2024) · Effective domain modeling

Connect

LinkedIn · tryalexandria.fr · fpaupier.fr


Credits

Banner designed using the github profile header generator, octocat shaped with the myOctocat creator.

Popular repositories Loading

  1. RapLyrics-Scraper RapLyrics-Scraper Public

    Data sourcing and pre-processing for raplyrics.eu - A rap music lyrics generation project

    Python 71 14

  2. tensorflow-serving_sidecar tensorflow-serving_sidecar Public archive

    Serve machine learning models using tensorflow serving

    Python 45 21

  3. gRPC-multiprocessing gRPC-multiprocessing Public

    A boilerplate to use multiprocessing for your gRPC server in your Python project

    Python 25 5

  4. telegrap telegrap Public

    Code for the @RapGeniusBot on the Telegram messaging service.

    Go 23 12

  5. cancerous_cells_scans_processing cancerous_cells_scans_processing Public archive

    Predict survival time from PET scans

    Python 9 6

  6. sidecar-pattern_tf-serving sidecar-pattern_tf-serving Public archive

    Introducing model_poller, a sidecar container for tensorflow/serving.

    Python 8 2