KhaifParis · Full remoteENFR

Si-ThamiKhaif

GenAI Engineer / LLM Systems / MLOps

I deploy LLM agents on your own infrastructure.Data sovereignty without the efficiency hit.

01about.me

  1. Auto-mate AgencyFounder and GenAI Engineer

    2023 to today · full-time since January 2026 · Paris, remote

    • Generative AI systems for clients, from scoping through to handover to their own teams.
    • RAG assistants over client document bases; automated extraction and summarisation of contracts, invoices and reports.
    • Advisory on safe adoption and on-premise LLM deployment, for teams whose data cannot leave the network.
    • Typical result: 40 hours of manual work removed per month.

    Python · LangChain · LangGraph · OpenAI · Mistral · Anthropic · vector databases

  2. Air Liquide R&D JapanMachine Learning Engineer, R&D

    June to December 2025 · Tokyo

    • Design, build and evaluation of a production RAG system, with hallucination rates measured and LLM-as-a-judge validation before go-live.
    • Rebuilt a MedTech facial measurement pipeline, from 2D photo to real-time 3D video, 468 landmarks per frame, 95% accuracy.
    • Owned the data foundation of a global ML project: scraping and preprocessing pipelines under CI/CD.

    LangChain · FAISS · Pinecone · LLM-as-a-judge · Python · MediaPipe · GitHub Actions

  3. Safran AerosystemsSoftware AI Developer, internship

    June to September 2024 · Plaisir

    • Fully on-premise RAG assistant deployed on the company SharePoint. 250+ active users, around two hours saved per person per week.
    • Delivery tracking dashboard and delay forecasting models for the supply chain team.

    LangChain · Ollama · SharePoint · Power BI · Power Query · VBA

  4. Safran Aircraft EnginesSoftware and Data Automation, internship

    June to September 2022

    • Python automation and a delivery tracking tool shared across teams.

    Python · Pandas · Scikit-learn

02my-projects

  1. 01Local vs cloud RAG (benchmark)

    code

    Question answering over a document corpus: every claim cites the passage backing it, and the system declines to answer rather than invent one. Nothing leaves the machine by default; a hosted backend puts a number on the local/cloud gap.

    Python · Ollama · Azure OpenAI · FastAPI · NumPy · rank-bm25 · sentence-transformers · Docker

    A question put to the system, then the answer writing itself out with numbered citations pointing back to the source passages.
  2. 02LLM credit scoring

    code

    Business loan applications arrive as free text. A local LLM extracts the figures, then plain Python scores them on an eight-criterion grid — the model never does the arithmetic. 97.3% field accuracy, no changed recommendation over 90 runs.

    Python · Ollama · Llama 3.1 8B · Streamlit · pytest

    A loan application typed as free text, then the score it returns — 42 out of 100, risk level and completeness — then the extracted figures laid out as editable fields, and the eight weighted criteria behind the result.
  3. 03Cyber AI NER

    code

    Named-entity extraction from English threat reports: 4,876 annotated sentences, a three-type BIO scheme. Two approaches compared, a bidirectional LSTM built from scratch against a fine-tuned pre-trained transformer. Done as a pair.

    Python · TensorFlow · Keras · simpletransformers · BERT · scikit-learn · Parquet

    Model output: cybersecurity entities picked out of a threat report.

03skills.md

LLM and GenAI

  • LangChain
  • LangGraph
  • LlamaIndex
  • MCP
  • OpenAI, Mistral and Anthropic APIs
  • Ollama
  • RAG
  • Multi-agent systems

ML and Data

  • PyTorch
  • TensorFlow
  • Scikit-learn
  • Pandas
  • NumPy
  • NLP
  • Computer vision
  • MediaPipe, OpenCV, YOLO

MLOps and evaluation

  • Docker
  • CI/CD, GitHub Actions
  • AWS
  • RAGAS
  • LangSmith
  • Langfuse
  • LLM-as-a-judge
  • Model versioning

Languages

  • Python
  • C++
  • JavaScript and TypeScript
  • Swift, SwiftUI
  • SQL

04contact.me

Scoping one LLM use case on your own data.

Deployed on your infrastructure, with nothing leaving your network.