Systems & AI engineer · South Brittany
I go one level down when a problem resists.
AI architectures, CPU inference, Rust, assembly when it is needed. I write the articles I wish I had read. Soon, the courses that go with them.
latest article

One level down: #000 an inference engine that writes its own code
I am opening a series on Herbert, an inference engine whose entire computation is written as machine code at startup, for the processor it runs on. What it is, what I am going to publish and what the series will measure.
series
One level down, or several...
AI agents methodology
CPU Inference
archive
2026
10-03One level down: #000 an inference engine that writes its own codeOne level down, or several...
06-02Friction or evaporation: two regimes of AI transformation in SMEs
05-04The model that drops digits: picking an LLM architecture per agentic link
04-10Markdown as insurance: why your agentic method must outlive the modelAI agents methodology
04-09Optimizing an AI agent's context: eco & logicalAI agents methodology
04-08One source of truth per role: the silent trap of agentic systemsAI agents methodology
04-07A methodology for building AI agentsAI agents methodology
04-04Open-source LLMs in April 2026: landscape and observations
03-15CPU Inference #0: understand before optimizingCPU Inference
02-20CPU Inference #1: The three regimesCPU Inference
02-17Qwen3 vs Qwen3.5: What Actually Changes
02-13From gpt2-experiments to qwen3-experiments-rs: The Memory Wall
02-07Tile-major layout and transformer memory allocations
02-05OPNI: when the Vibe Coder forgets the Stop button
+ 43 older articles
2026
01-30Generative AI and Software Liability: What Changes in 2026
01-20Beating PyTorch with Rust and 180 Lines of Assembly
01-15How tech is losing its craft
2025
11-16PriorityQueue - Rust #006 - async/awaitPriorityQueue in Rust
11-09PriorityQueue - Rust #005 - Graceful ShutdownPriorityQueue in Rust
11-02PriorityQueue - Rust #004 - Blocking DequeuePriorityQueue in Rust
11-02Cursor and Windsurf Run on Chinese AI Models
10-26PriorityQueue - Rust #003 - MultithreadingPriorityQueue in Rust
10-17PriorityQueue - Rust #002 - FairnessPriorityQueue in Rust
10-10PriorityQueue - Rust #001PriorityQueue in Rust
08-24Test Your Rust Knowledge
08-23Deep Dive into Safe Concurrency with Rust
08-22Rust and Miri: Beyond the Compiler for Memory Safety
08-17Real Costs of Matrix Multiplication (Unoptimized)video
08-15Virtual Memory, TLBs and NX Bit Simulation: A Journey into x86 32-bit
08-14Tenstorrent and Float Multiplication, and Why SLMs Are the Future
08-13Active Testing: Testing LLMs the Smart Way
08-10Qwen3 vs GPT-OSS: CPU Benchmark Without GPU
08-09MXFP4: Revolution or Evolution?
08-09Plan 9: An Operating System That Shaped My Career
08-08Why AI Slows Down as the Conversation Grows
08-05Business Debt: When Teams Shape Systems That Last
08-04KV Cache Optimization in LLMs
08-02Towards Energy-Efficient Conversational AI
08-01RAG: AI Without Hallucinations
07-31Kafka, Pulsar or Ray? A Comparison for Distributed AI
07-30Three-Body System: Thinking Distributed AI with Rust
07-29Why I Designed My Own LLM Orchestrator in Rust
07-28The Case for Sovereign and Frugal AI
07-25America Accelerates on AI: What About France?
07-25Why LLMs Sometimes Get Stuck in Loops
2019
06-11Go MySQL Tutorial [Part 2: Connection and Configuration]Go MySQL Tutorial
06-11Go MySQL Tutorial [Part 3: Data Transfer]Go MySQL Tutorial
06-11Go MySQL Tutorial [Part 1: Installation]Go MySQL Tutorial
06-11Go MySQL Tutorial [Part 4: Transactions]Go MySQL Tutorial
06-10Dive into Go API
02-28Cloud Act, Patriot Act, GDPR
01-15My readings: Space Opera
2018
10-21Sequence Models course certificate on Coursera
08-25Blogging with Hugo website generator
05-24'Neural Networks and Deep Learning' or why I like KISS
2015
11-25100 Years of General Relativity
04-01Plan9 Operating System