Systems & AI engineer · South Brittany

I go one level down when a problem resists.

AI architectures, CPU inference, Rust, assembly when it is needed. I write the articles I wish I had read. Soon, the courses that go with them.

latest article
October 3, 2026One level down, or several... #000
One level down: #000 an inference engine that writes its own code

I am opening a series on Herbert, an inference engine whose entire computation is written as machine code at startup, for the processor it runs on. What it is, what I am going to publish and what the series will measure.

series
ongoing1 / ∞
One level down, or several...
archive
2026
10-03One level down: #000 an inference engine that writes its own codeOne level down, or several... 06-02Friction or evaporation: two regimes of AI transformation in SMEs 05-04The model that drops digits: picking an LLM architecture per agentic link 04-10Markdown as insurance: why your agentic method must outlive the modelAI agents methodology 04-09Optimizing an AI agent's context: eco & logicalAI agents methodology 04-08One source of truth per role: the silent trap of agentic systemsAI agents methodology 04-07A methodology for building AI agentsAI agents methodology 04-04Open-source LLMs in April 2026: landscape and observations 03-15CPU Inference #0: understand before optimizingCPU Inference 02-20CPU Inference #1: The three regimesCPU Inference 02-17Qwen3 vs Qwen3.5: What Actually Changes 02-13From gpt2-experiments to qwen3-experiments-rs: The Memory Wall 02-07Tile-major layout and transformer memory allocations 02-05OPNI: when the Vibe Coder forgets the Stop button
+ 43 older articles
2026
01-30Generative AI and Software Liability: What Changes in 2026 01-20Beating PyTorch with Rust and 180 Lines of Assembly 01-15How tech is losing its craft
2025
11-16PriorityQueue - Rust #006 - async/awaitPriorityQueue in Rust 11-09PriorityQueue - Rust #005 - Graceful ShutdownPriorityQueue in Rust 11-02PriorityQueue - Rust #004 - Blocking DequeuePriorityQueue in Rust 11-02Cursor and Windsurf Run on Chinese AI Models 10-26PriorityQueue - Rust #003 - MultithreadingPriorityQueue in Rust 10-17PriorityQueue - Rust #002 - FairnessPriorityQueue in Rust 10-10PriorityQueue - Rust #001PriorityQueue in Rust 08-24Test Your Rust Knowledge 08-23Deep Dive into Safe Concurrency with Rust 08-22Rust and Miri: Beyond the Compiler for Memory Safety 08-17Real Costs of Matrix Multiplication (Unoptimized)video 08-15Virtual Memory, TLBs and NX Bit Simulation: A Journey into x86 32-bit 08-14Tenstorrent and Float Multiplication, and Why SLMs Are the Future 08-13Active Testing: Testing LLMs the Smart Way 08-10Qwen3 vs GPT-OSS: CPU Benchmark Without GPU 08-09MXFP4: Revolution or Evolution? 08-09Plan 9: An Operating System That Shaped My Career 08-08Why AI Slows Down as the Conversation Grows 08-05Business Debt: When Teams Shape Systems That Last 08-04KV Cache Optimization in LLMs 08-02Towards Energy-Efficient Conversational AI 08-01RAG: AI Without Hallucinations 07-31Kafka, Pulsar or Ray? A Comparison for Distributed AI 07-30Three-Body System: Thinking Distributed AI with Rust 07-29Why I Designed My Own LLM Orchestrator in Rust 07-28The Case for Sovereign and Frugal AI 07-25America Accelerates on AI: What About France? 07-25Why LLMs Sometimes Get Stuck in Loops
2019
06-11Go MySQL Tutorial [Part 2: Connection and Configuration]Go MySQL Tutorial 06-11Go MySQL Tutorial [Part 3: Data Transfer]Go MySQL Tutorial 06-11Go MySQL Tutorial [Part 1: Installation]Go MySQL Tutorial 06-11Go MySQL Tutorial [Part 4: Transactions]Go MySQL Tutorial 06-10Dive into Go API 02-28Cloud Act, Patriot Act, GDPR 01-15My readings: Space Opera
2018
10-21Sequence Models course certificate on Coursera 08-25Blogging with Hugo website generator 05-24'Neural Networks and Deep Learning' or why I like KISS
2015
11-25100 Years of General Relativity 04-01Plan9 Operating System