Spent years at Booking.com, including helping build its first AI agents and seeing firsthand how fragile, opaque, and hard to test these systems become at scale. Engineering background at ThoughtWorks before that; full-stack across React, TypeScript, GraphQL, and ML.
29 posts
July 5, 2026Developer findings
Background Agents on Slack: How we built our own Claude Tag before it was cool
February 3, 2026Integrations
How OpenClaw / ClawBot works behind the scenes - and why agent observability matter
February 2, 2026LLM Evals
LLM Evaluations Explained: Experiments, Online Evaluations, Guardrails, and when to use each in 2026
November 26, 2025Product Releases
Launch Week Day 5: Better Agents CLI: The reliability layer for the next wave of agent development
October 6, 2025Integrations
The Ultimate RAG Blueprint: Everything you need to know about RAG in 2025/2026
September 3, 2025LLM Evals
Essential LLM evaluation metrics for AI quality control: From error analysis to binary checks
August 7, 2025LLM Evals
LLM-as-a-Judge: Using the Panel of Judges Approach to Approximate Human Preference
June 21, 2025Agents
Best AI Agent Frameworks in 2025: Comparing LangGraph, DSPy, CrewAI, Agno, and More
April 22, 2025Product Releases


