Rogerio Chaves

Rogerio Chaves

Co-founder & CTO

Spent years at Booking.com, including helping build its first AI agents and seeing firsthand how fragile, opaque, and hard to test these systems become at scale. Engineering background at ThoughtWorks before that; full-stack across React, TypeScript, GraphQL, and ML.

29 posts

July 22, 2026Product Releases

Introducing Langy: Your Automated AI Engineer

July 8, 2026Article

Claude vs Codex: which is the better background agent?

July 5, 2026Developer findings

Background Agents on Slack: How we built our own Claude Tag before it was cool

July 5, 2026Governance AI

EU AI Act compliance: are you affected?

April 14, 2026Governance AI

Why AI Red teaming is broken (and how we fixed it)

March 25, 2026Governance AI

A Note on the LiteLLM Vulnerability

February 3, 2026Integrations

How OpenClaw / ClawBot works behind the scenes - and why agent observability matter

February 3, 2026Integrations

Instrumenting Your OpenClaw Agent with LangWatch via OpenTelemetry

February 3, 2026Integrations

How to Use Clawdbot + LangWatch to Monitor Your Agents in Production

February 2, 2026LLM Evals

LLM Evaluations Explained: Experiments, Online Evaluations, Guardrails, and when to use each in 2026

November 26, 2025Product Releases

Launch Week Day 5: Better Agents CLI: The reliability layer for the next wave of agent development

October 6, 2025Integrations

The Ultimate RAG Blueprint: Everything you need to know about RAG in 2025/2026

September 7, 2025LLM Evals

Are evals dead?

September 3, 2025LLM Evals

Essential LLM evaluation metrics for AI quality control: From error analysis to binary checks

August 7, 2025LLM Evals

LLM-as-a-Judge: Using the Panel of Judges Approach to Approximate Human Preference

June 27, 2025Agents

Real-time simulation visualization and debug mode

June 26, 2025Agents

Scripted simulations, evaluations, and guardrails

June 25, 2025Integrations

Test agents on Mastra, Agno, and 10+ other frameworks

June 24, 2025Product Releases

Introducing simulation-based agent testing

June 24, 2025Agents

Why LangWatch Scenarios represents the future of AI agent testing

June 21, 2025Agents

Best AI Agent Frameworks in 2025: Comparing LangGraph, DSPy, CrewAI, Agno, and More

April 22, 2025Product Releases

Introducing the Evaluations Wizard: How to evaluate your LLM: Building an LLM evaluation framework that actually works

April 8, 2025Product Releases

Introducing Scenario: Use an Agent to Test Your Agent

January 1, 2025Agents

7 Predictions for AI in 2025: A CTO's, Rogerio Chaves Perspective

December 10, 2024Product Releases

LangWatch Optimization Studio - Built for AI Engineers, by AI Engineers

July 3, 2024Agents

The complete guide for TDD with LLMs

June 27, 2024LLM Evals

Data Flywheel: Using your production data to build better LLM products

June 10, 2024LLM Evals

Unit Testing Your LLM: The Power of Datasets

June 3, 2024Product Releases

Introducing DSPy Visualizer