Technique

Agent Evaluation

Agent Evaluation involves methods for measuring the effectiveness of AI agents and their configurations. You’ll learn about self-training setups like Claude Code, explore benchmarks that test agent performance, and discover fast-track learning strategies if you’re starting with AI agents. Videos typically share real-world applications and insights into the best practices for evaluating and improving agent systems.

Stop watching · start building
Matched to Agent Evaluation

Makers residency 4th edition

Your idea can be anything. At our last workshops, people built plant-based nutrition businesses, a way to measure planetary resilience, crochet headwear and a way to connect lonely people living in the same building. An asset manager went on to raise an additional £25M on their fund within nine months. The fourth workshop is a full day at KOKO with Nick Sarafa, learning to build with Claude Code. Bring the thing you keep meaning to start. We’ll bring the pizza.

◆ Fri 13 Nov 2026 ◆ KOKO, London
Makers residency 4th edition
Live event
Makers residency 4th edition
Fri 13 Nov 2026
People watching Agent Evaluation also follow
The whole library, sorted

Browse by
topic.

CHOOSETO Studio

What would it take to build this?

Let us build it for you. Design and engineering from the people who shipped platforms to billions of users. AI-native, live in weeks, and yours outright at the end.