Cybiqon Lab
AboutRSScybiqon.in ↗

Notes from the workshop

What we are building, what broke, and what the numbers said. Written by hand, published when there is something worth saying — not on a schedule.

tagged Observability·show all

date
2026-08-12
words
4 749
read
24 min
sources
7
scenarios
41
runs
123
suite cost
$3.06
undercount
4.5x
untested tools
6/17
views
16
date
2026-08-12
·
words
4 749
·
read
24 min
·
sources
7 sources
·
scenarios
41 scenarios
·
runs
123 runs
·
suite cost
$3.06 suite cost
·
undercount
4.5x undercount
·
untested tools
6/17 untested tools
·
views
16 views

Six of our agent's seventeen tools had never run.

Six of seventeen agent tools had never once run in production, including both of the ones that unlock a contact and charge for it. This is the harness that finally tested them — a real model in a completely faked world, 41 scenarios, 123 runs, $3.06 — and the cost blind spot it uncovered on the way.

AI Agents · Evaluation · LLM

New posts by email

A few emails a year, when there is a new post. Nothing else, ever — and one click to leave.

Cybiqon LabWho writes thisRSSMSME blogcybiqon.in