lyuata.
Field notes
AI product & agent trust advisory · Europe

Your AI is shipping.
Is it working?

Everyone has an opinion. Nobody has evidence. In two weeks I get you a straight answer: fixed scope, fixed price.

Fit check first. If I'm not the right person, I'll tell you on the call.
evidence, not vibes
You see whether it works on real traffic, not in the demo.
an outcome, not hours
Scope and price agreed up front. No meters running.
a senior operator
15+ years in product, Experian & Progress. You work with me directly.
15+ years in product, including
ExperianProgress

At Experian, product manager for ML in high-risk credit decisioning systems (ModelOps)At Progress, product lead of the AI agent observability platform (GA 2026)#6 Product of the Day on Product HuntOpen-source agent-skills toolkit on GitHubWebinar host to developer audiences

When leadership asks “is our AI working?”
The usual answer
“We think so? The demo went really well…”
— a vibe, not an answer
After two weeks with me
“Yes. Here's the proof. And here's what we fixed.”
— evidence, in plain language
I turn the first answer into the second.
Fixed scope · fixed price · you work with me directly

Offers

Offer A

Agent Trust Sprint

For teams with an AI agent or LLM feature in production who can't confidently answer “is it working?”

  • Audit of architecture, prompts and tool calls
  • Failure analysis on real traffic
  • 3-5 evaluations set up as your regression safety net
  • Written findings + 90-minute team session
2 weeks, ~4 working days
€2,400 fixed
Pilot pricing available for the first two clients
Offer B

AI Adoption Sprint

For companies of 10-100 people that know AI should be saving them time but have no plan - or a shadow-ChatGPT problem.

  • Map the 3 highest-ROI uses for your actual workflows
  • Tool selection and setup, compliant options included
  • One-page AI usage policy + team training
  • Follow-up call in week 3
2 days + follow-up
€1,800 fixed
Pilot pricing available for the first two clients
How it works

30-minute call fixed quote start within 2 weeks

Who you'll work with

Hi, I'm Lyubo. By day I'm the product lead of Progress AI Observability, a platform for making AI agents trustworthy in production. On the side I'm the founder of Babuger, an AI SDR platform that runs outreach end-to-end. Before that I was a product manager at Experian, running ML in high-risk credit decisioning systems (ModelOps), where a model you can't explain or monitor doesn't ship. Which means I sit on every side of your problem: I've run regulated ML, I build the tooling that catches AI failures, and I run AI agents against real customers with my own money on the line.

You work with me directly. No handoff, no juniors, no account manager.


Proof of work

I work on this problem in public. The essays and free tools below are the same thinking you'd be buying, so judge it before you book.

Essay · No. 2 · Red Flags

Your AI Agent Has the Same Red Flags as Your Ex

A field guide to the five ways agents misbehave in production, and how a trace turns each one from a mystery into a span you can see.

Read the field notes →
Essay · No. 1 · The Agent Observability Gap

The Agent Observability Gap

Long-form explainer on why agents fail in production and what a useful stack actually covers.

Read the explainer →
Browse all field notes →
Free tool · LLM evals
Judgekit

Free generator for LLM-as-Judge evaluator prompts. Paste a trace, get a research-grounded prompt with a built-in stress test.

judgekit.lyuata.com
Free tool · MCP audit
MCP Trap

Free analyzer for MCP servers and tool-call schemas. Static + LLM-assisted findings cited to primary research, plus a portable test pack.

mcptrap.lyuata.com
Not ready yet?
Field notes

Agent trust, in your inbox

Occasional field notes on making AI agents provably work in production. No pitch, unsubscribe anytime.

Double opt-in via Buttondown.

FAQ

Why fixed price?

Because you're buying an outcome, not my hours.

Remote?

Yes. On-site available in Europe.

How soon can we start?

Two openings per month. After the fit call you get a fixed quote, and we start within 2 weeks.

Who is this not for?

Teams wanting a body-shop developer, or enterprises that need a vendor with a procurement portal.

Next step

If the fit isn't there, you'll know in 30 minutes.