AI Product Manager · Barcelona

I take LLM systems from demo to production — and prove they work.

8+ years product across B2B and B2C SaaS. Now owning a multi-agent platform with RAG, evals and human-in-the-loop in a regulated, high-liability domain.

8+
years in product
~30%
messages AI-drafted
1k+
users shipped solo
4
case studies
Selected work
All work →
How it is measured
Golden set · n=420
Retrieval recall0.91
Relevancy (judge)0.86
Edit distance (inv.)0.74

Normalized edit distance runs on every interaction — the passive quality signal that needs no reviewer.

Normalized edit distance is a passive quality signal on every single interaction — no reviewer required.
On designing the eval loop
Stack
Pydantic AILangfusepgvectorCohere RerankDeepEval
Track record
2026–
Studio Firn
Solo consumer products
2024–26
Senior PM — Settly
AI agent platform, immigration knowledge work
2017–23
Rentman
Specialist → Senior PO, B2B ERP