Roberto Aguirre Guardia
Senior Product Manager, AI · End to End Agentic AI Product Delivery · Agent Behavior, Tool Schemas & Human in the Loop · Evals & Guardrails · Prototypes with Claude Code · Consulting / Client Facing · English C1
Bogotá, Colombia · Remote (US overlap) aguirrerjg@gmail.com +57 317 675 0419 linkedin.com/in/roberto-javier-aguirre-guardia

Summary

AI product leader with 15+ years, now owning end to end delivery of GenAI and agentic AI. I define agent behavior (system prompts, tool schemas, context management, human in the loop boundaries and guardrails), run evals (task completion, tool selection, failure recovery), and set agent level metrics (cost per task, escalation, time to resolution) tied to business outcomes. I prototype and build hands on with Claude and Claude Code daily, read code (Python, TypeScript), and engage engineering on architecture, cost and latency tradeoffs. From consulting, I frame technology tradeoffs for executives up to a bank's presidency; and from AI leadership, I manage AI hype with realistic expectations. Systems Engineer with an MBA. English C1.

Experience

Head of AI · Agentic AI Product: Agent Behavior, Evals & Delivery · DigitalHubAssist LLC Oct 2024 · Present
Remote · Applied AI / SaaS (Anthropic partner program)
Consulting Manager · Client Facing Delivery & Executive Tradeoff Framing · NTT DATA Colombia Apr 2021 · Oct 2024
Bogotá, Colombia · Consultancy · Financial services (banking)
Project Manager / Delivery Lead · Client Software & Digital Products · VASS (2018·21) · GlobalHitss (2015·18) · Tecnocom (2012·15) 2012 · 2021
Bogotá, Colombia · Digital delivery

Education

Systems Engineering · U. de Lima (Top Third) · MBA, Tecnológico de Monterrey · Founding Professor of Innovation, U. Sergio Arboleda 1999 · 2004

Core Competencies

End to End Agentic AI Product Delivery (Discovery → Phased Rollout) Agent Behavior (System Prompts · Tool Schemas · Context · Working Memory) Human in the Loop Boundaries & Guardrails · Non Deterministic Systems Evals (Task Completion · Tool Selection · Failure Recovery · Safety) · Observability (Langfuse) Agent Level Metrics (Cost per Task · Escalation · Time to Resolution) + Business KPIs Prototyping with Claude · Claude Code · AI Forward · Reads Code (Python · TypeScript) Deep AI (LLMs · RAG · Agentic · Prompt/Context Engineering) · Cloud (Azure) Consulting / Client Facing · Executive Framing · Managing AI Hype · MBA