About Me
Solutions Architect and Certified Expert Site Reliability Engineer with 20+ years spanning both sides of the table: customer-facing technical consulting and hands-on ownership of the platforms underneath.
Seven of those years were spent in customer-facing technical roles. As Lead Solutions Architect at IBM I ran on-site requirement sessions with CIOs and technical stakeholders, designed hybrid and multi-cloud target architectures, presented and demoed them to executive and engineering audiences, and oversaw delivery through implementation and handoff. At VMware I owned VMware Lab Platform end to end — qualifying leads, running discovery calls, delivering live product demos, and guiding customers to the right subscription — while running the vSphere and NSX infrastructure the platform depended on.
The other side is production operations: reliability engineering, SLOs and error budgets, observability, Infrastructure as Code, incident response, and cloud cost optimization across AWS, Azure, GCP, IBM Cloud, Kubernetes, and OpenShift.
I built and teach the IBM SRE Academy, along with Kubernetes/OpenShift and FinOps bootcamps, and mentor engineers moving into reliability roles. Recognized as an IBM Profession Champion for contributions to the SRE profession through mentoring, speaking, and technical education. CKA certified and Open Group Certified Solution Architect. Native Spanish, fluent English.
Featured Work
SRE Academy at IBM
Built IBM's SRE training program — teaching reliability as a mindset, not a checklist. Structured curriculum covering CI/CD, observability, incident response, and chaos engineering through hands-on Kubernetes labs.
View the curriculumMentorship at Scale
Mentored engineers across continents into reliability roles — one systems thinker at a time. Focused on Kubernetes, observability, and teaching engineers to reason from reliability first principles.
My mentorship philosophyEmerging SRE
Agentic AI systems, production-ready multi-agent architectures on Kubernetes, and the next generation of distributed resilience — including Kagent SRE Companion, a working demo platform for autonomous cluster operations.
Explore the projectProjects
Kagent SRE Companion
AI-powered SRE companion platform with intelligent blue/green deployments, autonomous failover controllers, and conversational cluster operations.
SRE Academy
Hands-on SRE training series teaching reliability engineering through practical Kubernetes labs. Covers monitoring, automation, CI/CD, GitOps, rollbacks, and chaos engineering.
GCP SRE Hands-On Training
A structured professional development program that takes developers to production-ready SRE competency through systematic, hands-on exercises on Google Cloud.
CKA Study Guide
Preparation guide for the Certified Kubernetes Administrator exam, with hands-on examples, simulators, and strategies drawn from real-world operations.
AWS Instance Scheduler
Terraform automation for EC2 instance uptime scheduling, cutting operational spend through automated start/stop windows and resource optimization.
AWS Terraform + Packer
End-to-end AWS environment provisioning with Terraform and Packer, covering VPC, subnets, security groups, and an Nginx application behind a load balancer.