about me
building things, one yak at a time
I build systems behind other systems. From distributed infrastructure to internal developer platforms, I enjoy designing the foundations that help software run reliably to enable others to work effectively.
hot take
The "H" in Kubernetes stands for happiness.
Career
Where I've been.
Senior Platform Engineer
Building the platforms and internal tooling that let engineering teams ship reliable, scalable systems across the company.
Senior Data Reliability Engineer
Responsible for ensuring reliability and automation of stateful workloads in Kubernetes. Provided abstractions and automation for different product teams through Kubernetes operators, REST APIs, and integration services designed to accelerate adoption and migration from EC2 to Kubernetes (EKS).
Day-to-day work included Golang, Python, AWS, FluxCD, and Kubernetes controllers/operators to automate provisioning and customization of Kafka, RabbitMQ, Zookeeper, Elasticsearch, Spark, and Volcano for scheduling compute-intensive workloads.
Principal Engineer · Platform Core
Responsible for enhancing user experience and productivity of Platform and Software Engineering teams by assuring reliability and constant evolution of the core tooling used to support the Software Development Life Cycle and Infrastructure teams. Owned Configuration Management, Artifact Registry, Service Catalog and Service Discovery, Change Control, Secret Management, and the Internal Developer Platform (IDP).
Day-to-day stack: Golang, Python, Groovy, Jenkins, Azure Kubernetes Service, Artifactory, Consul, HashiCorp Vault, GitLab (on-prem), SaltStack, Terraform, and Kubernetes controllers/operators to automate installation and configuration of platform components.
Principal Engineer · Delivery Platform
Responsible for the CI/CD platform used by over 800 Farfetch developers to accelerate the Software Development Life Cycle. Assigned as principal engineer along with a team of seven engineers. Acted mostly on consolidation and modernization of the CI/CD stack, integrating different systems and architectures to promote best practices and help teams deliver products and features in a performant, reliable, and repeatable way.
Owned CI/CD, base container images, cluster auto-scaling, deployment of stateful/stateless workloads, and pre-install tasks such as DB schema migrations and CDN content publishing. Stack: ArgoCD, Argo Workflows, Argo Events, Jenkins, GitLab, Terraform, Golang, Python, Azure Kubernetes Service.
Senior Engineer · Developer Productivity
Assigned to the Developer Productivity team, focused on continuous integration and improving the developer experience with the right tooling. Responsible for maintaining flexibility and autonomy of over 3000 services and libraries to deliver value in a fast-paced environment, using scalable and tech-agnostic pipelines.
Day-to-day: Jenkins (CI/CD), Groovy shared libraries and declarative pipelines, Golang, Python, Prometheus exporters, API integrations, and automation. Also maintained the CI/CD Kubernetes clusters on Azure Kubernetes Service, including provisioning, patching, and customization.
Senior Engineer · Observability Platform
Part of the Observability platform team, responsible for supporting, scaling, and increasing reliability of the platform for analysis, detection, and investigation. Worked with open-source tools like Prometheus, Thanos, Grafana, Elasticsearch, and Kubernetes to help build a robust, multi-region platform that was both reliable and self-service, empowering users to understand their own systems.
Application Developer · M&EaaS
Acted as Application Developer on the M&EaaS (Monitoring & Event Automation as a Service) team. Responsible for creation and maintenance of an API synchronization micro-service written in Go to extract information from different monitoring tools and consolidate data from multiple sources into a single API.
Responsible for CI/CD tools such as Jenkins and UrbanCode Deploy, plus build, test, and deployment pipelines. Developed and supported all Helm Charts and Dockerfiles used for packaging and deploying applications. Created tools and scripts to increase team productivity, including CLI utilities for interacting with and load-testing different APIs, and Ansible playbooks for automating service deployment. Stack: Jenkins, UrbanCode Deploy, Kubernetes, OpenShift, Ansible, Helm, Golang.
Site Reliability Engineer · M&EaaS
Part of the M&EaaS (Monitoring and Event Automation as a Service) team, responsible for reliability of the solution that provided integrated monitoring services for multiple IBM customers across Europe and different IBM infrastructure teams across North America.
Responsible for deployment and reliability, supporting both Development and Ops teams by automating and developing better ways of working, while onboarding new customers. Stack: Linux, Kubernetes, Docker, Jenkins, UrbanCode Deploy, Go, Bash, Python, Ansible, Ansible Tower, Chef, Prometheus, Tivoli Monitoring, IPM8.
Lead System Engineer · Michelin North America
Responsible for developing new monitors and automation to support Michelin North America monitoring infrastructure, guaranteeing quality and availability using ITM6 and BlueCARE. Acted as technical lead of the level-2 support team (4 people). Developed new monitoring agents and provided tooling and automation to support Michelin North America's infrastructure.
Subject Matter Expert · Global Enterprise Automation
Subject Matter Expert for the Enterprise Automation team. Responsible for automating and supporting all monitoring infrastructure for Global offshore accounts using ITM5, ITM6, and BlueCARE. Supported accounts such as Whirlpool, Honda Motor Company, GAP, IBM Canada, Michelin North America, and Morgan Stanley.
IT Specialist · e-Business Hosting
Assigned to the Enterprise Automation team at IBM Brazil as a technical specialist on IT infrastructure monitoring and automation solutions. Provided support for the largest BMC Patrol environment at IBM with over 4000 monitoring agents integrated with different Tivoli monitoring tools used for observability of applications and multiple data centers' infrastructure.
Used Tivoli Configuration Management, Tivoli Distributed Monitoring, and ITM 6.2 solutions for a pool of over 100 offshore accounts across banking, health care, oil and gas, automotive, energy, retail, news media, and public sector. Responsible for development and delivery of new solutions focused on automation of manual tasks for cost reduction and revenue assurance.
Senior Consultant · Vale S.A.
Specialist in IT infrastructure solutions, part of the Technology & Solutions team. Responsible for helping with pre-sales and implementations of IT infrastructure projects at BearingPoint Brazil.
Acted as part of the Knowledge Management team at VALE, managing new demands and generating Key Performance Indicators for mining assets fuel supply.
Subject Matter Expert · Carrefour Brazil
Specialist in IT infrastructure solutions, part of the Brazilian Automation service line at IBM Brazil. Assigned as Subject Matter Expert on non-IBM solutions at Carrefour Brazil, part of the outsourcing transition project where most of Carrefour's infrastructure was migrated to IBM data centers.
Responsible for development and delivery of new solutions focused on automating manual tasks, reducing operational costs, and integrating customer infrastructure with IBM's internal systems. Also supported the sales team with proof-of-concepts and trials for proposed solutions.
Infrastructure Specialist · Automation & Tools
IT specialist focused on supporting EDS clients using CA Unicenter TNG / NSM, BMC PATROL, and BMC PATROL Performance. Part of the Automation and Tools offshore group supporting US accounts such as NASD and Nextel.
IT Consultant
IT consultant focused on Infrastructure Management and Service Desk solutions using BMC Software and Computer Associates solutions in large and medium accounts such as Petrobras, Aracruz Celulose, Light, Oi (Telemar), Blah! (TIM), Andima, Unimed, White Martins, and SITA.
Junior Software Engineer
Junior software engineer working on infrastructure management and service desk solutions.
Technical Support Specialist · InfoGlobo
Part of the service desk team, providing hardware and software support for Windows workstations and Linux servers.
Contact
Let's connect.
If you are interested in platforms, infrastructure, developer experience, self-hosting, or building useful systems, I would be happy to talk.