I’m a Site Reliability Engineer with 10+ years of experience across automation, reliability, security, observability, and cloud infrastructure. I focus on making systems resilient, fast, and easy to operate.

After getting my start at a mid-size company, I joined a startup as its first DevOps Engineer and built its infrastructure and CI/CD pipelines from scratch. Since then, I’ve worked on larger platforms, with experience spanning hands-on operations, reliability, and developer tooling.

Today I’m at Udemy, working on reliability at scale, developer tooling, and agentic systems.

What I do

I build reliable, observable platforms that teams can ship on with confidence. That includes making deployments safer, reducing incident risk, and automating the repetitive stuff so engineers can move faster.

01 / Reliability
SLOs, incident response, capacity planning
02 / Platform
Kubernetes, Helm, Skaffold, Terraform, Istio, Argo CD
03 / Observability
Datadog, Prometheus

Certifications

  • Certified Kubernetes Administrator CKA
  • AWS Certified Developer Associate
  • AWS Certified Solutions Architect Associate
  • Red Hat Certified Systems Administrator RHCSA
  • Certified Meraki Network Operator

Away from the keyboard

I love learning new things. Outside of work I’m into coding side projects, video games, sports, and spending time with my wife, two kids, and two French bulldogs.

Favorite shows: Mr. Robot, Silicon Valley, and Community. You can follow me on Twitter.

# start a conversation

Have something in mind?

Say hello