Robert Teah

DevOps Engineer

Cloud Engineer

Infra Engineer

Site Reliability Engineer

DevSecOps Engineer

Robert Teah

DevOps Engineer

Cloud Engineer

Infra Engineer

Site Reliability Engineer

DevSecOps Engineer

Senior DevOps / Platform / SRE Engineer with 10+ years of experience operating production infrastructure at enterprise scale across AWS and Kubernetes environments.

Robert Teah — Senior DevOps / Platform / SRE Engineer

Senior DevOps / Platform / SRE Engineer with 10+ years of experience operating production infrastructure at enterprise scale across AWS and Kubernetes environments. Built and ran mission-critical e-commerce systems at Luxottica (99.99% uptime) and multi-region SRE infrastructure at Home Depot (35% faster incident detection, 15% cost reduction). Now extending that foundation into AI infrastructure through AutoSRE, Reliai, and WebIntell.

  • Residence: Atlanta, GA
  • Status: Open to new opportunities
  • LinkedIn: linkedin.com/in/robertteah
What I Do
CI/CD & Release Engineering

GitHub Actions and CodePipeline workflows with Trivy scanning, Helm promotion, and blue/green rollouts across EKS.

Observability & SRE

SLIs/SLOs, Prometheus/Grafana pipelines, and blameless postmortems that turn reactive alerting into signal-driven ops.

Infrastructure as Code

Modular Terraform with S3 remote state and DynamoDB locking, cutting environment provisioning from hours to minutes.

AI Reliability Engineering

LLM observability, autonomous remediation, and evaluation guardrails for production-operable AI systems.

Fun Facts
100+ FIFA Won.
AI StartUp Founder
1 000+ Cans Of Monster
100+ Anime Watched
Resume
Experience
Oct 2021 – Apr 2025
AWS DevOps Engineer
Luxottica North America

Built end-to-end CI/CD (GitHub Actions, CodePipeline, Trivy, Helm) across EKS, cutting release cycle time by 50%. Standardized Terraform provisioning with S3 remote state and DynamoDB locking; introduced blue/green and canary rollouts for zero-downtime releases during high-volume retail events.

Jul 2019 – Oct 2021
AWS Site Reliability Engineer
Home Depot

Codified multi-region AWS infrastructure (EC2, RDS Multi-AZ, EKS) maintaining 99.99% uptime. Implemented SLIs/SLOs with Prometheus and Grafana, improving incident detection speed by 35% and cutting AWS costs 15% through right-sizing and autoscaling tuning.

Jun 2013 – Jul 2019
AWS Cloud Engineer
FutureMedia

Designed fault-tolerant AWS architectures for large-scale media streaming with defined RPO/RTO targets. Built Terraform automation migrating legacy on-prem workloads to AWS, and optimized streaming platform costs by 20% over two years.

Education
Clayton State University
B.S. in Information Technology

Bachelor of Science in Information Technology.

AWS Certified Solutions Architect – Associate
Amazon Web Services

Professional certification in AWS solutions architecture.

Udacity Nanodegrees
Cloud DevOps Engineer · AWS Architect · Site Reliability Engineering

Three nanodegree programs covering DevOps, AWS architecture, and SRE practice.

Skills
Coding
  • Python
  • Bash
  • YAML
  • FastAPI
CI/CD & IaC
  • Terraform
  • GitHub Actions
  • Docker / Helm
  • Jenkins / GitLab CI
Cloud & Platform
  • AWS (EKS, RDS, VPC, IAM)
    90%
  • Kubernetes
    90%
  • Linux
    85%
  • Prometheus / Grafana
    85%
AI Infrastructure & LLMOps
  • LLM Observability
  • Autonomous Remediation Workflows
  • RAG for Operational Knowledge Bases
  • Agentic AI Workflows
Works
AutoSRE
Content
Remotastic
Content
Locadize
Content
Reliai
Content
WebIntell
Content
Blog
September 27, 2026 What Broke First at Scale: The Synchronous Bottleneck Behind WebIntell

Load-testing the WebIntell ingestion pipeline exposed a problem that hadn’t shown up anywhere in normal operation: under 50 requests per…

September 27, 2026 The 503 That Route53 Didn’t Cause

At Luxottica, during a promotional campaign load spike, about 30% of product catalog requests started returning 503s — not all…

September 27, 2026 Why Your Readiness Probe Is Lying to You

At Luxottica, a new version of the Java Ad Service deployed to production EKS. The pipeline reported success, all six…

September 27, 2026 A Peak-Traffic RDS Failover Outage, and Why Disaster Recovery Isn’t a Configuration

At Home Depot, a Prometheus alert fired on elevated write-error rates across the checkout service during a peak shopping window…

Get in Touch
  • Address: Atlanta, GA
  • Email: robertteah@gmail.com
  • Phone: (404) 786-8146
  • LinkedIn: linkedin.com/in/robertteah
Contact Form