

Senior Site Reliability and Platform Engineer specializing in highly available AWS and Kubernetes environments, infrastructure automation, observability, and operational reliability. Experienced in designing and scaling cloud platforms, optimizing infrastructure costs, leading complex incident response, and building secure, multi-tenant systems in regulated environments. Strong expertise in AWS, Terraform, Kubernetes, and modern observability, with hands-on experience applying LLMs and autonomous agents to automate engineering and operational workflows.