I am a Site Reliability and DevOps Engineer with more than eight years of hands-on experience building reliable, scalable, and automated cloud platforms. I specialize in AWS, Kubernetes, infrastructure as code, CI/CD, GitOps, observability, and production operations.
I design and operate highly available cloud infrastructure using AWS services including EKS, VPC, IAM, EC2, RDS, S3, load balancers, and Auto Scaling. I focus on creating secure, efficient platforms that enable teams to deploy and operate workloads confidently.
I build repeatable infrastructure and configuration workflows with Terraform, Ansible, Helm, and Kustomize. I have modernized delivery processes through GitHub Actions, GitLab CI/CD, Jenkins, Argo CD, and GitOps practices, improving release consistency and rollback capabilities.
I am experienced in reliability engineering, including SLI/SLO management, incident response, root cause analysis, capacity planning, disaster recovery, and MTTR/MTTD improvement. I implement comprehensive observability using Prometheus, Grafana, OpenTelemetry, Alertmanager, CloudWatch, and ELK.
I also apply DevSecOps principles throughout delivery and infrastructure workflows. My experience includes container security, vulnerability scanning with Trivy, code-quality controls with SonarQube, secrets management with HashiCorp Vault, cloud security controls, and AWS cost optimization.
I enjoy simplifying complex operational problems, improving engineering workflows, and building self-service platform capabilities. I bring strong Linux administration, scripting, troubleshooting, and cross-functional platform engineering expertise to production environments.