Site Reliability Engineer
Current- Manage cloud services in AWS, GCP, Azure, and Heroku to ensure optimal performance and scalability.- Automate operational and deployment processes using various tools and scripts for efficiency (Jenkins, BitbucketCI, Github Action, ArgoCD).- Investigated and resolved technical issues, providing root cause analysis to improve system reliability. I also take part of on-call rotation.- Operate Kubernetes and Docker servers in production and development environments for seamless deployment- Actively find new way to reduce infrastructure cost and improve performance, efficency, and reliability of the cloud infrastructure.- Creatively create automations to eliminate operation toils. Usually using bash script, python, or with workflow automation tools.