Currently working as a Senior Site Reliability Engineer in the Platform SRE team at Shopify 🛍️, serving as an Incident Manager On Call (IMOC) and driving end-to-end resolution of high-severity production incidents across a globally distributed commerce platform. My work spans real-time incident triage, root cause analysis, WAF/DDoS mitigation, database performance optimization, and building AI-powered incident response tooling. I focus on not just resolving incidents but driving permanent fixes, improving operational playbooks, and reducing the cost of future incidents.
Throughout my career, I've focused on SRE and operational excellence with 9+ years of industrial experience, covering cloud-native areas such as Kubernetes, Go, and software architecture.
Also, I am a tech speaker who's given 60+ conference/event presentations about SRE, CI/CD, Pattern language, AWS, Terraform, TDD, etc. Additionally, I contributed to 2 technical magazines, covering Docker and unit test best practices.
I spent many years in the e-commerce and financial domains before my current role. My first engineering job was in cloud infrastructure with AWS. Then, as a backend engineer, I joined the payment team at BASE, one of the biggest e-commerce platforms in Japan. After that, I was transferred to a subsidiary company that offered financial services at an early stage as its first engineer. I was promoted to Tech Lead and Engineering Manager of 10 team members. After that, I worked as a Senior Site Reliability Engineer at Autify, a Series B startup headquartered in San Francisco, which serves AI-powered quality assurance platforms. I was mainly responsible for developing and operating complicated infrastructure systems and maintaining cloud-native platforms such as Kubernetes, AWS, and Google Cloud.





