Senior SRE

freee

SRE

AWS

Docker

About this role

This senior SRE position leads cloud infrastructure design and operations for a multi-product SaaS environment built around integrated cloud software for small businesses and individual customers. The role addresses increasing operational complexity across multiple products and microservices, reducing engineer cognitive load and operational toil while providing support suited to each product domain.

Responsibilities include leading cloud architecture and operations using AWS and EKS; building AI-native operations and AIOps practices that strategically use AI and LLMs; introducing SLI and SLO practices and helping product teams adopt them; designing standardized infrastructure as code and self-service environments that support rapid development; making company-wide technical decisions and preparing a long-term technology roadmap; leading responses to major incidents and directing root-cause prevention work; and solving user-facing problems through application and database performance tuning. Assigned duties may change depending on business circumstances and the individual’s suitability.

Required experience includes leading system architecture design through operations using cloud technologies such as AWS, GCP, or Azure; designing and operating production services with Docker and Kubernetes, including decisions about technology selection and configuration policy; leading incident response from cause analysis through implementation of recurrence-prevention measures; leading the design and improvement of CI/CD pipelines that automate build, test, and deployment while improving deployment frequency or lead time; and either application development experience or knowledge gained by solving problems through application or database performance tuning.

Preferred qualifications include leading the design or migration of microservices or cloud-native architectures; designing and implementing capacity planning or availability improvements for large-scale products; establishing SRE or Platform Engineering organizations and spreading a reliability culture; coordinating or leading technical issues across multiple teams; and technically developing or mentoring team members. The posting also seeks someone who can define essential problems in uncertain and complex systems, work across infrastructure and product-development boundaries, and choose appropriately between deterministic approaches and AI rather than assuming AI is always the answer.

The position is full-time with a three-month probationary period. The work location is the Tokyo headquarters in Osaki, Shinagawa-ku, Tokyo, at Art Village Osaki Central Tower, 21F. The work system is a discretionary labor system with a deemed eight-hour workday. Holidays include Saturdays, Sundays, public holidays, and the year-end and New Year period. Paid leave is granted upon joining, and six paid sick-leave days are provided annually. Benefits include employment insurance, workers’ compensation insurance, employees’ pension insurance, and health insurance. Smoking is prohibited indoors. Online interviews are available.

This is an AI-generated summary of the employer's original posting — details can be incomplete, out of date or simply wrong. Always confirm everything on the official posting before applying.