English summary for screening — check the original posting before applying.
Stock is seeking an SRE to design and implement reliability for their rapidly growing SaaS products, "Stock" and "ナレカン". This role focuses on preventing failures and ensuring high availability, performance, and operational efficiency as the services scale.
Must-haves
- Experience as an SRE, or infrastructure design/construction/operation in a cloud environment (AWS etc.), or system design/operation considering availability, fault tolerance, and scalability.
- Experience in incident response, root cause analysis, and establishing recurrence prevention mechanisms, OR experience leading operational improvements and automation.
- Must be able to work onsite in Tokyo (hybrid model).
Nice-to-haves
- Experience with container orchestration (ECS/Fargate, Kubernetes)
- Experience with cloud network/security design and operation (VPC, IAM, WAF/CDN)
- Experience with managed data stores (RDS, OpenSearch) performance tuning
- Experience with auto-scaling, load testing, and cloud cost optimization
- Infrastructure as Code (IaC) experience
- CI/CD pipeline design/improvement experience
- Experience developing/operating BtoB SaaS products
- Experience in a startup or growth-stage company
Tech stack
AWSECSFargateKubernetesDockerCloudflareRuby on RailsReactTypeScriptPostgreSQLRedisOpenSearchNode.jsGitHub ActionsSentryPapertrail
Work style
Hybrid work: 3 days in office, 2 days remote. Location: Tokyo.
Other notes
Annual salary: ¥8,500,000 - ¥11,000,000.