Transforming global manufacturers' engineering drawings, documents, and supply-chain data management. Full reliability posture, automation-first infrastructure, and direct partnership with engineering. 9+ years of hands-on software development experience and 7+ years in SRE or a closely related role.
Key Highlights
Key Responsibilities
Technical Skills Required
Benefits & Perks
Nice to Have
Job Description
🔧 Hiring: Senior Site Reliability Engineer (SRE) 📍 Tokyo, Japan | Full-time | Full remote
My client is a product team within a B2B SaaS company transforming how global manufacturers manage engineering drawings, documents, and supply-chain data. Their platform is trusted across the manufacturing industry, and they're now preparing for the full-scale launch of a new product built on top of it.
What you'll own:
- Full reliability posture of the product: monitoring, alerting, SLIs/SLOs, incident response, and post-mortem culture
- Automation-first infrastructure on GCP/GKE so the team scales without operational drag
- CI/CD pipelines with a focus on delivery safety and developer productivity
- Direct partnership with engineering to embed reliability into product design
- Reliability culture and leadership
Interested in remote work opportunities in Development & Programming? Discover Development & Programming Remote Jobs featuring exclusive positions from top companies that offer flexible work arrangements.
What they're looking for:
- 9+ years of hands-on software development experience
- 7+ years in SRE, platform engineering, or a closely related role
- Strong cloud infrastructure background with IaC (Terraform or equivalent)
- Production-grade Kubernetes experience
- Experience designing and running CI/CD pipelines with reliability in mind
- Web application development and production troubleshooting experience
Browse our curated collection of remote jobs across all categories and industries, featuring positions from top companies worldwide.
Bonus points for: GCP hands-on experience, Datadog or observability platforms, experience scaling SRE culture in a 50+ engineer org, and business-level Japanese (JLPT N2 or equivalent).
Tech stack highlights:
GCP · GKE · Terraform · ArgoCD · GitHub Actions · Helm · Kustomize · Istio · Datadog · Sentry · AlloyDB · BigQuery · Cloud Pub/Sub · Rust · TypeScript · Cloudflare · Auth0
SRE #SiteReliabilityEngineering #DevOps #GCP #Kubernetes #Infrastructure
Similar Jobs
Explore other opportunities that match your interests