KOMOJU (by Degica) is the leading cross-border payment gateway for Japan. We power payments for companies like video game distribution platform Steam and the popular mobile app TikTok. Today we help thousands of merchants by providing them with the payment infrastructure they need through developer-friendly API’s to integrations on popular platforms like Shopify and Wix; we help our merchants grow in all markets they are expanding.
As our systems grow in complexity, scale, and traffic , maintaining their reliability and availability becomes increasingly challenging—and critical. We're looking for a Site Reliability Engineer (SRE) with a focus for observability to help us meet these demands.
In this role, you'll be at the forefront of ensuring that our infrastructure is not just running, but understandable and measurable . Observability is a core pillar of our reliability strategy—it's how we detect issues before they impact our merchants and users, quickly understand the root causes of incidents, and continuously improve our systems performance and reliability.
You’ll design and evolve our observability platform, including metrics, logging, tracing, and alerting , and partner with development teams to embed observability into every stage of the software lifecycle. Your work will directly impact our ability to scale confidently and respond to incidents swiftly.
This is a key role for someone who wants to build resilient systems , empower teams with actionable insights , and make a real difference in how we operate at scale.
While we are a remote-first company, this position is based in Tokyo, and we expect candidates to be willing to relocate to Japan.
Design, implement, and maintain our observability stack (metrics, logging, tracing, dashboards).
Define and monitor SLIs/SLOs to ensure service health and reliability.
Correspond with engineering teams to instrument applications for better visibility.
Build and maintain dashboards and alerts that provide actionable insights and minimize alert fatigue.
Troubleshoot system performance and reliability issues using observability data.
Educate and guide engineering teams on best practices in monitoring, alerting, and incident response.
Contribute to postmortems and continuously improve system transparency and resiliency.
Knowledge of CI/CD pipelines and integrating observability into build and deploy processes.
Familiarity with incident response , on-call rotations, and post-incident reviews.
Business-level Japanese.
Job DescriptionInsight Global is looking for a talented Learning Content Editor/Writer to join our team and lead the standardization of content templates used across various learning materials. This role involves reviewing and standardizing templates for onboarding, success...
...Join Our Team as a Window Installer Immediate Opening! Are you looking for a new opportunity? We want to hear from you! TruHome is a rapidly growing home remodeling company, and we are looking for a Window Installer to join our team in the Madison, WI area. We...
...Job Posting Title: Lab Manager, Department of Civil, Architectural and Environmental Engineering ---- Hiring Department: Fariborz Maseeh Department of Civil, Architectural and Environmental Engineering ---- Position Open To: All Applicants ---- Weekly...
...designs; posts and submits materials to County website and media outlets. Receives and responds to media inquiries and questions: consults with County staff and management to provide appropriate response to media inquiries; provides and/or coordinates interviews as...
...that matters! Certified Nursing Assistants (CNA) are at the heart of what we do, providing... ...Aide Registry or completion of approved training with competency evaluation. If this... ...is committed to providing a drug-free, tobacco-free, healthy, safe workplace for...