Appealing Points
- A Developer-SRE hybrid role where you write and ship production code directly to the core product stack
- The chance to design carrier-grade, cloud-native orchestration and storage tools for global telecom partners
- Direct collaboration with the core engineering team, with your fixes contributed permanently to the main repository
Annual Salary: ¥9 million and above
Job Description
As a Forward Deployed Software Engineer, you will bridge core product development, site reliability and enterprise customer support. This is a highly technical Developer-SRE hybrid position, not a traditional operational role. You will be an active member of the development cycle, writing, contributing and shipping production code directly to the core product stack. At the same time, you will manage complex production escalations for global telecom partners. You will use your deep systems engineering expertise to design carrier-grade, cloud-native orchestration tools.
Job Responsibilities
- Production Development: Write and ship production code in Go, Python and C to develop new features, licensing tools and automation frameworks for the core platform.
- Core Orchestration and Storage: Design and develop core components of the orchestration and storage stack, with deep analysis across the Kubernetes CSI, CRI and CNI layers.
- Escalation Management: Rapidly resolve critical customer escalations by providing low-level workarounds, code modifications and thorough Root Cause Analysis (RCA).
- Software-Defined Storage (SDS): Architect and manage SDS frameworks, optimizing block, file and object storage infrastructure for production-grade telco applications.
- Lifecycle Management: Deploy new customer environments, manage upgrades, and develop and test robust upgrade procedures.
- Workflow Automation: Create application workflows (deployment, upgrades, backups) using Helm charts and Kubernetes operators.
- Distributed Systems: Ensure high availability, split-brain mitigation, consensus synchronization and state management for large-scale distributed setups.
- Workload Orchestration: Manage workloads across virtualization tiers, optimizing resource isolation and performance between containerized and virtualized environments.
- Capacity Planning: Design and execute data-driven cluster capacity planning strategies that maximize hardware utilization while meeting strict telco SLA and performance constraints.
- Patching and Debugging: Reproduce complex customer issues in a dedicated test lab, identify bugs and develop code patches to resolve them.
- Engineering Collaboration: Work directly with the core engineering team to root-cause deep-system product defects and contribute permanent fixes back to the main repository.
Required Qualifications
- Experience: 5–10 years of professional experience, with a strong background in software engineering, system design and handling complex customer escalations
- Education: Bachelor's or Master's degree in Computer Science, Engineering or an equivalent field
- Programming Skills: Advanced proficiency in C (or C++), Go and Python for low-level systems automation and feature development
- Linux and Kernel: Expert Linux system administration background, with hands-on experience in kernel-level troubleshooting, tracing and performance tuning
- Storage Expertise: Deep, practical understanding of Software-Defined Storage (SDS) across block, file and object storage. Robust knowledge of low-level data paths (DAS, Linux SCSI subsystem) and network storage fabrics (NFS, etc.)
- Kubernetes Mastery: In-depth knowledge of the Kubernetes stack, distributed system paradigms, containerization and public cloud architectures (specifically the CSI, CRI and CNI layers)
- Databases: Familiarity with relational, distributed and NoSQL databases
Language Requirements: Advanced Business level English
About Company:
The largest eCommerce company in Japan, and the third-largest eCommerce marketplace company worldwide. The organization provides a variety of consumer and business-focused services including e-commerce, e-reading, travel, banking, securities, credit card, e-money, portal and media, online marketing, and professional sports. The company is expanding globally and currently has operations throughout Asia, Western Europe, and the Americas.
[Measures against passive smoking]
No smoking indoors allowed
Designated smoking area
. Skillset Required: Go, Python, C, C++, Software-Defined Storage, Kubernetes CSI, Kubernetes CRI, Kubernetes CNI, Linux system administration, Kernel troubleshooting, Kernel tracing, Performance tuning, DAS, Linux SCSI subsystem, NFS, Distributed system paradigms, Containerization, Public cloud architectures, Relational databases, Distributed databases, NoSQL databases, Production code shipping, Root Cause Analysis, Capacity Planning, Workflow Automation, Helm charts, Kubernetes operators, Distributed Systems, Workload Orchestration, Virtualization, Software engineering, System design, Customer escalations, Low-level systems automation, Feature development