アラートの作成
似たような求人をメールで送る

Site Reliability Engineer - Global Tech - Up to ¥10M Hybrid

We are seeking an experienced Site Reliability Engineer (SRE) to design, build, and operate highly available infrastructure supporting large-scale consumer-facing digital services. This role focuses on reliability engineering, cloud infrastructure, automation, and continuous improvement while partnering closely with development teams to deliver resilient and scalable platforms.
Client Details
Our client is a leading global technology company operating high-volume digital platforms that process millions of user interactions and transactions every day. They are investing in modern cloud-based infrastructure, automation, and Site Reliability Engineering practices to deliver highly reliable customer experiences.
Description
* Define and continuously improve Service Level Indicators (SLIs) and Service Level Objectives (SLOs) in collaboration with engineering teams.
* Design highly available and resilient systems with redundancy, failover, disaster recovery, and backup strategies.
* Perform infrastructure capacity planning to balance scalability, performance, and cost efficiency.
* Design, build, and maintain Kubernetes-based infrastructure and CI/CD pipelines using Infrastructure as Code.
* Manage production operations, including releases, configuration changes, maintenance, and operational documentation.
* Participate in incident response, on-call rotations, root cause analysis, and blameless postmortems.
* Implement security best practices including vulnerability management, patching, access control, and secret management.
* Build automation tools and reduce operational overhead through scripting and AI-powered solutions.
* Collaborate with development teams to improve deployment processes, platform reliability, and developer productivity.
* Drive continuous improvements in observability, automation, and operational excellence.
Profile
* Bachelor's degree in Computer Science, Software Engineering, or equivalent practical experience.
* 8+ years of experience in Site Reliability Engineering, DevOps, or Infrastructure Engineering.
* 5+ years of hands-on experience designing and operating Kubernetes environments.
* Strong experience managing Linux-based production infrastructure.
* Experience designing scalable infrastructure for large-scale web or mobile applications.
* Solid understanding of networking technologies including DNS, CDN, load balancing, TLS, and firewalls.
* Experience designing and maintaining CI/CD pipelines using Jenkins, GitHub Actions, CircleCI, or similar tools.
* Proficiency with Infrastructure as Code tools such as Terraform or Pulumi.
* Hands-on experience with public cloud platforms including AWS, GCP, or Azure.
* Strong scripting and automation skills using Python, Shell, or similar languages.
* Experience using AI-assisted engineering tools such as GitHub Copilot, Claude Code, or similar.
* Strong knowledge of systems architecture, networking, and information security.
* Excellent communication, documentation, and stakeholder management skills.
* A proactive mindset with a passion for improving systems and engineering practices.
Job Offer
* Opportunity to build and operate mission-critical services at massive scale.
* Opportunity to work in a hybrid environment.
* Competitive salary and comprehensive employee benefits.
* Excellent opportunities for technical leadership and long-term career growth within a global technology organization.
To apply online please click the 'Apply' button below. For a confidential discussion about this role please contact Aahan Rawat on +81368328627.
似たような求人

Site Reliability Engineer - Global Tech - Up to ¥10M Hybrid

今すぐ適用する
Back to search page