スキルハウスの採用情報

Senior Site Reliability Engineer(8152)

役職:

Senior Site Reliability Engineer(8152)

雇用形態:

派遣

給与:

0.00

勤務時間:

Japanese Level - None,English Level - Advanced (TOEIC 860)

職務内容

 

 


A Global IT Service firm is looking for a Senior Site Reliability Engineer. As an SRE (Site Reliability Engineer), you will be required to be involved in activities to keep our systems reliable. Your core mission is to design and monitor the metrics in your target services and to make improvements when the targets (= SLO) are not achieved.

Responsibilities:
- Design of the SLI/SLOs: Define appropriate Service Level Indicators and Objectives for target services in collaboration with product and development teams, and continuously review them based on user
- Reliability & Availability Design: System Capacity Planning: Design systems for failure with redundancy, failover, backup/restore, and DR/BCP strategies, ensuring data integrity and recoverability for point/coupon transactions
- System Capacity Planning:  Forecast future resource demand based on traffic trends and business growth, and plan compute/storage/network capacity to ensure both reliability and cost
efficiency (FinOps)
- Server infra / CI/CD pipeline construction: Design and build server infrastructure and CI/CD pipelines using IaC, enabling safe, fast, and repeatable delivery
- Operations in the Production environment: Operate production systems including release management, configuration changes, and routine maintenance while minimizing user impact, and maintain runbooks/playbooks for consistent operations
- Incident management & On-call: Participate in the on-call rotation, respond to production incidents with clear escalation, and act as incident coordinator when needed.
- Production trouble shooting and postmotem: Respond to production incidents, lead root-cause analysis, and drive blameless postmortems to prevent recurrence and improve reliability
- Security & Compliance: Drive vulnerability response, security patching, secret management, and access control in cooperation with security teams, and support audit requirements for a service handling monetary value.
- Toil reduction & Developer enablement: Reduce operational toil and provide reliability guidance, golden paths, and self-service tooling to development teams to lower their cognitive load
- AI & Automation: Identify and automate repetitive development lifecycle tasks using automation scripts and AI agents

Required skills:
- 8+ years of hands-on experience in SRE, DevOps, or infrastructure engineering
- 5+ years of hands-on experience in Kubernetes (design, deployment, scaling, and maintenance)
- 5+ years of hands-on experience in production system operations on Linux OS (release, server configuration change, database operations and security patching)
- Strong experience designing the server infrastructure environment from scratch, with considering the network efficiency and the machine capacity
- Strong experience with networking fundamentals for large-scale web/app services (DNS, CDN, load balancing, TLS, firewalls etc.)
- Strong experience designing and maintaining CI/CD pipelines (Jenkins, GitHub Actions, CircleCI, etc.) and system monitoring environments (Prometheus, Grafana, ELK Stack, OpenTelemetry etc.)
- Proficiency in IaC (= Infrastructure as a Code) tools (Terraform, Pulumi, or similar)
- Proficiency in deploying and operating services on public cloud (GCP, AWS, Azure etc.)
- Proficiency in scripting/automation (Python, Shell, or similar) for operational tooling
- Proficiency in AI-agentic coding/operation (GitHub Copilot, Claude Code, Codex etc.)
- Enough knowledge about computer architecture, network configuration and information security
- Strong communication, technical writing and stakeholder involvement skills to lead a project to its success
- Mindset to proactively improve the status quo to make the team better

Why should you apply:
- This is a long-term opportunity with a chance to become a permanent employee 
- You will be working with international team members 
- Free breakfast, lunch and dinner at the cafeteria 


Company Details: 
A global company with a strong presence in multiple business areas. It has achieved sustained growth both domestically and internationally, including in the U.S. and Europe. The company boasts a diverse and international environment and is committed to equal opportunity, offering a wealth of career opportunities. Due to the diverse nature of our business, we handle a wide range of technologies! You can also choose the environment you are most comfortable with, such as Windows/Mac! Meals in the company cafeteria are also free. Our chefs are always coming up with new menu items, so you can enjoy your meal without getting bored!

Working hours: 9:00 - 17:30 (Mon-Fri)
Working Style: Hybrid (4 days in office, 1 day work from home)
Holidays:  Saturday, Sunday, and National Holidays, Year-end and New Year Holidays, Paid Holidays, Other Special Holidays
Services/Benefits:  Social insurance, DC Pension Plan, Transportation Fee, Skillhouse University, Test payback system, and more


スキルハウスで共に成長し、学び、成功を目指しませんか。

スキルハウスのこのポジションにご興味のある方は、ご連絡先をご記入の上、履歴書を添付してください。

 ※右記個人情報は、採用選考のみに利用されます

東京都港区虎ノ門3-8-27巴町アネックス2号館

internalcareers@skillhouse.co.jp

Internal Vacancy Form