T

Mid-Level Site Reliability Engineer

TEKsystems Japan
Visa Sponsorship
Apply
AI Summary

We are looking for a mid-level Site Reliability Engineer who can balance software engineering and Kubernetes platform operations. This role focuses on running and improving production-grade container platforms while building automation and tooling to support internal engineering teams. The ideal candidate will have experience operating Kubernetes in production environments at scale and software development experience with statically typed languages.

Key Highlights
Operate and maintain production Kubernetes clusters and nodes
Design and develop automation and platform tools using Go
Support internal users and assist with platform and service migrations
Key Responsibilities
Operate and maintain production Kubernetes clusters and nodes
Respond to incidents, alerts, and participate in on-call rotations
Perform OS, middleware, and security-related updates
Execute high-risk production changes during off-peak hours
Design and develop automation and platform tools using Go
Create and follow strict operational procedures and runbooks
Support internal users and assist with platform and service migrations
Technical Skills Required
Kubernetes Go Linux
Benefits & Perks
Hybrid work environment
Visa sponsorship available for overseas candidates
International environment
Nice to Have
Strong proficiency in Go
Experience automating large-scale infrastructure
Private and/or public cloud experience
Monitoring and observability tools (Prometheus, Grafana, ELK, etc.)
Contributions to open-source projects

Job Description


  • Hybrid Work Environment
  • Visa sponsorship available for overseas candidates
  • International environment

We are looking for a mid‑level Site Reliability Engineer who can balance software engineering and Kubernetes platform operations. This role focuses on running and improving production‑grade container platforms while building automation and tooling to support internal engineering teams.

Responsibilities

  • Operate and maintain production Kubernetes clusters and nodes
  • Respond to incidents, alerts, and participate in on‑call rotations
  • Perform OS, middleware, and security‑related updates
  • Execute high‑risk production changes during off‑peak hours
  • Design and develop automation and platform tools using Go
  • Create and follow strict operational procedures and runbooks
  • Support internal users and assist with platform and service migrations

Required Skills & Experience

  • 3+ years operating Kubernetes in production environments at scale
  • 3+ years software development experience with statically typed languages (Go, Java, C++, Rust, etc.)
  • Solid understanding of Linux and basic networking (TCP/IP)
  • Experience working in environments with strong operational processes and controls
  • Willingness to support late‑night or early‑morning operational work when required
  • CKA certification (or ability to obtain within 3 months)
  • English communication skills (business level)

Nice to Have

  • Strong proficiency in Go
  • Experience automating large‑scale infrastructure
  • Private and/or public cloud experience
  • Monitoring and observability tools (Prometheus, Grafana, ELK, etc.)
  • Contributions to open‑source projects

Similar Jobs

Explore other opportunities that match your interests

LLM Engineer

Programming
1h ago
Visa Sponsorship Relocation Remote
Job Type Full-time
Experience Level Mid-Senior level

Superstars

Japan

Machine Learning Engineer

Programming
15h ago

Premium Job

Sign up is free! Login or Sign up to view full details.

•••••• •••••• ••••••
Job Type ••••••
Experience Level ••••••

appier group (東証プライム:41...

Japan
Visa Sponsorship Relocation Remote
Job Type Full-time
Experience Level Mid-Senior level

hennge

Japan

Subscribe our newsletter

New Things Will Always Update Regularly