Senior Service Reliability Engineer
PlayStation Global
| Company | PlayStation Global |
| Category | Engineering |
| Location | Germany |
| Remote | On-site (inferred) |
| Employment | Not stated |
| Level | Senior |
| Salary | Not stated by the employer |
| Posted | 20 Apr 2026 |
| Last verified | 2 Aug 2026 |
| Source | Employer ATS (greenhouse) |
Description
Why Sony Interactive Entertainment?
Sony Interactive Entertainment isn’t just the Best Place to Play — it’s also the Best Place to Work. Sony Interactive Entertainment (SIE) is the company behind the PlayStation brand. As a subsidiary of Sony Group Corporation, we’re part of a proud legacy of innovation and excellence. SIE is a dynamic technology company, delivering cutting-edge hardware and network services to more than 100 million people and an entertainment leader, home to some of the most beloved and recognizable intellectual properties (IP) in the world. Our role at SIE is to create and nurture the experiences under the PlayStation brand, a name synonymous with entertainment excellence and creativity. Senior Service Reliability Engineer
As a part of Sony Interactive Entertainment, the Gaming, Developer & Future Technology Group ( GDFT ) is leading the cloud gaming revolution, putting console-quality video games on any device, from TVs to consoles to mobile devices and beyond.
Our Site Reliability Engineering team plays a significant role in delivering on the promise of a great cloud gaming experience to our customers. We do this by influencing design and operational decisions towards the overall stability of the gaming service. Our SREs focus on three main things: overall ownership of production, production code quality, and deployments. The successful candidate will be self-directed and able to participate in the way we make decisions at different levels.
We expect our SREs to have opinions on the state of our service and provide critical feedback during different phases of the operational lifecycle. We are engaged throughout the software development lifecycle, ensuring operational readiness and stability.
What you'll be doing:
Taking a leadership role in ongoing improvements in Reliability and Scalability
Work closely with SRE Management to define KPIs, processes and drive continuous improvement
Influence the architecture and implementation of solutions within the division
Mentor more junior SRE staff and enable them for success
Act as a voice to represent SRE in the wider organisation
Represent the operational scalability of solutions in the wider division
Lead small-scale projects from inception to implementation
Design platform-wide solutions and provide technical leadership during their implementation
Demonstrate a high-level of organizational skills and initiative in the role
What we're looking for:
Minimum of 7+ years working experience in Software Development and/or Linux Systems Administration role.
Strong interpersonal, written and verbal communication skills.
Available to be scheduled in on-call rotation.
Skills & Knowledge:
Proficient as a Linux Production Systems Engineer, with experience managing large scale Web Services infrastructure.
Development experience in one or more of the following programming languages:
Python (preferred)
Bash, Go, Java, C++, or Rust
In addition, experience with at least 3 of the following topics:
Distributed data storage at scale (Hadoop, Ceph)
NoSQL at scale (MongoDB, Redis, Cassandra)
Data aggregation technologies. (ElasticSearch, Kafka)
Scaling and running traditional RDBMS (PostgreSQL, MySQL) with High Availability
Monitoring & alerting (Prometheus, Grafana), and Incident Management toolsets
Kubernetes and/or AWS (deployment and management)
Software distribution (Package management and distribution at scale)
Configuration management (ansible, saltstack, puppet, chef
2,088,683 openings. Erioun finds yours.Scored against your own profile, every hour.Try the radar →