Site Reliability Engineer
Hack The Box
| Company | Hack The Box |
| Category | Engineering |
| Location | Palaio Faliro |
| Remote | On-site (inferred) |
| Employment | Full-time |
| Level | Not stated |
| Salary | Not stated by the employer |
| Posted | 15 Jun 2026 |
| Last verified | 30 Jul 2026 |
| Source | Employer career page (workable) |
Description
Welcome! Super excited you dropped by 🥳 Let's redefine cyber security expertise standards and connect business - community through highly engaging hacking experiences. (Find out more insights about Hack The Box culture in our career site). ✨ The Core Mission of the Site Reliability Engineer (SRE): As a Site Reliability Engineer at Hack The Box, your paramount mission is to empower our Content Engineering team by providing reliable, scalable, and automated cloud infrastructure for our hands-on learning experiences. Over the next 6 months, you will participate in enhancing and simplifying the systems, services, and tools that enable Content Engineers to build and operate cloud labs efficiently. You will focus on infrastructure automation, observability, operational excellence, and continuous improvement of the workflows that support content creation and delivery. In parallel, you will contribute to the broader Site Reliability Engineering practice by helping maintain production services, improving platform reliability, supporting observability initiatives, and collaborating with fellow SREs on operational excellence efforts. While your primary focus will be the Content Engineering domain, you will remain closely aligned with the team’s reliability, automation, and infrastructure standards. ⚔️ Technology Tools & Weapons You’ll Be Using: Infrastructure as Code (Terraform): Automate the provisioning and management of cloud resources. Cloud Platforms (Google Cloud Platform, Microsoft Azure, AWS): Design, deploy, and operate infrastructure powering our cloud labs. Observability & Monitoring (Prometheus, Grafana, Mimir, Loki, Tempo): Maintain visibility into platform health and reliability. CI/CD & Automation: Improve and automate existing workflows and deployment processes. Collaboration & Enablement: Work closely with Content Engineers to improve developer experience and platform adoption. 🚀 The Adventures That Await Your Life Becoming a Site Reliability Engineer at Hack The Box: Heavily contribute to the reliability and scalability of the infrastructure powering Hack The Box cloud labs. Partner with Content Engineers to improve workflows, remove operational friction, and enable faster content delivery. Train and facilitate engineers on infrastructure best practices, Infrastructure as Code, and cloud-native technologies. Design, implement, and maintain Terraform-based infrastructure across multiple cloud providers. Build and enhance observability capabilities that improve operational visibility and incident response. Support production environments through maintenance, troubleshooting, and continuous improvement initiatives. Collaborate with the broader SRE team, contributing to shared platform reliability efforts when needed. Drive automation initiatives that reduce manual effort and improve consistency across systems and processes. 🏆 Skills, Knowledge, and Experience Points Required to Unlock the Role of SRE at Hack The Box: Hands-on experience with Terraform and Infrastructure as Code practices. Experience operating and supporting workloads in Microsoft Azure and/or Google Cloud Platform (GCP). Strong scripting and automation skills, ideally in Go but open for Python, Bash, or similar. Experience with monitoring, observability, and operational troubleshooting in production environments. Familiarity with CI/CD pipelines and developer enablement practices. Excellent communication and collaboration skills, with the ability to work closely with cross-functional engineering teams. Bonus Points: Previous participation in on-call rotations or incident response processes. Software development experience and familiarity with application development workflows. Background in cybersecurity, penetration testing, or security-focused environments. Experience contributing to realistic cloud architectures and operational workflows that support cybersecurity training scenari