Site Reliability Engineer
ARCHIVED
We can't find an active application page for this role right now. It may reopen or be listed elsewhere. Use Next Steps to search for an active apply link and similar live jobs.
The Core Mission of the Site Reliability Engineer (SRE):
As a Site Reliability Engineer at Hack The Box, your paramount mission is to empower our Content Engineering team by providing reliable, scalable, and automated cloud infrastructure for our hands‑on learning experiences. Over the next 6 months, you will enhance and simplify the systems, services, and tools that enable Content Engineers to build and operate cloud labs efficiently, focusing on infrastructure automation, observability, operational excellence, and continuous improvement.
In parallel, you will contribute to the broader Site Reliability Engineering practice by maintaining production services, improving platform reliability, supporting observability initiatives, and collaborating with fellow SREs on operational excellence efforts.
Technology Tools & Weapons You’ll Be Using:
Infrastructure as Code (Terraform): Automate the provisioning and management of cloud resources.
Cloud Platforms (Google Cloud Platform, Microsoft Azure, AWS): Design, deploy, and operate infrastructure powering our cloud labs.
Observability & Monitoring (Prometheus, Grafana, Mimir, Loki, Tempo): Maintain visibility into platform health and reliability.
CI/CD & Automation: Improve and automate existing workflows and deployment processes.
Collaboration & Enablement: Work closely with Content Engineers to improve developer experience and platform adoption.
The Adventures That Await Your Life Becoming a Site Reliability Engineer at Hack The Box:
Contribute heavily to the reliability and scalability of the infrastructure powering Hack The Box cloud labs.
Partner with Content Engineers to improve workflows, remove operational friction, and enable faster content delivery.
Train and facilitate engineers on infrastructure best practices, Infrastructure as Code, and cloud‑native technologies.
Design, implement, and maintain Terraform‑based infrastructure across multiple cloud providers.
Build and enhance observability capabilities that improve operational visibility and incident response.
Support production environments through maintenance, troubleshooting, and continuous improvement initiatives.
Collaborate with the broader SRE team, contributing to shared platform reliability efforts when needed.
Drive automation initiatives that reduce manual effort and improve consistency across systems and processes.
Skills, Knowledge, and Experience Points Required to Unlock the Role of SRE at Hack The Box:
Hands‑on experience with Terraform and Infrastructure as Code practices.
Experience operating and supporting workloads in Microsoft Azure and/or Google Cloud Platform (GCP).
Strong scripting and automation skills, ideally in Go but open for Python, Bash, or similar.
Experience with monitoring, observability, and operational troubleshooting in production environments.
Familiarity with CI/CD pipelines and developer enablement practices.
Excellent communication and collaboration skills, with the ability to work closely with cross‑functional engineering teams.
Bonus Points:
Previous participation in on‑call rotations or incident response processes.
Software development experience and familiarity with application development workflows.
Background in cybersecurity, penetration testing, or security‑focused environments.
Experience contributing to realistic cloud architectures and operational workflows that support cybersecurity training scenarios.
Experience with Kubernetes, containers, and cloud‑native technologies.
The gems you’ll be enjoying as a Site Reliability Engineer:
Private health care
Paid paternity leave
25 annual leave days
Free lunch & snacks at the office
120€ Ticket Restaurant by Edenred
Dedicated budget for training and professional development, participation in conferences
Full access to the Hack The Box lab offerings; so you can learn how to hack
State‑of‑the‑art equipment (mac, iPhone, and mobile plan)
Flexible WFH (Hybrid Model) – Fully Remote is also an option if you're not an Attica resident
Our benefits package is designed to provide strong support to our team, but it may vary depending on location and type of employment (e.g., UK, Greece, or engagement through an Employer of Record).
At Hack The Box, we are committed to fostering a diverse, inclusive, and equitable workplace. We believe that diversity enriches our performance, services, and the communities we serve. As such, we ensure that all job applications are considered solely based on merit, skills, and qualifications. We do not discriminate on grounds of race, color, religion, gender, gender identity or expression, sexual orientation, national origin, genetics, disability, age, or veteran status. We are dedicated to providing a fair and respectful work environment that reflects our values.
#J-18808-Ljbffr