Senior DevOps Engineer - Highload, Cloud & Data-Intensive Systems (EU / Remote)
ARCHIVED
We can't find an active application page for this role right now. It may reopen or be listed elsewhere. Use Next Steps to search for an active apply link and similar live jobs.
About The Project The team develops and maintains distributed services around analytics, APIs, and transaction monitoring. The systems process very large volumes of data — terabytes of storage, trillions of records, continuously growing load.Infrastructure100 servers (bare metal + VPS)active use of IaCKubernetes clusters in productionfocus on stability, observability, and automationThe project is long-term — not a hype startup, but a mature product with real users.What The Work Looks Like This is a hands‐on role with a clear time allocation:60% — operations and incidents (including helping teams)20% — infrastructure automation20% — prototyping, improvements, technical initiativesThere is on‐call responsibility, but normally after‐hours incidents happen 2‐3 times a year, not every week.ResponsibilitiesOperation of production services and infrastructure (server provisioning/decommissioning, updates, replacements, performance troubleshooting)Support and development of Infrastructure as Code (Terraform / Ansible: modules, roles, standards, reviews)Monitoring, alerting, backups, and regular recovery checksDevelopment of service and infrastructure automationDevelopment of CI/CD and release proceduresIncident diagnosis and resolution, support for product teamsTraffic analytics, bot and attack protection toolsResponsibility for 24/7 platform stabilityRequirements4+ years of experience operating Linux/Ubuntu infrastructure and production servicesStrong understanding of networking and troubleshootingKubernetes (cluster operations), Rancher, Docker / containerdHands‐on experience with Ansible and TerraformMonitoring: Prometheus / Thanos / Telegraf / Grafana / SentryCI/CD: JenkinsAutomation: Bash, PythonExperience working with LVMNice to haveExperience working with blockchain nodesDiagnosis and tuning of ClickHouse and MongoDB in high‐load clustersProviders: Hetzner / OVHcloudCloudflare (edge, DDoS), experience with AWSHandling abuse tickets with hosting providersTechnology stackVPN: WireGuard, OpenVPNDatabases: ClickHouse, MongoDB, Redis, PostgreSQLApplications: Node.js (pm2), php-fpm, Lua, TarantoolSupporting services: Go (operatorSDK), Ruby, Node.js, PHPBenefits5,000 - 8,000 € netFormat: office / hybrid / remoteLocation: Spain (Barcelona and suburbs) or remote (CET ±2)Full-timeOpportunity to genuinely influence architecture and processesMature engineering team and reasonable expectations#J-18808-Ljbffr