{"schemaVersion":"jobsearcher.job.v1","id":"b13faf9bbe5c7efadef703f5","url":"https://jobsearcher.com/jobs/b13faf9bbe5c7efadef703f5","canonicalUrl":"https://jobsearcher.com/jobs/b13faf9bbe5c7efadef703f5","title":"Senior DevOps Engineer - Highload, Cloud & Data-Intensive Systems (EU / Remote)","description":"About The Project The team develops and maintains distributed services around analytics, APIs, and transaction monitoring. The systems process very large volumes of data — terabytes of storage, trillions of records, continuously growing load.Infrastructure100 servers (bare metal + VPS)active use of IaCKubernetes clusters in productionfocus on stability, observability, and automationThe project is long-term — not a hype startup, but a mature product with real users.What The Work Looks Like This is a hands‐on role with a clear time allocation:60% — operations and incidents (including helping teams)20% — infrastructure automation20% — prototyping, improvements, technical initiativesThere is on‐call responsibility, but normally after‐hours incidents happen 2‐3 times a year, not every week.ResponsibilitiesOperation of production services and infrastructure (server provisioning/decommissioning, updates, replacements, performance troubleshooting)Support and development of Infrastructure as Code (Terraform / Ansible: modules, roles, standards, reviews)Monitoring, alerting, backups, and regular recovery checksDevelopment of service and infrastructure automationDevelopment of CI/CD and release proceduresIncident diagnosis and resolution, support for product teamsTraffic analytics, bot and attack protection toolsResponsibility for 24/7 platform stabilityRequirements4+ years of experience operating Linux/Ubuntu infrastructure and production servicesStrong understanding of networking and troubleshootingKubernetes (cluster operations), Rancher, Docker / containerdHands‐on experience with Ansible and TerraformMonitoring: Prometheus / Thanos / Telegraf / Grafana / SentryCI/CD: JenkinsAutomation: Bash, PythonExperience working with LVMNice to haveExperience working with blockchain nodesDiagnosis and tuning of ClickHouse and MongoDB in high‐load clustersProviders: Hetzner / OVHcloudCloudflare (edge, DDoS), experience with AWSHandling abuse tickets with hosting providersTechnology stackVPN: WireGuard, OpenVPNDatabases: ClickHouse, MongoDB, Redis, PostgreSQLApplications: Node.js (pm2), php-fpm, Lua, TarantoolSupporting services: Go (operatorSDK), Ruby, Node.js, PHPBenefits5,000 - 8,000 € netFormat: office / hybrid / remoteLocation: Spain (Barcelona and suburbs) or remote (CET ±2)Full-timeOpportunity to genuinely influence architecture and processesMature engineering team and reasonable expectations#J-18808-Ljbffr","company":"Alex Staff","rawCompany":"alex staff","isRemote":true,"isActive":false,"createdAt":"2026-04-09T12:03:45.221Z","occupations":[{"code":"15-1299.08","title":"Computer Systems Engineers/Architects","slug":"computer-systems-engineers-architects"},{"code":"15-1252.00","title":"Software Developers","slug":"software-developers"},{"code":"15-1244.00","title":"Network and Computer Systems Administrators","slug":"network-and-computer-systems-administrators"}],"industries":[{"code":"541512","title":"Computer Systems Design Services","slug":"computer-systems-design-services"},{"code":"541511","title":"Custom Computer Programming Services","slug":"custom-computer-programming-services"},{"code":"513210","title":"Software Publishers","slug":"software-publishers"}],"jobPosting":{"@context":"https://schema.org","@type":"JobPosting","title":"Senior DevOps Engineer - Highload, Cloud & Data-Intensive Systems (EU / Remote)","description":"About The Project The team develops and maintains distributed services around analytics, APIs, and transaction monitoring. The systems process very large volumes of data — terabytes of storage, trillions of records, continuously growing load.Infrastructure100 servers (bare metal + VPS)active use of IaCKubernetes clusters in productionfocus on stability, observability, and automationThe project is long-term — not a hype startup, but a mature product with real users.What The Work Looks Like This is a hands‐on role with a clear time allocation:60% — operations and incidents (including helping teams)20% — infrastructure automation20% — prototyping, improvements, technical initiativesThere is on‐call responsibility, but normally after‐hours incidents happen 2‐3 times a year, not every week.ResponsibilitiesOperation of production services and infrastructure (server provisioning/decommissioning, updates, replacements, performance troubleshooting)Support and development of Infrastructure as Code (Terraform / Ansible: modules, roles, standards, reviews)Monitoring, alerting, backups, and regular recovery checksDevelopment of service and infrastructure automationDevelopment of CI/CD and release proceduresIncident diagnosis and resolution, support for product teamsTraffic analytics, bot and attack protection toolsResponsibility for 24/7 platform stabilityRequirements4+ years of experience operating Linux/Ubuntu infrastructure and production servicesStrong understanding of networking and troubleshootingKubernetes (cluster operations), Rancher, Docker / containerdHands‐on experience with Ansible and TerraformMonitoring: Prometheus / Thanos / Telegraf / Grafana / SentryCI/CD: JenkinsAutomation: Bash, PythonExperience working with LVMNice to haveExperience working with blockchain nodesDiagnosis and tuning of ClickHouse and MongoDB in high‐load clustersProviders: Hetzner / OVHcloudCloudflare (edge, DDoS), experience with AWSHandling abuse tickets with hosting providersTechnology stackVPN: WireGuard, OpenVPNDatabases: ClickHouse, MongoDB, Redis, PostgreSQLApplications: Node.js (pm2), php-fpm, Lua, TarantoolSupporting services: Go (operatorSDK), Ruby, Node.js, PHPBenefits5,000 - 8,000 € netFormat: office / hybrid / remoteLocation: Spain (Barcelona and suburbs) or remote (CET ±2)Full-timeOpportunity to genuinely influence architecture and processesMature engineering team and reasonable expectations#J-18808-Ljbffr","datePosted":"2026-04-09T12:03:45.221Z","dateModified":"2026-04-09T12:03:45.221Z","hiringOrganization":{"@type":"Organization","name":"Alex Staff","sameAs":"https://jobsearcher.com"},"jobLocationType":"TELECOMMUTE","applicantLocationRequirements":{"@type":"Country","name":"US"},"identifier":{"@type":"PropertyValue","name":"JobSearcher","value":"b13faf9bbe5c7efadef703f5"},"url":"https://jobsearcher.com/jobs/b13faf9bbe5c7efadef703f5"}}