{"schemaVersion":"jobsearcher.job.v1","id":"14eece55f6c52a16e242f9a3","url":"https://jobsearcher.com/jobs/14eece55f6c52a16e242f9a3","canonicalUrl":"https://jobsearcher.com/jobs/14eece55f6c52a16e242f9a3","title":"Site Reliability Engineer","description":"Hybrid — Berkeley, CA\n 1 Year Contract Assignment with possibility of extension based on performance and organizational needs.\n $80/hr\n\nEver wondered what powers breakthrough research in energy, physics, materials science, and chemistry? You're looking at it. This national HPC facility supports 11,000+ scientists pushing the boundaries of what's possible, and we need a sharp, self-motivated SRE to help keep that engine running without interruption.\n\nIf you love solving real problems on live infrastructure, thrive on ownership, and want your work to directly enable world-class science, this is your seat.\n\nWhat You'll Own\nMonitor and triage alerts across compute, storage, network, and facility systems in real time\nBuild automation that prevents issues before they become outages\nDevelop new tools and integrations across the monitoring pipeline (APIs alerts action)\nWalk the data center floor to keep power, cooling, and environmental systems humming\nCoordinate maintenance activities across teams and keep incidents accurately tracked\nDig into complex, ambiguous problems and drive them to resolution\nWhat You Bring\nComfort working Owl shift (12am–8am), 5 days/week, hybrid onsite in Berkeley, CA\nSolid Linux/command-line (SSH) chops\nProgramming/scripting experience: Python, C, C++, Perl, or Java\nA self-starter mindset, eager to pick up Kubernetes, Prometheus/VictoriaMetrics, Alertmanager, and building management/cooling systems\nNetwork security fundamentals (ACLs, firewalls)\nStrong cross-team communication and collaboration skills\nNice to Have\nExperience building or deploying Agentic AI / autonomous automation for technical workflows\nServiceNow implementation experience\nITSM best-practice know-how\nlmuc9HhFwu","company":"Global","rawCompany":"global","city":"Berkeley","state":"CA","isRemote":false,"isActive":false,"createdAt":"2026-09-04T10:33:05.049Z","occupations":[{"code":"15-1299.08","title":"Computer Systems Engineers/Architects","slug":"computer-systems-engineers-architects"},{"code":"15-1244.00","title":"Network and Computer Systems Administrators","slug":"network-and-computer-systems-administrators"},{"code":"15-1211.00","title":"Computer Systems Analysts","slug":"computer-systems-analysts"}],"industries":[{"code":"541512","title":"Computer Systems Design Services","slug":"computer-systems-design-services"},{"code":"541519","title":"Other Computer Related Services","slug":"other-computer-related-services"},{"code":"518210","title":"Computing Infrastructure Providers, Data Processing, Web Hosting, and Related Services","slug":"computing-infrastructure-providers-data-processing-web-hosting-and-related-services"}],"jobPosting":{"@context":"https://schema.org","@type":"JobPosting","title":"Site Reliability Engineer","description":"Hybrid — Berkeley, CA\n 1 Year Contract Assignment with possibility of extension based on performance and organizational needs.\n $80/hr\n\nEver wondered what powers breakthrough research in energy, physics, materials science, and chemistry? You're looking at it. This national HPC facility supports 11,000+ scientists pushing the boundaries of what's possible, and we need a sharp, self-motivated SRE to help keep that engine running without interruption.\n\nIf you love solving real problems on live infrastructure, thrive on ownership, and want your work to directly enable world-class science, this is your seat.\n\nWhat You'll Own\nMonitor and triage alerts across compute, storage, network, and facility systems in real time\nBuild automation that prevents issues before they become outages\nDevelop new tools and integrations across the monitoring pipeline (APIs alerts action)\nWalk the data center floor to keep power, cooling, and environmental systems humming\nCoordinate maintenance activities across teams and keep incidents accurately tracked\nDig into complex, ambiguous problems and drive them to resolution\nWhat You Bring\nComfort working Owl shift (12am–8am), 5 days/week, hybrid onsite in Berkeley, CA\nSolid Linux/command-line (SSH) chops\nProgramming/scripting experience: Python, C, C++, Perl, or Java\nA self-starter mindset, eager to pick up Kubernetes, Prometheus/VictoriaMetrics, Alertmanager, and building management/cooling systems\nNetwork security fundamentals (ACLs, firewalls)\nStrong cross-team communication and collaboration skills\nNice to Have\nExperience building or deploying Agentic AI / autonomous automation for technical workflows\nServiceNow implementation experience\nITSM best-practice know-how\nlmuc9HhFwu","datePosted":"2026-09-04T10:33:05.049Z","dateModified":"2026-09-04T10:33:05.049Z","hiringOrganization":{"@type":"Organization","name":"Global","sameAs":"https://jobsearcher.com"},"jobLocation":{"@type":"Place","address":{"@type":"PostalAddress","addressLocality":"Berkeley","addressRegion":"CA","addressCountry":"US"}},"identifier":{"@type":"PropertyValue","name":"JobSearcher","value":"14eece55f6c52a16e242f9a3"},"url":"https://jobsearcher.com/jobs/14eece55f6c52a16e242f9a3"}}