{"schemaVersion":"jobsearcher.job.v1","id":"0891c55da5420b2a0c374788","url":"https://jobsearcher.com/jobs/0891c55da5420b2a0c374788","canonicalUrl":"https://jobsearcher.com/jobs/0891c55da5420b2a0c374788","title":"Site Reliability Engineer","description":"ABOUT NEXTPOINT\nNextpoint builds transformative software and services for the legal industry — making eDiscovery, case management, and litigation prep simple, fluid, and affordable for law firms of all sizes. Our secure, cloud-based platform lets teams start document review in minutes, backed by powerful analytics, an intuitive interface, and best-in-class security at every point.\nWe're problem solvers, simplifiers, and challenge seekers, united by a shared goal: a great team culture and satisfied clients. We're headquartered in Chicago's Ravenswood neighborhood and proud to have been named one of Built In's Best Startups to Work for in Chicago five years running (2022–2026).\nABOUT THE ROLE\nNextpoint's platform handles massive, unpredictable volumes of sensitive legal data — unlimited-upload document review, AI-assisted analysis, and secure electronic production — all running on AWS with zero downtime tolerance for firms in active litigation. We're looking for a Site Reliability Engineer to help maintain the reliability, scalability, and security posture of that platform as we expand our AI capabilities (built on Amazon Bedrock) and grow our customer base.\nThis is a hands-on role for someone who wants to be the first line of response for production infrastructure at a company where reliability is a customer-trust issue, not just an engineering metric.\nRESPONSIBILITIES\nInfrastructure & Automation (50%)\n\nMaintain and extend existing infrastructure-as-code (Terraform/CloudFormation/CDK) following established patterns and standards\nSupport and operate CI/CD pipelines; implement improvements as directed\nMonitor cloud cost trends and flag optimization opportunities for review\n\nReliability & Operations (25%)\n\nMonitor uptime, latency, and performance SLOs/SLIs for production systems supporting document upload, processing, review, and production workflows\nParticipate in the on-call rotation and serve as first responder for production incidents during US business hours\nTriage, troubleshoot, and resolve incoming infrastructure requests and incidents; **escalate** and coordinate on complex root-cause work\nWrite clear post-incident reports and help investigate recurring incidents and cost overruns\nMaintain and extend existing monitoring, alerting, and observability tooling\n\nCross-Functional Collaboration (15%)\n\nPartner with engineering teams to build reliability, scalability, and observability into new features from design through launch\nDocument runbooks, architecture decisions, and operational procedures for the broader engineering team\nCommunicate incident status and technical issues clearly to engineering and non-technical stakeholders\nParticipate in design reviews to flag reliability or operational concerns early\n\nSecurity & Compliance (10%)\n\nSupport SOC 2 compliance activities and help uphold encryption, access-control, and audit-trail standards across all environments\nImplement and maintain security best practices for infrastructure handling confidential legal and client data, including AI workloads on Amazon Bedrock\nSupport security reviews, vulnerability management, and patching cadences across production systems\nMaintain permissions-based access controls and comprehensive audit logging in line with client security commitments\n\nQUALIFICATIONS\n\n5+ years of experience in Site Reliability Engineering, DevOps, or Infrastructure Engineering roles, ideally in a B2B SaaS environment\nDeep hands-on experience with AWS (EC2, S3, RDS, Lambda, VPC, IAM, CloudWatch, or equivalent services)\nWorking knowledge of infrastructure-as-code (Terraform, CloudFormation, or CDK) and configuration management\nProficiency in at least one scripting/programming language (Python, Go, or similar) for automation and tooling\nExperience with containerization and orchestration (Docker, Kubernetes, or ECS)\nTrack record of operating CI/CD pipelines\nExperience with monitoring/observability stacks (Datadog, CloudWatch, Prometheus/Grafana, or similar)\nFamiliarity with security compliance frameworks (SOC 2, HIPAA, or similar) and encryption/access-control best practices\nExperience supporting systems handling large-scale, variable-volume data processing is a plus\nExposure to AI/ML infrastructure (e.g., Amazon Bedrock, model-serving pipelines) is a plus, given our growing AI feature set\nExperience using ClaudeCode, Kiro, OpenCode or similar Agentic AI\nBachelor's degree in Computer Science, Engineering, or related field (or equivalent experience)\nStrong written and verbal communication skills, with the ability to write clear runbooks and explain technical tradeoffs to non-technical stakeholders\nComfortable being part of an on-call rotation\n\nEQUAL OPPORTUNITY EMPLOYER\nNextpoint is an equal opportunity employer. We actively work to build a diverse team and encourage candidates of all backgrounds to apply. All applicants are considered without regard to race, color, religion, sex, sexual orientation, gender identity, national origin, veteran status, disability, or any other characteristic protected by applicable law. \n#J-18808-Ljbffr","company":"Nextpoint","rawCompany":"nextpoint","city":"Eastern","state":"KY","isRemote":false,"isActive":false,"createdAt":"2026-10-03T03:44:59.436Z","occupations":[{"code":"15-1299.08","title":"Computer Systems Engineers/Architects","slug":"computer-systems-engineers-architects"},{"code":"15-1252.00","title":"Software Developers","slug":"software-developers"},{"code":"15-1211.00","title":"Computer Systems Analysts","slug":"computer-systems-analysts"}],"industries":[{"code":"541512","title":"Computer Systems Design Services","slug":"computer-systems-design-services"},{"code":"518210","title":"Computing Infrastructure Providers, Data Processing, Web Hosting, and Related Services","slug":"computing-infrastructure-providers-data-processing-web-hosting-and-related-services"},{"code":"541519","title":"Other Computer Related Services","slug":"other-computer-related-services"}],"jobPosting":{"@context":"https://schema.org","@type":"JobPosting","title":"Site Reliability Engineer","description":"ABOUT NEXTPOINT\nNextpoint builds transformative software and services for the legal industry — making eDiscovery, case management, and litigation prep simple, fluid, and affordable for law firms of all sizes. Our secure, cloud-based platform lets teams start document review in minutes, backed by powerful analytics, an intuitive interface, and best-in-class security at every point.\nWe're problem solvers, simplifiers, and challenge seekers, united by a shared goal: a great team culture and satisfied clients. We're headquartered in Chicago's Ravenswood neighborhood and proud to have been named one of Built In's Best Startups to Work for in Chicago five years running (2022–2026).\nABOUT THE ROLE\nNextpoint's platform handles massive, unpredictable volumes of sensitive legal data — unlimited-upload document review, AI-assisted analysis, and secure electronic production — all running on AWS with zero downtime tolerance for firms in active litigation. We're looking for a Site Reliability Engineer to help maintain the reliability, scalability, and security posture of that platform as we expand our AI capabilities (built on Amazon Bedrock) and grow our customer base.\nThis is a hands-on role for someone who wants to be the first line of response for production infrastructure at a company where reliability is a customer-trust issue, not just an engineering metric.\nRESPONSIBILITIES\nInfrastructure & Automation (50%)\n\nMaintain and extend existing infrastructure-as-code (Terraform/CloudFormation/CDK) following established patterns and standards\nSupport and operate CI/CD pipelines; implement improvements as directed\nMonitor cloud cost trends and flag optimization opportunities for review\n\nReliability & Operations (25%)\n\nMonitor uptime, latency, and performance SLOs/SLIs for production systems supporting document upload, processing, review, and production workflows\nParticipate in the on-call rotation and serve as first responder for production incidents during US business hours\nTriage, troubleshoot, and resolve incoming infrastructure requests and incidents; **escalate** and coordinate on complex root-cause work\nWrite clear post-incident reports and help investigate recurring incidents and cost overruns\nMaintain and extend existing monitoring, alerting, and observability tooling\n\nCross-Functional Collaboration (15%)\n\nPartner with engineering teams to build reliability, scalability, and observability into new features from design through launch\nDocument runbooks, architecture decisions, and operational procedures for the broader engineering team\nCommunicate incident status and technical issues clearly to engineering and non-technical stakeholders\nParticipate in design reviews to flag reliability or operational concerns early\n\nSecurity & Compliance (10%)\n\nSupport SOC 2 compliance activities and help uphold encryption, access-control, and audit-trail standards across all environments\nImplement and maintain security best practices for infrastructure handling confidential legal and client data, including AI workloads on Amazon Bedrock\nSupport security reviews, vulnerability management, and patching cadences across production systems\nMaintain permissions-based access controls and comprehensive audit logging in line with client security commitments\n\nQUALIFICATIONS\n\n5+ years of experience in Site Reliability Engineering, DevOps, or Infrastructure Engineering roles, ideally in a B2B SaaS environment\nDeep hands-on experience with AWS (EC2, S3, RDS, Lambda, VPC, IAM, CloudWatch, or equivalent services)\nWorking knowledge of infrastructure-as-code (Terraform, CloudFormation, or CDK) and configuration management\nProficiency in at least one scripting/programming language (Python, Go, or similar) for automation and tooling\nExperience with containerization and orchestration (Docker, Kubernetes, or ECS)\nTrack record of operating CI/CD pipelines\nExperience with monitoring/observability stacks (Datadog, CloudWatch, Prometheus/Grafana, or similar)\nFamiliarity with security compliance frameworks (SOC 2, HIPAA, or similar) and encryption/access-control best practices\nExperience supporting systems handling large-scale, variable-volume data processing is a plus\nExposure to AI/ML infrastructure (e.g., Amazon Bedrock, model-serving pipelines) is a plus, given our growing AI feature set\nExperience using ClaudeCode, Kiro, OpenCode or similar Agentic AI\nBachelor's degree in Computer Science, Engineering, or related field (or equivalent experience)\nStrong written and verbal communication skills, with the ability to write clear runbooks and explain technical tradeoffs to non-technical stakeholders\nComfortable being part of an on-call rotation\n\nEQUAL OPPORTUNITY EMPLOYER\nNextpoint is an equal opportunity employer. We actively work to build a diverse team and encourage candidates of all backgrounds to apply. All applicants are considered without regard to race, color, religion, sex, sexual orientation, gender identity, national origin, veteran status, disability, or any other characteristic protected by applicable law. \n#J-18808-Ljbffr","datePosted":"2026-10-03T03:44:59.436Z","dateModified":"2026-10-03T03:44:59.436Z","hiringOrganization":{"@type":"Organization","name":"Nextpoint","sameAs":"https://jobsearcher.com"},"jobLocation":{"@type":"Place","address":{"@type":"PostalAddress","addressLocality":"Eastern","addressRegion":"KY","addressCountry":"US"}},"identifier":{"@type":"PropertyValue","name":"JobSearcher","value":"0891c55da5420b2a0c374788"},"url":"https://jobsearcher.com/jobs/0891c55da5420b2a0c374788"}}