{"schemaVersion":"jobsearcher.job.v1","id":"188de8ae8b48d669444f7247","url":"https://jobsearcher.com/jobs/188de8ae8b48d669444f7247","canonicalUrl":"https://jobsearcher.com/jobs/188de8ae8b48d669444f7247","title":"Staff Software Engineer - Serverless","description":"We’re looking for a Staff/Senior Software Engineer to technically lead and build Runware’s serverless platform. You’ll be part of a small team working closely with senior leadership to bring a new product line to market.\n\nRunware is evolving beyond model APIs into a platform where developers can deploy models and run AI workloads without managing GPUs, servers or infrastructure. You’ll build the software layer that connects our proprietary high-performance sonic inference engine and turns complex infrastructure into a compelling developer experience.\n\nThis is a role for someone who enjoys technical leadership, hard platform problems and clean, simple systems. Your work will help enterprises, model labs and developers move from idea to production faster, then scale without needing to build and operate the infrastructure themselves.\n\nWhat you’ll do\nBuild the core systems behind Runware’s serverless platform, including workload execution, routing, scheduling, isolation and scaling leveraging our inference engine\nMake it simple for developers to deploy models and run AI workloads without managing GPUs, servers, queues or infrastructure through a simple SDK interface\nDesign and improve the control plane for serverless execution, including APIs, workers, lifecycle management, retries and failure handling\nWork closely with our infrastructure and ML teams to improve workload startup time, GPU utilisation, model warm-up, caching and placement\nBuild observability that makes serverless workloads easy to monitor, debug and operate at scale globally\nLead the technical design, mentor other engineers and help define the engineering standards for a new product area\n\nRequirements\n\nStrong experience as a Staff Engineer, Senior Software Engineer, Backend Engineer, Platform Engineer or similar\nExperience building backend services, distributed systems, developer platforms or workload orchestration systems\nStrong understanding of async processing, queues, scheduling, retries, back pressure and failure handling\nComfortable working across APIs, control planes, workers, databases and observability systems\nStrong engineering fundamentals in one or more backend languages ideally Go\nGood judgement around trade-offs between reliability, latency, scale, cost and developer experience\nClear communication, strong ownership and the ability to lead technical direction in a fast-moving environment\nNice to have\nExperience building serverless platforms, job execution systems, container platforms or compute orchestration systems\nExperience with GPU-backed workloads, AI/ML inference, model serving, batch processing or high-performance compute\nFamiliarity with technologies such as vLLM, TensorRT, Triton, Kubernetes, Nomad or Knative\nExperience improving workload performance through batching, autoscaling, model warm-up, caching, request routing or queue management\nExperience with multi-tenant isolation, sandboxing, quotas, rate limits, resource accounting or usage-based billing\n\nBenefits\n\nWe’re a remote-first collective, meeting in person twice a year to plan, brainstorm, celebrate wins, and enjoy some face-to-face time. We have core hours for cooperative working and calls, but outside of that your calendar is yours. Work the hours that let you perform at your peak while also building a healthy life.\n\nOur release cycles are fast and intense, but they’re followed by real downtime. After big pushes we expect the team to unplug, recharge, and come back ready & stronger than ever for the next leap.\n\nGenerous paid time off – vacation, sick days, public holidays\nMeaningful stock options – share in the upside you create\nRemote-first setup – work from home anywhere we can employ you\nFlexible hours – own your schedule outside core collaboration blocks\nFamily leave – paid maternity, paternity, and caregiver time\nCompany retreats – twice-yearly gatherings in inspiring locations","company":"Runware","rawCompany":"runware","city":"Denver","state":"CO","isRemote":false,"isActive":false,"createdAt":"2026-08-27T09:02:37.580Z","occupations":[{"code":"15-1252.00","title":"Software Developers","slug":"software-developers"},{"code":"15-1299.08","title":"Computer Systems Engineers/Architects","slug":"computer-systems-engineers-architects"},{"code":"15-1299.00","title":"Computer Occupations, All Other","slug":"computer-occupations-all-other"}],"industries":[{"code":"513210","title":"Software Publishers","slug":"software-publishers"},{"code":"541511","title":"Custom Computer Programming Services","slug":"custom-computer-programming-services"},{"code":"541512","title":"Computer Systems Design Services","slug":"computer-systems-design-services"}],"jobPosting":{"@context":"https://schema.org","@type":"JobPosting","title":"Staff Software Engineer - Serverless","description":"We’re looking for a Staff/Senior Software Engineer to technically lead and build Runware’s serverless platform. You’ll be part of a small team working closely with senior leadership to bring a new product line to market.\n\nRunware is evolving beyond model APIs into a platform where developers can deploy models and run AI workloads without managing GPUs, servers or infrastructure. You’ll build the software layer that connects our proprietary high-performance sonic inference engine and turns complex infrastructure into a compelling developer experience.\n\nThis is a role for someone who enjoys technical leadership, hard platform problems and clean, simple systems. Your work will help enterprises, model labs and developers move from idea to production faster, then scale without needing to build and operate the infrastructure themselves.\n\nWhat you’ll do\nBuild the core systems behind Runware’s serverless platform, including workload execution, routing, scheduling, isolation and scaling leveraging our inference engine\nMake it simple for developers to deploy models and run AI workloads without managing GPUs, servers, queues or infrastructure through a simple SDK interface\nDesign and improve the control plane for serverless execution, including APIs, workers, lifecycle management, retries and failure handling\nWork closely with our infrastructure and ML teams to improve workload startup time, GPU utilisation, model warm-up, caching and placement\nBuild observability that makes serverless workloads easy to monitor, debug and operate at scale globally\nLead the technical design, mentor other engineers and help define the engineering standards for a new product area\n\nRequirements\n\nStrong experience as a Staff Engineer, Senior Software Engineer, Backend Engineer, Platform Engineer or similar\nExperience building backend services, distributed systems, developer platforms or workload orchestration systems\nStrong understanding of async processing, queues, scheduling, retries, back pressure and failure handling\nComfortable working across APIs, control planes, workers, databases and observability systems\nStrong engineering fundamentals in one or more backend languages ideally Go\nGood judgement around trade-offs between reliability, latency, scale, cost and developer experience\nClear communication, strong ownership and the ability to lead technical direction in a fast-moving environment\nNice to have\nExperience building serverless platforms, job execution systems, container platforms or compute orchestration systems\nExperience with GPU-backed workloads, AI/ML inference, model serving, batch processing or high-performance compute\nFamiliarity with technologies such as vLLM, TensorRT, Triton, Kubernetes, Nomad or Knative\nExperience improving workload performance through batching, autoscaling, model warm-up, caching, request routing or queue management\nExperience with multi-tenant isolation, sandboxing, quotas, rate limits, resource accounting or usage-based billing\n\nBenefits\n\nWe’re a remote-first collective, meeting in person twice a year to plan, brainstorm, celebrate wins, and enjoy some face-to-face time. We have core hours for cooperative working and calls, but outside of that your calendar is yours. Work the hours that let you perform at your peak while also building a healthy life.\n\nOur release cycles are fast and intense, but they’re followed by real downtime. After big pushes we expect the team to unplug, recharge, and come back ready & stronger than ever for the next leap.\n\nGenerous paid time off – vacation, sick days, public holidays\nMeaningful stock options – share in the upside you create\nRemote-first setup – work from home anywhere we can employ you\nFlexible hours – own your schedule outside core collaboration blocks\nFamily leave – paid maternity, paternity, and caregiver time\nCompany retreats – twice-yearly gatherings in inspiring locations","datePosted":"2026-08-27T09:02:37.580Z","dateModified":"2026-08-27T09:02:37.580Z","hiringOrganization":{"@type":"Organization","name":"Runware","sameAs":"https://jobsearcher.com"},"jobLocation":{"@type":"Place","address":{"@type":"PostalAddress","addressLocality":"Denver","addressRegion":"CO","addressCountry":"US"}},"identifier":{"@type":"PropertyValue","name":"JobSearcher","value":"188de8ae8b48d669444f7247"},"url":"https://jobsearcher.com/jobs/188de8ae8b48d669444f7247"}}