{"schemaVersion":"jobsearcher.job.v1","id":"32565db5719fc85fcaaaf3f4","url":"https://jobsearcher.com/jobs/32565db5719fc85fcaaaf3f4","canonicalUrl":"https://jobsearcher.com/jobs/32565db5719fc85fcaaaf3f4","title":"Machine Learning Engineer","description":"Overview Our client is a Seattle AI research lab with $60M raised, building photorealistic, real-time AI avatars that actually understand emotion. Their 19-person team holds PhDs from MIT, UW, Oxford, CMU and Johns Hopkins, with industry experience from Apple, Meta, Amazon AGI and Discord. They are training foundation models from the ground up for full-duplex audiovisual conversation: a system that listens, speaks, reacts and interrupts like a real person. Small team, unsolved problems, real work.\r\nThe role Today's conversational avatars take 2 to 5 seconds to respond. Natural conversation needs sub-500ms. Closing that 10x gap means rethinking the entire serving stack, and this role owns it.\r\nResponsibilities Own and build the serving stack for our multimodal AI workloads, optimizing for latency, throughput and cost\r\nArchitect and manage systems for real-time, long-lived WebRTC connections that keep video and audio smooth\r\nBuild and orchestrate robust data pipelines for large-scale offline processing, evaluation and training using frameworks like Dagster or Ray\r\nConfigure, maintain and optimize our GPU clusters with Kubernetes and Terraform\r\nDevelop CI/CD, evaluation and versioning systems that enable safe, zero-downtime model deployments and rapid iteration\r\nWork closely with ML researchers and product engineers to build the foundational infrastructure powering visual conversational AI\r\nQualifications 2+ years of full-time experience building and maintaining production-level ML systems\r\n0 to 1 experience building ML infra at a VC-backed startup (Series C or earlier), on LLM inference systems, or on multimodal systems for video, audio or multimedia models\r\nExtensive experience building data pipelines as distributed systems\r\nProficiency in Python and either Rust or Go\r\nStrong practical experience with Kubernetes, Terraform and cloud platforms\r\nA track record of optimizing systems for latency, throughput and cost, plus the ability to own complex projects and debug distributed systems\r\nReal-time video or audio streaming experience (WebRTC, low-latency infrastructure)\r\nBreadth across inference infrastructure, streaming and data engineering rather than depth in one narrow stack\r\nCompensation and benefits $250,000 to $450,000 base, depending on seniority and experience\r\nRelocation assistance is available\r\nVisa sponsorship is available, including new H-1B applications\r\nLocation On-site in Seattle, WA, 5 days per week. Full-time. Two openings.\r\nJ-18808-Ljbffr","company":"Raydar","rawCompany":"raydar","city":"Seattle","state":"WA","isRemote":false,"isActive":false,"createdAt":"2026-08-08T01:42:06.253Z","occupations":[{"code":"15-1299.08","title":"Computer Systems Engineers/Architects","slug":"computer-systems-engineers-architects"},{"code":"15-1252.00","title":"Software Developers","slug":"software-developers"},{"code":"15-1221.00","title":"Computer and Information Research Scientists","slug":"computer-and-information-research-scientists"}],"industries":[{"code":"541512","title":"Computer Systems Design Services","slug":"computer-systems-design-services"},{"code":"541511","title":"Custom Computer Programming Services","slug":"custom-computer-programming-services"},{"code":"513210","title":"Software Publishers","slug":"software-publishers"}],"jobPosting":{"@context":"https://schema.org","@type":"JobPosting","title":"Machine Learning Engineer","description":"Overview Our client is a Seattle AI research lab with $60M raised, building photorealistic, real-time AI avatars that actually understand emotion. Their 19-person team holds PhDs from MIT, UW, Oxford, CMU and Johns Hopkins, with industry experience from Apple, Meta, Amazon AGI and Discord. They are training foundation models from the ground up for full-duplex audiovisual conversation: a system that listens, speaks, reacts and interrupts like a real person. Small team, unsolved problems, real work.\r\nThe role Today's conversational avatars take 2 to 5 seconds to respond. Natural conversation needs sub-500ms. Closing that 10x gap means rethinking the entire serving stack, and this role owns it.\r\nResponsibilities Own and build the serving stack for our multimodal AI workloads, optimizing for latency, throughput and cost\r\nArchitect and manage systems for real-time, long-lived WebRTC connections that keep video and audio smooth\r\nBuild and orchestrate robust data pipelines for large-scale offline processing, evaluation and training using frameworks like Dagster or Ray\r\nConfigure, maintain and optimize our GPU clusters with Kubernetes and Terraform\r\nDevelop CI/CD, evaluation and versioning systems that enable safe, zero-downtime model deployments and rapid iteration\r\nWork closely with ML researchers and product engineers to build the foundational infrastructure powering visual conversational AI\r\nQualifications 2+ years of full-time experience building and maintaining production-level ML systems\r\n0 to 1 experience building ML infra at a VC-backed startup (Series C or earlier), on LLM inference systems, or on multimodal systems for video, audio or multimedia models\r\nExtensive experience building data pipelines as distributed systems\r\nProficiency in Python and either Rust or Go\r\nStrong practical experience with Kubernetes, Terraform and cloud platforms\r\nA track record of optimizing systems for latency, throughput and cost, plus the ability to own complex projects and debug distributed systems\r\nReal-time video or audio streaming experience (WebRTC, low-latency infrastructure)\r\nBreadth across inference infrastructure, streaming and data engineering rather than depth in one narrow stack\r\nCompensation and benefits $250,000 to $450,000 base, depending on seniority and experience\r\nRelocation assistance is available\r\nVisa sponsorship is available, including new H-1B applications\r\nLocation On-site in Seattle, WA, 5 days per week. Full-time. Two openings.\r\nJ-18808-Ljbffr","datePosted":"2026-08-08T01:42:06.253Z","dateModified":"2026-08-08T01:42:06.253Z","hiringOrganization":{"@type":"Organization","name":"Raydar","sameAs":"https://jobsearcher.com"},"jobLocation":{"@type":"Place","address":{"@type":"PostalAddress","addressLocality":"Seattle","addressRegion":"WA","addressCountry":"US"}},"identifier":{"@type":"PropertyValue","name":"JobSearcher","value":"32565db5719fc85fcaaaf3f4"},"url":"https://jobsearcher.com/jobs/32565db5719fc85fcaaaf3f4"}}