{"schemaVersion":"jobsearcher.job.v1","id":"1f610e670c9ece30fe896ee9","url":"https://jobsearcher.com/jobs/1f610e670c9ece30fe896ee9","canonicalUrl":"https://jobsearcher.com/jobs/1f610e670c9ece30fe896ee9","title":"Research Evaluation Lead","description":"At Rhoda AI, we’re building the next generation of generalist intelligent robots. We own the full robotics stack from high-performance hardware and robot systems to the infrastructure and state-of-the-art foundation world models that control our robots. Our robots are designed to be generalists capable of operating in complex, real-world environments and handling long-tail edge cases, made possible by our cutting edge research and end-to-end system design. We've raised over $450M and are investing aggressively in model research, infrastructure, hardware development, and manufacturing scale-up to make generalist robotics a reality.\r\nMissionOwn end-to-end robot model evaluation for Research. Turn research questions into consistent, high-quality, repeatable evals that enable fast iteration and trusted results.\r\nResponsibilitiesTranslate research intent into clear eval protocols, trial plans, and success criteria.\r\nOwn execution end-to-end: model handoff → station readiness → pilot execution → QA → results.\r\nTrain and manage eval pilots; ensure consistency across people, shifts, and stations.\r\nDistinguish model failures from hardware, setup, operator, or data-quality issues.\r\nMaintain eval setups, resets, randomization, metadata, and experiment traceability.\r\nTrack quality, throughput, and bottlenecks; continuously improve the eval process.\r\nPartner closely with Research, Robot Data, and Eval Platform teams.\r\nWhat we’re looking forSome understanding of robotics / ML experimentation.\r\nComputer science background or hands-on experience with coding\r\nRigorous, detail-oriented, and able to understand the intent behind an experiment, not just execute instructions.\r\nStrong hands-on execution and ownership.\r\nExperience in robotics testing, data collection, lab operations, or QA preferred.\r\nSuccess looks likeA researcher can hand off a model and research question and receive a trusted, standardized eval result with sufficient trials and QA.#J-18808-Ljbffr","company":"Socketdev","rawCompany":"socketdev","city":"Mountain View","state":"CA","isRemote":false,"isActive":false,"createdAt":"2026-09-26T01:22:36.437Z","occupations":[{"code":"17-2199.08","title":"Robotics Engineers","slug":"robotics-engineers"},{"code":"15-1221.00","title":"Computer and Information Research Scientists","slug":"computer-and-information-research-scientists"},{"code":"17-3024.01","title":"Robotics Technicians","slug":"robotics-technicians"}],"industries":[{"code":"541715","title":"Research and Development in the Physical, Engineering, and Life Sciences (except Nanotechnology and Biotechnology)","slug":"research-and-development-in-the-physical-engineering-and-life-sciences-except-nanotechnology-and-biotechnology"},{"code":"541690","title":"Other Scientific and Technical Consulting Services","slug":"other-scientific-and-technical-consulting-services"},{"code":"541720","title":"Research and Development in the Social Sciences and Humanities","slug":"research-and-development-in-the-social-sciences-and-humanities"}],"jobPosting":{"@context":"https://schema.org","@type":"JobPosting","title":"Research Evaluation Lead","description":"At Rhoda AI, we’re building the next generation of generalist intelligent robots. We own the full robotics stack from high-performance hardware and robot systems to the infrastructure and state-of-the-art foundation world models that control our robots. Our robots are designed to be generalists capable of operating in complex, real-world environments and handling long-tail edge cases, made possible by our cutting edge research and end-to-end system design. We've raised over $450M and are investing aggressively in model research, infrastructure, hardware development, and manufacturing scale-up to make generalist robotics a reality.\r\nMissionOwn end-to-end robot model evaluation for Research. Turn research questions into consistent, high-quality, repeatable evals that enable fast iteration and trusted results.\r\nResponsibilitiesTranslate research intent into clear eval protocols, trial plans, and success criteria.\r\nOwn execution end-to-end: model handoff → station readiness → pilot execution → QA → results.\r\nTrain and manage eval pilots; ensure consistency across people, shifts, and stations.\r\nDistinguish model failures from hardware, setup, operator, or data-quality issues.\r\nMaintain eval setups, resets, randomization, metadata, and experiment traceability.\r\nTrack quality, throughput, and bottlenecks; continuously improve the eval process.\r\nPartner closely with Research, Robot Data, and Eval Platform teams.\r\nWhat we’re looking forSome understanding of robotics / ML experimentation.\r\nComputer science background or hands-on experience with coding\r\nRigorous, detail-oriented, and able to understand the intent behind an experiment, not just execute instructions.\r\nStrong hands-on execution and ownership.\r\nExperience in robotics testing, data collection, lab operations, or QA preferred.\r\nSuccess looks likeA researcher can hand off a model and research question and receive a trusted, standardized eval result with sufficient trials and QA.#J-18808-Ljbffr","datePosted":"2026-09-26T01:22:36.437Z","dateModified":"2026-09-26T01:22:36.437Z","hiringOrganization":{"@type":"Organization","name":"Socketdev","sameAs":"https://jobsearcher.com"},"jobLocation":{"@type":"Place","address":{"@type":"PostalAddress","addressLocality":"Mountain View","addressRegion":"CA","addressCountry":"US"}},"identifier":{"@type":"PropertyValue","name":"JobSearcher","value":"1f610e670c9ece30fe896ee9"},"url":"https://jobsearcher.com/jobs/1f610e670c9ece30fe896ee9"}}