{"schemaVersion":"jobsearcher.job.v1","id":"fd389eeabd30b8ebe83663d0","url":"https://jobsearcher.com/jobs/fd389eeabd30b8ebe83663d0","canonicalUrl":"https://jobsearcher.com/jobs/fd389eeabd30b8ebe83663d0","title":"Research Engineer","description":"Research Engineer\nYou'll build the evaluation systems that tell us whether Firecrawl actually works. That sounds simple. It isn't. Our core promise, convert any URL into clean, structured, LLM-ready data reliably, is hard to measure rigorously across millions of different websites, formats, and edge cases. As the systems we're measuring get more complex, the question \"did that work?\" gets harder, not easier.\nThis isn't an eval role where you inherit a framework and run benchmarks. You'll design the metrics, build the pipelines, generate the datasets, and own the feedback loop from output quality back to model and product decisions. If you care about what \"good\" actually means and have the engineering depth to measure it, this is the role.\nSalary Range: $210,000–$275,000/year (Range shown is for U.S.-based employees in San Francisco, CA. Compensation outside the U.S. is adjusted fairly based on your country's cost of living.)\nEquity Range: Competitive equity — details shared during the process.\nLocation: San Francisco, CA (Hybrid, on-site required)\nJob Type: Full-Time\nExperience: 4+ years in ML, research engineering, or data-heavy backend, with real evaluation work\nVisa: Must be legally authorized to work in the United States. We're not able to sponsor visas right now, though that may change down the line.\nAbout Firecrawl\nFirecrawl is the easiest way to turn the web into data AI agents can use. One API call converts any URL into clean, LLM-ready markdown or structured data - the boring-hard problem everyone building with LLMs eventually hits, solved.\nWe hit 8 figures in ARR in year one and more than doubled it in year two. We have 147k+ GitHub stars, and developers, agents, and category-defining AI companies build on us every day. Growth like this is rare, and we're just getting started.\nWe're a small team punching far above our weight. Everyone here owns a real piece of the product and company, end to end, and runs it themselves - no hiding behind process or headcount.\nThis is a place for people who want to work at the frontier: an AI company building the infrastructure other AI companies run on, not one bolting AI onto an existing product. We move fast, go deep, and are building the tools superintelligence will rely on to gather data from the web.\nWhat You'll Do\nDesign the metrics that define what \"good output\" actually means across millions of sites, formats, and edge cases\nBuild the pipelines and harnesses that measure quality rigorously and at scale\nGenerate and curate the datasets that make evaluation trustworthy\nOwn the feedback loop from output quality back to model and product decisions\nTurn \"did that work?\" into an answer the whole team can act on\nWhat We're Looking For\nYou have the engineering depth to build real evaluation systems, not just run existing ones\nYou care deeply about what \"good\" means and how to measure it rigorously\nYou're comfortable owning ambiguous problems where the metric itself has to be invented\nYou move fast and close the loop - you'd rather ship, measure, and iterate than perfect on paper\nWhat We're NOT Looking For\nSomeone who only wants to run benchmarks someone else designed\nA pure researcher who won't build the systems, or a pure engineer who won't think about methodology\nSomeone who needs a fully-specced ticket to start\nA Note On Pace\nWe operate at an absurd level of urgency because the window for what we're building won't stay open forever. If that excites you, keep reading. If it doesn't, no hard feelings — but this role probably isn't for you.\nBenefits & Perks\nAvailable to all employees\nSalary that makes sense — $210,000-$275,000/year (U.S.-based), based on impact, not tenure\nOwn a piece — Gain competitive equity in what you're helping build\nGenerous PTO — 15 days mandatory, anything after 24 days, just ask (holidays excluded); take the time you need to recharge\nParental leave — 12 weeks fully paid, for all parents\nWellness stipend — $100/month for the gym, therapy, massages, or whatever keeps you human\nLearning & Development — Expense up to $1,000/year toward anything that helps you grow professionally\nTeam offsites — A change of scenery, minus the trust falls\nSabbatical — 3 paid months off after 4 years, do something fun and new\nAvailable to US-based full-time employees\nFull coverage, no red tape — Medical, dental, and vision (100% for employees, 50% for spouse/kids) — no weird loopholes, just care that works\nLife & Disability insurance — Employer-paid short-term disability, long-term disability, and life insurance — coverage for life's curveballs\nSupplemental options — Optional accident, critical illness, hospital indemnity, and voluntary life insurance for extra peace of mind\nDoctegrity telehealth — Talk to a doctor from your couch\n401(k) plan — Retirement might be a ways off, but future-you will thank you\nPre-tax benefits — Access to FSAs and commuter benefits to help your wallet out a bit\nPet insurance — Because fur babies are family too\nAvailable to SF-based employees\nSF HQ perks — Snacks, drinks, team lunches, intense ping pong, and peak startup energy\nE-Bike transportation — A loaner electric bike to get you around the city, on us\nInterview Process\nApplication Review — Send us your work and a quick note on why this excites you. Show us what you've built — eval systems, metrics you designed, datasets you created, quality problems you measured. We care about what you've shipped, not where you went to school.\nIntro Chat (~25 min) — A quick conversation to get to know each other before we go deep. We'll talk about what you've been working on, what drew you to Firecrawl, and what you're looking for in your next role. Time for your questions too.\nTechnical Chat (~45 min) — We'll dig into a real problem from our world — how you'd measure whether messy web output is actually \"good,\" and build the system to prove it. Come ready to think out loud; we care how you reason, not whether you memorized the answer.\nFounder Chat (~25 min) — Culture, pace, ownership, and how you like to work. Time for your questions too.\nPaid Work Trial (1-2 weeks) — Work with the team on a real, scoped evaluation problem — paid at a contractor rate. It's the truest signal for both sides: you see what building at Firecrawl actually feels like, and we see how you ship. Remote-friendly, and we'll flex around your current commitments.\nDecision — We move fast after the trial.\nIf you care about what \"good\" actually means and have the engineering depth to measure it, you should join us.\nApply now.\nCompensation Range: $210K - $275K","company":"Firecrawl","rawCompany":"firecrawl","city":"Millbrae","state":"CA","isRemote":false,"isActive":false,"createdAt":"2026-08-04T20:38:36.923Z","occupations":[{"code":"15-1221.00","title":"Computer and Information Research Scientists","slug":"computer-and-information-research-scientists"},{"code":"17-2199.00","title":"Engineers, All Other","slug":"engineers-all-other"},{"code":"15-1252.00","title":"Software Developers","slug":"software-developers"}],"industries":[{"code":"513210","title":"Software Publishers","slug":"software-publishers"},{"code":"541715","title":"Research and Development in the Physical, Engineering, and Life Sciences (except Nanotechnology and Biotechnology)","slug":"research-and-development-in-the-physical-engineering-and-life-sciences-except-nanotechnology-and-biotechnology"},{"code":"541511","title":"Custom Computer Programming Services","slug":"custom-computer-programming-services"}],"jobPosting":{"@context":"https://schema.org","@type":"JobPosting","title":"Research Engineer","description":"Research Engineer\nYou'll build the evaluation systems that tell us whether Firecrawl actually works. That sounds simple. It isn't. Our core promise, convert any URL into clean, structured, LLM-ready data reliably, is hard to measure rigorously across millions of different websites, formats, and edge cases. As the systems we're measuring get more complex, the question \"did that work?\" gets harder, not easier.\nThis isn't an eval role where you inherit a framework and run benchmarks. You'll design the metrics, build the pipelines, generate the datasets, and own the feedback loop from output quality back to model and product decisions. If you care about what \"good\" actually means and have the engineering depth to measure it, this is the role.\nSalary Range: $210,000–$275,000/year (Range shown is for U.S.-based employees in San Francisco, CA. Compensation outside the U.S. is adjusted fairly based on your country's cost of living.)\nEquity Range: Competitive equity — details shared during the process.\nLocation: San Francisco, CA (Hybrid, on-site required)\nJob Type: Full-Time\nExperience: 4+ years in ML, research engineering, or data-heavy backend, with real evaluation work\nVisa: Must be legally authorized to work in the United States. We're not able to sponsor visas right now, though that may change down the line.\nAbout Firecrawl\nFirecrawl is the easiest way to turn the web into data AI agents can use. One API call converts any URL into clean, LLM-ready markdown or structured data - the boring-hard problem everyone building with LLMs eventually hits, solved.\nWe hit 8 figures in ARR in year one and more than doubled it in year two. We have 147k+ GitHub stars, and developers, agents, and category-defining AI companies build on us every day. Growth like this is rare, and we're just getting started.\nWe're a small team punching far above our weight. Everyone here owns a real piece of the product and company, end to end, and runs it themselves - no hiding behind process or headcount.\nThis is a place for people who want to work at the frontier: an AI company building the infrastructure other AI companies run on, not one bolting AI onto an existing product. We move fast, go deep, and are building the tools superintelligence will rely on to gather data from the web.\nWhat You'll Do\nDesign the metrics that define what \"good output\" actually means across millions of sites, formats, and edge cases\nBuild the pipelines and harnesses that measure quality rigorously and at scale\nGenerate and curate the datasets that make evaluation trustworthy\nOwn the feedback loop from output quality back to model and product decisions\nTurn \"did that work?\" into an answer the whole team can act on\nWhat We're Looking For\nYou have the engineering depth to build real evaluation systems, not just run existing ones\nYou care deeply about what \"good\" means and how to measure it rigorously\nYou're comfortable owning ambiguous problems where the metric itself has to be invented\nYou move fast and close the loop - you'd rather ship, measure, and iterate than perfect on paper\nWhat We're NOT Looking For\nSomeone who only wants to run benchmarks someone else designed\nA pure researcher who won't build the systems, or a pure engineer who won't think about methodology\nSomeone who needs a fully-specced ticket to start\nA Note On Pace\nWe operate at an absurd level of urgency because the window for what we're building won't stay open forever. If that excites you, keep reading. If it doesn't, no hard feelings — but this role probably isn't for you.\nBenefits & Perks\nAvailable to all employees\nSalary that makes sense — $210,000-$275,000/year (U.S.-based), based on impact, not tenure\nOwn a piece — Gain competitive equity in what you're helping build\nGenerous PTO — 15 days mandatory, anything after 24 days, just ask (holidays excluded); take the time you need to recharge\nParental leave — 12 weeks fully paid, for all parents\nWellness stipend — $100/month for the gym, therapy, massages, or whatever keeps you human\nLearning & Development — Expense up to $1,000/year toward anything that helps you grow professionally\nTeam offsites — A change of scenery, minus the trust falls\nSabbatical — 3 paid months off after 4 years, do something fun and new\nAvailable to US-based full-time employees\nFull coverage, no red tape — Medical, dental, and vision (100% for employees, 50% for spouse/kids) — no weird loopholes, just care that works\nLife & Disability insurance — Employer-paid short-term disability, long-term disability, and life insurance — coverage for life's curveballs\nSupplemental options — Optional accident, critical illness, hospital indemnity, and voluntary life insurance for extra peace of mind\nDoctegrity telehealth — Talk to a doctor from your couch\n401(k) plan — Retirement might be a ways off, but future-you will thank you\nPre-tax benefits — Access to FSAs and commuter benefits to help your wallet out a bit\nPet insurance — Because fur babies are family too\nAvailable to SF-based employees\nSF HQ perks — Snacks, drinks, team lunches, intense ping pong, and peak startup energy\nE-Bike transportation — A loaner electric bike to get you around the city, on us\nInterview Process\nApplication Review — Send us your work and a quick note on why this excites you. Show us what you've built — eval systems, metrics you designed, datasets you created, quality problems you measured. We care about what you've shipped, not where you went to school.\nIntro Chat (~25 min) — A quick conversation to get to know each other before we go deep. We'll talk about what you've been working on, what drew you to Firecrawl, and what you're looking for in your next role. Time for your questions too.\nTechnical Chat (~45 min) — We'll dig into a real problem from our world — how you'd measure whether messy web output is actually \"good,\" and build the system to prove it. Come ready to think out loud; we care how you reason, not whether you memorized the answer.\nFounder Chat (~25 min) — Culture, pace, ownership, and how you like to work. Time for your questions too.\nPaid Work Trial (1-2 weeks) — Work with the team on a real, scoped evaluation problem — paid at a contractor rate. It's the truest signal for both sides: you see what building at Firecrawl actually feels like, and we see how you ship. Remote-friendly, and we'll flex around your current commitments.\nDecision — We move fast after the trial.\nIf you care about what \"good\" actually means and have the engineering depth to measure it, you should join us.\nApply now.\nCompensation Range: $210K - $275K","datePosted":"2026-08-04T20:38:36.923Z","dateModified":"2026-08-04T20:38:36.923Z","hiringOrganization":{"@type":"Organization","name":"Firecrawl","sameAs":"https://jobsearcher.com"},"jobLocation":{"@type":"Place","address":{"@type":"PostalAddress","addressLocality":"Millbrae","addressRegion":"CA","addressCountry":"US"}},"identifier":{"@type":"PropertyValue","name":"JobSearcher","value":"fd389eeabd30b8ebe83663d0"},"url":"https://jobsearcher.com/jobs/fd389eeabd30b8ebe83663d0"}}