{"schemaVersion":"jobsearcher.job.v1","id":"1f7f522b6397e9fe9debdb85","url":"https://jobsearcher.com/jobs/1f7f522b6397e9fe9debdb85","canonicalUrl":"https://jobsearcher.com/jobs/1f7f522b6397e9fe9debdb85","title":"Software Engineer, Distributed Systems","description":"Location\n\nRemote - Global\n\nEmployment Type\n\nFull time\n\nDepartment\n\nEngineeringPlatform\nfal is the generative media ecosystem powering the next generation of AI products. We build the infrastructure, tools, and model access that teams need to move from idea to production, and do it at scale without compromise. For developers and enterprises, fal is the foundation that makes generative media not just possible, but practical: a unified platform where high-performance inference, orchestration, and observability come together to unlock new categories of AI-native products.\n\nAs generative media reshapes industries across a market projected to grow by hundreds of billions over the next decade, fal is becoming the ecosystem that ambitious teams build on.\n\nAbout this role:\n\nYou are an experienced software engineer who thrives on building large-scale computing platforms. You have deep expertise in large scale distributed systems that deal with high complexity, a lot of traffic and data. You know how to achieve reliability and scale with minimum operational load.\n\nWhat you'll do:\n\nBuild our core Python/Rust platform: request routing, AI workload orchestration, scheduling, GPU autoscaling, large scale file storage, queueing, etc\n\nProduce forward designs for platform evolution as we scale to 100x current traffic and need to provide low latency across the world\n\nLeverage AI to an extreme level to automate the mundane parts of building complex but reliable systems\n\nProfile and tune low level CPU and memory performance\n\nQualifications/Nice to have:\n\n5+ years experience building distributed compute and orchestration platforms in Python or Rust\n\nStrong understanding of distributed systems fundamentals: consensus, scheduling, fault tolerance, capacity planning\n\nDeep understanding of computational complexity and memory allocation\n\nTrack record of designing systems that scale under real production load\n\nExperience building and using observability to drive performance and reliability decisions\n\nExcellent communication and ability to drive technical decisions across teams\n\nSelf-starter who executes quickly, takes ownership, and constantly seeks improvement\n\nExperience with AI/ML inference or training infrastructure\n\nExperience with high-performance systems programming (async runtimes, zero-copy, memory-safe concurrency)\n\nBackground in building multi-tenant compute platforms\n\nUnderstanding of networking fundamentals and performance characteristics\n\nFamiliarity with GPU workload characteristics and scheduling constraints\n\nWhat we offer at fal\n\nInteresting and challenging work\n\nA lot of learning and growth opportunities\n\nRegular team events and offsites\n\nU.S. EQUAL EMPLOYMENT OPPORTUNITY INFORMATION:\n\nfal provides equal employment opportunities to applicants and employees without regard to race, color, religion, sex, sexual orientation, gender identity, national origin, protected veteran status, disability, or any other classification protected by applicable law.","company":"Fal Features Labels","rawCompany":"fal features labels","city":"Myrtle Point","state":"OR","isRemote":false,"isActive":false,"createdAt":"2026-09-28T11:48:15.767Z","occupations":[{"code":"15-1299.08","title":"Computer Systems Engineers/Architects","slug":"computer-systems-engineers-architects"},{"code":"15-1252.00","title":"Software Developers","slug":"software-developers"},{"code":"15-1299.00","title":"Computer Occupations, All Other","slug":"computer-occupations-all-other"}],"industries":[{"code":"541512","title":"Computer Systems Design Services","slug":"computer-systems-design-services"},{"code":"513210","title":"Software Publishers","slug":"software-publishers"},{"code":"541511","title":"Custom Computer Programming Services","slug":"custom-computer-programming-services"}],"jobPosting":{"@context":"https://schema.org","@type":"JobPosting","title":"Software Engineer, Distributed Systems","description":"Location\n\nRemote - Global\n\nEmployment Type\n\nFull time\n\nDepartment\n\nEngineeringPlatform\nfal is the generative media ecosystem powering the next generation of AI products. We build the infrastructure, tools, and model access that teams need to move from idea to production, and do it at scale without compromise. For developers and enterprises, fal is the foundation that makes generative media not just possible, but practical: a unified platform where high-performance inference, orchestration, and observability come together to unlock new categories of AI-native products.\n\nAs generative media reshapes industries across a market projected to grow by hundreds of billions over the next decade, fal is becoming the ecosystem that ambitious teams build on.\n\nAbout this role:\n\nYou are an experienced software engineer who thrives on building large-scale computing platforms. You have deep expertise in large scale distributed systems that deal with high complexity, a lot of traffic and data. You know how to achieve reliability and scale with minimum operational load.\n\nWhat you'll do:\n\nBuild our core Python/Rust platform: request routing, AI workload orchestration, scheduling, GPU autoscaling, large scale file storage, queueing, etc\n\nProduce forward designs for platform evolution as we scale to 100x current traffic and need to provide low latency across the world\n\nLeverage AI to an extreme level to automate the mundane parts of building complex but reliable systems\n\nProfile and tune low level CPU and memory performance\n\nQualifications/Nice to have:\n\n5+ years experience building distributed compute and orchestration platforms in Python or Rust\n\nStrong understanding of distributed systems fundamentals: consensus, scheduling, fault tolerance, capacity planning\n\nDeep understanding of computational complexity and memory allocation\n\nTrack record of designing systems that scale under real production load\n\nExperience building and using observability to drive performance and reliability decisions\n\nExcellent communication and ability to drive technical decisions across teams\n\nSelf-starter who executes quickly, takes ownership, and constantly seeks improvement\n\nExperience with AI/ML inference or training infrastructure\n\nExperience with high-performance systems programming (async runtimes, zero-copy, memory-safe concurrency)\n\nBackground in building multi-tenant compute platforms\n\nUnderstanding of networking fundamentals and performance characteristics\n\nFamiliarity with GPU workload characteristics and scheduling constraints\n\nWhat we offer at fal\n\nInteresting and challenging work\n\nA lot of learning and growth opportunities\n\nRegular team events and offsites\n\nU.S. EQUAL EMPLOYMENT OPPORTUNITY INFORMATION:\n\nfal provides equal employment opportunities to applicants and employees without regard to race, color, religion, sex, sexual orientation, gender identity, national origin, protected veteran status, disability, or any other classification protected by applicable law.","datePosted":"2026-09-28T11:48:15.767Z","dateModified":"2026-09-28T11:48:15.767Z","hiringOrganization":{"@type":"Organization","name":"Fal Features Labels","sameAs":"https://jobsearcher.com"},"jobLocation":{"@type":"Place","address":{"@type":"PostalAddress","addressLocality":"Myrtle Point","addressRegion":"OR","addressCountry":"US"}},"identifier":{"@type":"PropertyValue","name":"JobSearcher","value":"1f7f522b6397e9fe9debdb85"},"url":"https://jobsearcher.com/jobs/1f7f522b6397e9fe9debdb85"}}