{"schemaVersion":"jobsearcher.job.v1","id":"f8c38eedccf1f90e3dbd1d44","url":"https://jobsearcher.com/jobs/f8c38eedccf1f90e3dbd1d44","canonicalUrl":"https://jobsearcher.com/jobs/f8c38eedccf1f90e3dbd1d44","title":"Software Engineer (C++ Systems)","description":"CompanyThunder Compute is building the VMware for GPUs. We have raised over $17M from Matrix Partners, Y Combinator, and leading angels from Coreweave, Microsoft, Cognition, and Anthropic.Deployed GPU fleets are currently only 5–20% utilized. Leading solutions for underutilization sit at the workload layer and are therefore only able to optimize specific use cases. We believe the ideal cluster optimization solution must be invisible to developers and compatible with all workloads; hence, it must sit at the systems layer.We are a team of systems researchers productionizing cutting-edge GPU virtualization research to build this general-purpose optimization layer.Concretely, our virtualization library abstracts GPUs across TCP networking. We use a userspace shim library, loaded through LD_PRELOAD, to intercept CUDA calls and send them over gRPC to a host server connected to a physical GPU elsewhere in the data center.This enables something like “Ceph for GPUs”: GPUs become network resources that can be abstracted, pooled, and dynamically allocated across a cluster to improve utilization without requiring developers to modify their workloads.RoleYour work will focus on building the core C++ systems behind our virtualization layer. This includes low-latency performance optimization, distributed systems debugging, production reliability, and research into new techniques for improving GPU utilization.You will take ownership of complex systems from early experimentation through production deployment. Example projects may include:Profiling and reducing latency across remote CUDA operationsBuilding high-performance networking and data-transfer pathsDebugging failures across customer processes, our userspace runtime, the network, and remote GPU serversImproving support for process forking, signals, multithreading, dynamic linking, and unusual application behaviorDesigning systems for GPU allocation, scheduling, failure recovery, and observabilityResearching and productionizing new GPU virtualization and oversubscription techniquesExpanding compatibility across CUDA applications, frameworks, and GPU architecturesYou will spend your days bouncing between the weeds of complex, performance-critical systems that are live in production. One week, you may be tracing a synchronization bug across a distributed CUDA workload; the next, you may be redesigning a hot data path to remove microseconds of overhead.This work is not easy. It blends the hardest parts of systems research and production engineering.We look for exceptional low-level engineering talent, strong work ethic, and extreme attention to detail. We must move quickly while shipping high-quality, reliable systems code.Core Technical SkillsExceptional modern C++ ability, including memory management, concurrency, performance optimization, and systems-level abstraction designDeep understanding of operating systems, low-level networking, compilers, distributed systems, or computer architectureExperience building and operating performance-critical C++ systems in productionStrong Linux systems programming and debugging abilityAbility to reason through unfamiliar systems across multiple layers of the stackMust HavesStrong work ethic and the ability to independently push a project from an experimental prototype through 100% completion under tight deadlinesAttention to detail and the ability to deliver production-ready, thoroughly tested code without significant oversightStrong ownership over correctness, reliability, performance, and operational outcomesAbility to debug ambiguous problems without a clear reproduction, existing playbook, or obvious ownerWillingness to work directly with customers and investigate difficult production failuresPreferredExperience with CUDA, GPU systems, compilers, runtime interception, dynamic linking, high-performance networking, or distributed computingExperience at a trading firm such as Citadel Securities or Jane Street; a hardware or AI infrastructure company such as NVIDIA or SambaNova; a systems research group; or a similarly demanding engineering environmentStrong computer science fundamentals demonstrated through academic work, systems research, competitive programming, open-source contributions, or exceptional professional experienceExperience taking new systems research from a paper or prototype into a reliable production systemWhy JoinYou will join early enough to meaningfully shape the architecture, engineering standards, and technical direction of the company.You will work directly with the founders on a category-defining systems problem, with a short path between writing code and seeing it run in production. The systems you build will form the foundation of a new infrastructure layer for GPU computing.LogisticsYou will report to co-founder and CTO Brian Model, formerly a Quantitative Developer at Citadel SecuritiesThis role is full-time and in person, five days per week, at our office in downtown San FranciscoRelocation support and visa sponsorship are availableBenefitsCompetitive salary and meaningful equityDaily lunch, snacks, and coffeeTeam dinners and events401(k)Health, dental, and vision insuranceCompensation Range: $200K - $300K","company":"Thunder Compute","rawCompany":"thunder compute","city":"Millbrae","state":"CA","isRemote":false,"isActive":false,"createdAt":"2026-04-30T01:12:23.125Z","occupations":[{"code":"15-1299.08","title":"Computer Systems Engineers/Architects","slug":"computer-systems-engineers-architects"},{"code":"15-1252.00","title":"Software Developers","slug":"software-developers"},{"code":"15-1299.00","title":"Computer Occupations, All Other","slug":"computer-occupations-all-other"}],"industries":[{"code":"513210","title":"Software Publishers","slug":"software-publishers"},{"code":"541511","title":"Custom Computer Programming Services","slug":"custom-computer-programming-services"},{"code":"541512","title":"Computer Systems Design Services","slug":"computer-systems-design-services"}],"jobPosting":{"@context":"https://schema.org","@type":"JobPosting","title":"Software Engineer (C++ Systems)","description":"CompanyThunder Compute is building the VMware for GPUs. We have raised over $17M from Matrix Partners, Y Combinator, and leading angels from Coreweave, Microsoft, Cognition, and Anthropic.Deployed GPU fleets are currently only 5–20% utilized. Leading solutions for underutilization sit at the workload layer and are therefore only able to optimize specific use cases. We believe the ideal cluster optimization solution must be invisible to developers and compatible with all workloads; hence, it must sit at the systems layer.We are a team of systems researchers productionizing cutting-edge GPU virtualization research to build this general-purpose optimization layer.Concretely, our virtualization library abstracts GPUs across TCP networking. We use a userspace shim library, loaded through LD_PRELOAD, to intercept CUDA calls and send them over gRPC to a host server connected to a physical GPU elsewhere in the data center.This enables something like “Ceph for GPUs”: GPUs become network resources that can be abstracted, pooled, and dynamically allocated across a cluster to improve utilization without requiring developers to modify their workloads.RoleYour work will focus on building the core C++ systems behind our virtualization layer. This includes low-latency performance optimization, distributed systems debugging, production reliability, and research into new techniques for improving GPU utilization.You will take ownership of complex systems from early experimentation through production deployment. Example projects may include:Profiling and reducing latency across remote CUDA operationsBuilding high-performance networking and data-transfer pathsDebugging failures across customer processes, our userspace runtime, the network, and remote GPU serversImproving support for process forking, signals, multithreading, dynamic linking, and unusual application behaviorDesigning systems for GPU allocation, scheduling, failure recovery, and observabilityResearching and productionizing new GPU virtualization and oversubscription techniquesExpanding compatibility across CUDA applications, frameworks, and GPU architecturesYou will spend your days bouncing between the weeds of complex, performance-critical systems that are live in production. One week, you may be tracing a synchronization bug across a distributed CUDA workload; the next, you may be redesigning a hot data path to remove microseconds of overhead.This work is not easy. It blends the hardest parts of systems research and production engineering.We look for exceptional low-level engineering talent, strong work ethic, and extreme attention to detail. We must move quickly while shipping high-quality, reliable systems code.Core Technical SkillsExceptional modern C++ ability, including memory management, concurrency, performance optimization, and systems-level abstraction designDeep understanding of operating systems, low-level networking, compilers, distributed systems, or computer architectureExperience building and operating performance-critical C++ systems in productionStrong Linux systems programming and debugging abilityAbility to reason through unfamiliar systems across multiple layers of the stackMust HavesStrong work ethic and the ability to independently push a project from an experimental prototype through 100% completion under tight deadlinesAttention to detail and the ability to deliver production-ready, thoroughly tested code without significant oversightStrong ownership over correctness, reliability, performance, and operational outcomesAbility to debug ambiguous problems without a clear reproduction, existing playbook, or obvious ownerWillingness to work directly with customers and investigate difficult production failuresPreferredExperience with CUDA, GPU systems, compilers, runtime interception, dynamic linking, high-performance networking, or distributed computingExperience at a trading firm such as Citadel Securities or Jane Street; a hardware or AI infrastructure company such as NVIDIA or SambaNova; a systems research group; or a similarly demanding engineering environmentStrong computer science fundamentals demonstrated through academic work, systems research, competitive programming, open-source contributions, or exceptional professional experienceExperience taking new systems research from a paper or prototype into a reliable production systemWhy JoinYou will join early enough to meaningfully shape the architecture, engineering standards, and technical direction of the company.You will work directly with the founders on a category-defining systems problem, with a short path between writing code and seeing it run in production. The systems you build will form the foundation of a new infrastructure layer for GPU computing.LogisticsYou will report to co-founder and CTO Brian Model, formerly a Quantitative Developer at Citadel SecuritiesThis role is full-time and in person, five days per week, at our office in downtown San FranciscoRelocation support and visa sponsorship are availableBenefitsCompetitive salary and meaningful equityDaily lunch, snacks, and coffeeTeam dinners and events401(k)Health, dental, and vision insuranceCompensation Range: $200K - $300K","datePosted":"2026-04-30T01:12:23.125Z","dateModified":"2026-04-30T01:12:23.125Z","hiringOrganization":{"@type":"Organization","name":"Thunder Compute","sameAs":"https://jobsearcher.com"},"jobLocation":{"@type":"Place","address":{"@type":"PostalAddress","addressLocality":"Millbrae","addressRegion":"CA","addressCountry":"US"}},"identifier":{"@type":"PropertyValue","name":"JobSearcher","value":"f8c38eedccf1f90e3dbd1d44"},"url":"https://jobsearcher.com/jobs/f8c38eedccf1f90e3dbd1d44"}}