{"schemaVersion":"jobsearcher.job.v1","id":"76ea8a8c20a2a99f0a52980b","url":"https://jobsearcher.com/jobs/76ea8a8c20a2a99f0a52980b","canonicalUrl":"https://jobsearcher.com/jobs/76ea8a8c20a2a99f0a52980b","title":"AI Systems Architect","description":"AI Systems ArchitectLocation: SFO, CA & San Leandro, CA (Hybrid)Duration: Contract for 12 + MonthsJob DescriptionWe are seeking an experienced AI Systems Architect to design, build, and scale high-performance distributed AI systems. The ideal candidate will have deep expertise in GenAI, LLMs, and cloud-native architectures, along with hands-on experience in building enterprise-scale AI/ML platforms and agent-based systems.Must-Have SkillsStrong experience in designing and implementing high-performance, large-scale distributed systemsProven experience in implementing and deploying AI/ML platforms at scaleExpertise in building agent-based architectures, evaluation frameworks, and prompt/context engineeringKnowledge of MCP (Model Context Protocol) serversHands-on experience in LLM inference optimization, including batching and caching strategiesStrong experience with Kubernetes and cloud infrastructure (AWS/Azure/GCP)Proficiency in at least one programming language (Python, Java, Go, etc.)Expertise in designing agent data stacks & retrieval systems, including:Vector databasesHybrid searchData freshness strategiesMemory systemsGraph reasoningBM25 and advanced retrieval techniquesKey ResponsibilitiesArchitect and deliver scalable, high-performance distributed systemsDesign and deploy AI/ML and GenAI platforms at enterprise scaleBuild and manage agent-based architectures, including:Prompt and context engineeringMCP serversEvaluation frameworksOptimize LLM inference pipelines for latency, throughput, and efficiencyDesign and implement agent data & retrieval systems (vector DBs, hybrid search, memory, graph-based reasoning)Lead Kubernetes-based, cloud-native deploymentsProvide technical leadership, architecture governance, and hands-on mentoring to engineering teamsNice to HaveExperience with RAG (Retrieval-Augmented Generation) frameworksFamiliarity with multi-agent systems and orchestration frameworksExposure to real-time data pipelines and streaming architectures","company":"Software Technology","rawCompany":"software technology","city":"San Leandro","state":"CA","isRemote":false,"isActive":true,"createdAt":"2026-06-26T03:31:00.406Z","occupations":[{"code":"15-1299.08","title":"Computer Systems Engineers/Architects","slug":"computer-systems-engineers-architects"},{"code":"15-1252.00","title":"Software Developers","slug":"software-developers"},{"code":"15-1243.00","title":"Database Architects","slug":"database-architects"}],"industries":[{"code":"541512","title":"Computer Systems Design Services","slug":"computer-systems-design-services"},{"code":"541511","title":"Custom Computer Programming Services","slug":"custom-computer-programming-services"},{"code":"513210","title":"Software Publishers","slug":"software-publishers"}],"jobPosting":{"@context":"https://schema.org","@type":"JobPosting","title":"AI Systems Architect","description":"AI Systems ArchitectLocation: SFO, CA & San Leandro, CA (Hybrid)Duration: Contract for 12 + MonthsJob DescriptionWe are seeking an experienced AI Systems Architect to design, build, and scale high-performance distributed AI systems. The ideal candidate will have deep expertise in GenAI, LLMs, and cloud-native architectures, along with hands-on experience in building enterprise-scale AI/ML platforms and agent-based systems.Must-Have SkillsStrong experience in designing and implementing high-performance, large-scale distributed systemsProven experience in implementing and deploying AI/ML platforms at scaleExpertise in building agent-based architectures, evaluation frameworks, and prompt/context engineeringKnowledge of MCP (Model Context Protocol) serversHands-on experience in LLM inference optimization, including batching and caching strategiesStrong experience with Kubernetes and cloud infrastructure (AWS/Azure/GCP)Proficiency in at least one programming language (Python, Java, Go, etc.)Expertise in designing agent data stacks & retrieval systems, including:Vector databasesHybrid searchData freshness strategiesMemory systemsGraph reasoningBM25 and advanced retrieval techniquesKey ResponsibilitiesArchitect and deliver scalable, high-performance distributed systemsDesign and deploy AI/ML and GenAI platforms at enterprise scaleBuild and manage agent-based architectures, including:Prompt and context engineeringMCP serversEvaluation frameworksOptimize LLM inference pipelines for latency, throughput, and efficiencyDesign and implement agent data & retrieval systems (vector DBs, hybrid search, memory, graph-based reasoning)Lead Kubernetes-based, cloud-native deploymentsProvide technical leadership, architecture governance, and hands-on mentoring to engineering teamsNice to HaveExperience with RAG (Retrieval-Augmented Generation) frameworksFamiliarity with multi-agent systems and orchestration frameworksExposure to real-time data pipelines and streaming architectures","datePosted":"2026-06-26T03:31:00.406Z","dateModified":"2026-06-26T03:31:00.406Z","hiringOrganization":{"@type":"Organization","name":"Software Technology","sameAs":"https://jobsearcher.com"},"jobLocation":{"@type":"Place","address":{"@type":"PostalAddress","addressLocality":"San Leandro","addressRegion":"CA","addressCountry":"US"}},"identifier":{"@type":"PropertyValue","name":"JobSearcher","value":"76ea8a8c20a2a99f0a52980b"},"url":"https://jobsearcher.com/jobs/76ea8a8c20a2a99f0a52980b"}}