{"schemaVersion":"jobsearcher.job.v1","id":"2cc49a63057871f81d1e5ca5","url":"https://jobsearcher.com/jobs/2cc49a63057871f81d1e5ca5","canonicalUrl":"https://jobsearcher.com/jobs/2cc49a63057871f81d1e5ca5","title":"Solution Lead/Java Technical Lead","description":"Job DescriptionMust Have Technical/Functional Skills• 13+ years of experience with IT• Build and productionize cloud native backend services and AI/LLM inference pipelines.•• Design and develop Python-based APIs and microservices (FastAPI, async patterns) and agentic AI workflows using LangChain/LangGraph.•• Implement and optimize LLM capabilities including embeddings, RAG, vector search, prompt/context engineering, and model versioning.•• Package, serve, and monitor models for real-time and batch inference, ensuring operational readiness and performance.•• Build event driven, resilient integrations and containerized services, with hands-on Kubernetes debugging and Helm-based deployments.•• Establish observability, SLOs, CI/CD automation, testing•• Apply strong systems design principles (concurrency, caching, reliability, rate limiting) and robust data engineering practices.•• Cloud exposure preferred (Azure/AKS, managed services), with bonus experience in performance tuning, frontend collaboration, and model governance/monitoring.Roles & Responsibilities• Build and productionize cloud native backend services and AI/LLM inference pipelines.•• Design and develop Python-based APIs and microservices (FastAPI, async patterns) and agentic AI workflows using LangChain/LangGraph.•• Implement and optimize LLM capabilities including embeddings, RAG, vector search, prompt/context engineering, and model versioning.•• Package, serve, and monitor models for real-time and batch inference, ensuring operational readiness and performance.•• Build event driven, resilient integrations and containerized services, with hands-on Kubernetes debugging and Helm-based deployments.•• Establish observability, SLOs, CI/CD automation, testing•• Apply strong systems design principles (concurrency, caching, reliability, rate limiting) and robust data engineering practices.•• Cloud exposure preferred (Azure/AKS, managed services), with bonus experience in performance tuning, frontend collaboration, and model governance/monitoring.","company":"Hireon Tech","rawCompany":"hireon tech","city":"Cary","state":"NC","isRemote":false,"isActive":false,"createdAt":"2026-08-11T11:28:14.255Z","occupations":[{"code":"15-1252.00","title":"Software Developers","slug":"software-developers"},{"code":"15-1299.08","title":"Computer Systems Engineers/Architects","slug":"computer-systems-engineers-architects"},{"code":"15-1211.00","title":"Computer Systems Analysts","slug":"computer-systems-analysts"}],"industries":[{"code":"541512","title":"Computer Systems Design Services","slug":"computer-systems-design-services"},{"code":"541511","title":"Custom Computer Programming Services","slug":"custom-computer-programming-services"},{"code":"513210","title":"Software Publishers","slug":"software-publishers"}],"jobPosting":{"@context":"https://schema.org","@type":"JobPosting","title":"Solution Lead/Java Technical Lead","description":"Job DescriptionMust Have Technical/Functional Skills• 13+ years of experience with IT• Build and productionize cloud native backend services and AI/LLM inference pipelines.•• Design and develop Python-based APIs and microservices (FastAPI, async patterns) and agentic AI workflows using LangChain/LangGraph.•• Implement and optimize LLM capabilities including embeddings, RAG, vector search, prompt/context engineering, and model versioning.•• Package, serve, and monitor models for real-time and batch inference, ensuring operational readiness and performance.•• Build event driven, resilient integrations and containerized services, with hands-on Kubernetes debugging and Helm-based deployments.•• Establish observability, SLOs, CI/CD automation, testing•• Apply strong systems design principles (concurrency, caching, reliability, rate limiting) and robust data engineering practices.•• Cloud exposure preferred (Azure/AKS, managed services), with bonus experience in performance tuning, frontend collaboration, and model governance/monitoring.Roles & Responsibilities• Build and productionize cloud native backend services and AI/LLM inference pipelines.•• Design and develop Python-based APIs and microservices (FastAPI, async patterns) and agentic AI workflows using LangChain/LangGraph.•• Implement and optimize LLM capabilities including embeddings, RAG, vector search, prompt/context engineering, and model versioning.•• Package, serve, and monitor models for real-time and batch inference, ensuring operational readiness and performance.•• Build event driven, resilient integrations and containerized services, with hands-on Kubernetes debugging and Helm-based deployments.•• Establish observability, SLOs, CI/CD automation, testing•• Apply strong systems design principles (concurrency, caching, reliability, rate limiting) and robust data engineering practices.•• Cloud exposure preferred (Azure/AKS, managed services), with bonus experience in performance tuning, frontend collaboration, and model governance/monitoring.","datePosted":"2026-08-11T11:28:14.255Z","dateModified":"2026-08-11T11:28:14.255Z","hiringOrganization":{"@type":"Organization","name":"Hireon Tech","sameAs":"https://jobsearcher.com"},"jobLocation":{"@type":"Place","address":{"@type":"PostalAddress","addressLocality":"Cary","addressRegion":"NC","addressCountry":"US"}},"identifier":{"@type":"PropertyValue","name":"JobSearcher","value":"2cc49a63057871f81d1e5ca5"},"url":"https://jobsearcher.com/jobs/2cc49a63057871f81d1e5ca5"}}