{"schemaVersion":"jobsearcher.job.v1","id":"4dd050bfa49ca1d6bc3a0d70","url":"https://jobsearcher.com/jobs/4dd050bfa49ca1d6bc3a0d70","canonicalUrl":"https://jobsearcher.com/jobs/4dd050bfa49ca1d6bc3a0d70","title":"Python AI Engineer","description":"Job Title: Python AI EngineerLocation: Salt Lake City, UT or Phoenix, AZ or Sunrise, FLWorksite: OnsiteAbout WCTWCT is a global talent solutions partner committed to delivering high-impact technology and engineering talent to some of the world's most innovative companies. As a WCT employee, you'll be part of a dynamic, growth-oriented culture that values collaboration, continuous learning, and excellence in execution.Job SummaryWe are hiring a Python Platform Engineer to operate and evolve the infrastructure behind an enterprise knowledge base platform that is moving from a Confluence-focused RAG chatbot into a broader agentic knowledge system.Today, the platform supports Confluence/GitHub ingestion → chunking → pgvector → RAG retrieval → FastAPI serving. Over the next phase, we are expanding toward hybrid retrieval (vector + sparse + graph), multi-source ingestion, evaluation pipelines, agent infrastructure, harness and shared chat platform primitives.What You'll Own:Architect and implement backend services in Python 3.11, FastAPI, Pydantic, SQLAlchemy async, and asyncpgDesign retrieval and orchestration trade-offs around quality, latency, cost, safety, and operational simplicityBuild production-grade agent runtime capabilities: memory boundaries, tool sandboxing, permissions, and budget controlsImprove answer grounding, failure analysis, and citation enforcement rather than optimizing for demo behaviorCreate observability and operational feedback loops with OpenTelemetry, Prometheus/Grafana, Docker/Helm, and GitHub ActionsWork closely with product and engineering partners to support multiple conversational surfaces through one knowledge platformIngestion infrastructure across current and future content sourcesObservability across application, pipeline, database, and model-serving behaviorCost, latency, throughput, and failure-mode management for AI-heavy workloadsRelease workflows that validate AI behavior changes, not just code compilationQualifications:Strong hands-on experience with Python in platform, automation, or infrastructure-heavy environmentsExperience building CLI tools using Python, Golang or Rust.Hands-on experience with LangGraph, LangChain, pgvector, and modern retrieval pipelinesExperience designing evaluation frameworks for LLM-backed systems, including regression detection and quality measurementStrong experience with Docker, Helm, GitHub Actions, and Kubernetes-oriented workflowsFamiliarity with the operational characteristics of embedding pipelines, vector search, and LLM-backed systemsStrong observability skills across metrics, tracing, dashboards, alerting, and log analysisExperience with ingestion, ETL, or content-processing pipelines at scaleAbility to think in terms of reliability, cost, latency, throughput, and recoveryNice to Have:Experience with Qdrant, Neo4j, or other vector/graph infrastructureExperience supporting RAG, search, evaluation, or agent platformsExperience in enterprise or regulated environmentsFamiliarity with Vault, Splunk, Artifactory, ECRComfort using AI-assisted engineering workflows in day-to-day workCompensation / Salary Range: The typical pay range for this role is: USD $110,000/Yearly - $125,000/Yearly. Factors that may affect pay within or outside of this range may include but not limited to geography/market, skills, education, experience, and other qualifications of the successful candidate.Benefits: Medical, dental, Vision, Life, PTO, Holidays, 401(k) benefits and ancillaries may be available for eligible WCT employees and may vary depending on the nature of your employment.WCT will accept applications and processes offers for these roles until the role is filled.Equal Employment Opportunity Declaration:WCT is an equal opportunity employer. All qualified applicants will receive consideration for employment without regard to age, ancestry, color, family or medical care leave, gender identity or expression, genetic information, marital status, medical condition, national origin, physical or mental disability, political affiliation, protected veteran status, race, religion, sex (including pregnancy), sexual orientation, or any other characteristic protected by applicable laws, regulations and ordinances.","company":"Waferwire","rawCompany":"waferwire","city":"Phoenix","state":"AZ","isRemote":false,"isActive":true,"createdAt":"2026-09-19T13:00:32.788Z","occupations":[{"code":"15-1252.00","title":"Software Developers","slug":"software-developers"},{"code":"15-1299.08","title":"Computer Systems Engineers/Architects","slug":"computer-systems-engineers-architects"},{"code":"15-1243.01","title":"Data Warehousing Specialists","slug":"data-warehousing-specialists"}],"industries":[{"code":"541512","title":"Computer Systems Design Services","slug":"computer-systems-design-services"},{"code":"541511","title":"Custom Computer Programming Services","slug":"custom-computer-programming-services"},{"code":"513210","title":"Software Publishers","slug":"software-publishers"}],"jobPosting":{"@context":"https://schema.org","@type":"JobPosting","title":"Python AI Engineer","description":"Job Title: Python AI EngineerLocation: Salt Lake City, UT or Phoenix, AZ or Sunrise, FLWorksite: OnsiteAbout WCTWCT is a global talent solutions partner committed to delivering high-impact technology and engineering talent to some of the world's most innovative companies. As a WCT employee, you'll be part of a dynamic, growth-oriented culture that values collaboration, continuous learning, and excellence in execution.Job SummaryWe are hiring a Python Platform Engineer to operate and evolve the infrastructure behind an enterprise knowledge base platform that is moving from a Confluence-focused RAG chatbot into a broader agentic knowledge system.Today, the platform supports Confluence/GitHub ingestion → chunking → pgvector → RAG retrieval → FastAPI serving. Over the next phase, we are expanding toward hybrid retrieval (vector + sparse + graph), multi-source ingestion, evaluation pipelines, agent infrastructure, harness and shared chat platform primitives.What You'll Own:Architect and implement backend services in Python 3.11, FastAPI, Pydantic, SQLAlchemy async, and asyncpgDesign retrieval and orchestration trade-offs around quality, latency, cost, safety, and operational simplicityBuild production-grade agent runtime capabilities: memory boundaries, tool sandboxing, permissions, and budget controlsImprove answer grounding, failure analysis, and citation enforcement rather than optimizing for demo behaviorCreate observability and operational feedback loops with OpenTelemetry, Prometheus/Grafana, Docker/Helm, and GitHub ActionsWork closely with product and engineering partners to support multiple conversational surfaces through one knowledge platformIngestion infrastructure across current and future content sourcesObservability across application, pipeline, database, and model-serving behaviorCost, latency, throughput, and failure-mode management for AI-heavy workloadsRelease workflows that validate AI behavior changes, not just code compilationQualifications:Strong hands-on experience with Python in platform, automation, or infrastructure-heavy environmentsExperience building CLI tools using Python, Golang or Rust.Hands-on experience with LangGraph, LangChain, pgvector, and modern retrieval pipelinesExperience designing evaluation frameworks for LLM-backed systems, including regression detection and quality measurementStrong experience with Docker, Helm, GitHub Actions, and Kubernetes-oriented workflowsFamiliarity with the operational characteristics of embedding pipelines, vector search, and LLM-backed systemsStrong observability skills across metrics, tracing, dashboards, alerting, and log analysisExperience with ingestion, ETL, or content-processing pipelines at scaleAbility to think in terms of reliability, cost, latency, throughput, and recoveryNice to Have:Experience with Qdrant, Neo4j, or other vector/graph infrastructureExperience supporting RAG, search, evaluation, or agent platformsExperience in enterprise or regulated environmentsFamiliarity with Vault, Splunk, Artifactory, ECRComfort using AI-assisted engineering workflows in day-to-day workCompensation / Salary Range: The typical pay range for this role is: USD $110,000/Yearly - $125,000/Yearly. Factors that may affect pay within or outside of this range may include but not limited to geography/market, skills, education, experience, and other qualifications of the successful candidate.Benefits: Medical, dental, Vision, Life, PTO, Holidays, 401(k) benefits and ancillaries may be available for eligible WCT employees and may vary depending on the nature of your employment.WCT will accept applications and processes offers for these roles until the role is filled.Equal Employment Opportunity Declaration:WCT is an equal opportunity employer. All qualified applicants will receive consideration for employment without regard to age, ancestry, color, family or medical care leave, gender identity or expression, genetic information, marital status, medical condition, national origin, physical or mental disability, political affiliation, protected veteran status, race, religion, sex (including pregnancy), sexual orientation, or any other characteristic protected by applicable laws, regulations and ordinances.","datePosted":"2026-09-19T13:00:32.788Z","dateModified":"2026-09-19T13:00:32.788Z","hiringOrganization":{"@type":"Organization","name":"Waferwire","sameAs":"https://jobsearcher.com"},"jobLocation":{"@type":"Place","address":{"@type":"PostalAddress","addressLocality":"Phoenix","addressRegion":"AZ","addressCountry":"US"}},"identifier":{"@type":"PropertyValue","name":"JobSearcher","value":"4dd050bfa49ca1d6bc3a0d70"},"url":"https://jobsearcher.com/jobs/4dd050bfa49ca1d6bc3a0d70"}}