{"schemaVersion":"jobsearcher.job.v1","id":"f375acbfd8bfcf1bb00106b7","url":"https://jobsearcher.com/jobs/f375acbfd8bfcf1bb00106b7","canonicalUrl":"https://jobsearcher.com/jobs/f375acbfd8bfcf1bb00106b7","title":"Artificial Intelligence Engineer","description":"Senior Software Engineer — AI/MLPlano, TX | On-siteAbout the CompanyA tier-one financial services enterprise is building next-generation AI infrastructure to power intelligent conversational systems—from chatbots handling routine transactions to voicebots managing complex workflows. They're scaling from proof-of-concept to production systems handling millions of interactions monthly.About the RoleThis is a senior engineering role focused on shipping robust, reliable AI applications on AWS, with direct impact on how millions of customers interact with financial services. You'll own the full stack: from designing RAG pipelines and prompt strategies, to orchestrating LLM workflows, to deploying containerized services that run 24/7. You'll work with leading models (including Claude via AWS Bedrock), define guardrails and responsible AI practices, and mentor mid-level engineers on production AI patterns.ResponsibilitiesDesign and build production RAG systems — from chunking strategy and embedding model selection, to vector store architecture (pgvector, OpenSearch, Pinecone, or FAISS), to retrieval optimization for accuracy and latencyOrchestrate complex LLM workflows using frameworks like LangChain, LlamaIndex, or Semantic Kernel; implement multi-step reasoning, tool use, and structured output patternsDevelop conversational AI systems across text (chatbots) and speech (voicebots with STT/TTS integration); manage latency and streaming constraints in real-time voiceLeverage AWS Bedrock and Claude models directly; design effective system prompts, few-shot examples, and chain-of-thought reasoning; iterate based on real-world performanceBuild and scale AWS infrastructure — Lambda functions, API Gateway, Step Functions, DynamoDB, SQS, S3; implement CI/CD pipelines with Docker/ECS or EKS; write infrastructure-as-codeDefine quality and safety guardrails — implement content filtering, responsible AI practices, and evaluation frameworks to catch model drift and output degradation before production impactOwn API design (REST, GraphQL) and microservices architecture; build event-driven systems that integrate seamlessly with enterprise systemsMentor and unblock junior engineers; establish patterns and best practices for how your team ships AI features reliablyQualifications10+ years shipping production software in Python or Java (Python strongly preferred)2+ years hands-on building and deploying AI/ML applications in production — not coursework, not personal projects; real systems in a real environmentDeep RAG expertise — you've built chunking strategies, selected and tuned embedding models, chosen vector stores based on use case, and optimized retrieval for productionProficiency with AI orchestration frameworks (LangChain, LlamaIndex, Semantic Kernel, or CrewAI); you can explain the trade-offs between them and when to use eachHands-on AWS Bedrock and Claude experience — direct invocation, token counting, structured outputs, prompt optimization; familiarity with other LLM APIs (OpenAI, Azure) is transferable but AWS Bedrock depth is expectedAdvanced prompt engineering skills — system prompts, few-shot learning, chain-of-thought reasoning, tool use, and structured outputs aren't abstract concepts to you; you've debugged them in productionConversational AI systems — you've shipped text-based chatbots AND ideally have experience with voice systems (speech-to-text, text-to-speech, real-time streaming)Senior AWS proficiency — Lambda, API Gateway, Step Functions, DynamoDB, SQS, S3; you build without hand-holding and reason through cost and performance trade-offsCI/CD and containerization — solid grasp of Docker, ECS/EKS, infrastructure-as-code; shipping changes frequently and safely is table stakesAPI and systems design — you can articulate why you chose REST vs. GraphQL, how events flow through microservices, and where synchronous vs. asynchronous makes senseEvaluation and safety mindset — you've built evaluation frameworks for LLM outputs, implemented guardrails, and thought deeply about what can go wrong at scalePreferred SkillsExperience with model fine-tuning or retrieval optimization (RAG vs. in-context learning trade-offs)GraphQL implementation in productionFamiliarity with additional vector stores (Weaviate, Chroma, pgvector)Background in financial services or regulated industriesPublished writing or talks on AI systems architectureExperience with multi-modal models or agent frameworks beyond basic tool usePay range and compensation packageCompetitive salary, equity, and benefits package commensurate with experience.Equal Opportunity StatementWe're committed to building a diverse team and welcome applications from underrepresented groups in engineering. We do not discriminate based on race, color, religion, sex, national origin, age, disability, veteran status, sexual orientation, or any other protected characteristic.How to ApplyIf this role fits your background and you're ready to ship production AI systems at scale, please apply with:Your resume — include 2–3 specific examples of AI/ML production work (system, your role, outcome)Optional: Link to GitHub, blog, or publicly shipped workWe review applications on a rolling basis and move quickly with serious candidates.","company":"Operaxis","rawCompany":"operaxis","city":"Garland","state":"TX","isRemote":false,"isActive":false,"createdAt":"2026-07-14T07:38:38.114Z","occupations":[{"code":"15-1252.00","title":"Software Developers","slug":"software-developers"},{"code":"15-1299.08","title":"Computer Systems Engineers/Architects","slug":"computer-systems-engineers-architects"},{"code":"15-1251.00","title":"Computer Programmers","slug":"computer-programmers"}],"industries":[{"code":"541512","title":"Computer Systems Design Services","slug":"computer-systems-design-services"},{"code":"541511","title":"Custom Computer Programming Services","slug":"custom-computer-programming-services"},{"code":"513210","title":"Software Publishers","slug":"software-publishers"}],"jobPosting":{"@context":"https://schema.org","@type":"JobPosting","title":"Artificial Intelligence Engineer","description":"Senior Software Engineer — AI/MLPlano, TX | On-siteAbout the CompanyA tier-one financial services enterprise is building next-generation AI infrastructure to power intelligent conversational systems—from chatbots handling routine transactions to voicebots managing complex workflows. They're scaling from proof-of-concept to production systems handling millions of interactions monthly.About the RoleThis is a senior engineering role focused on shipping robust, reliable AI applications on AWS, with direct impact on how millions of customers interact with financial services. You'll own the full stack: from designing RAG pipelines and prompt strategies, to orchestrating LLM workflows, to deploying containerized services that run 24/7. You'll work with leading models (including Claude via AWS Bedrock), define guardrails and responsible AI practices, and mentor mid-level engineers on production AI patterns.ResponsibilitiesDesign and build production RAG systems — from chunking strategy and embedding model selection, to vector store architecture (pgvector, OpenSearch, Pinecone, or FAISS), to retrieval optimization for accuracy and latencyOrchestrate complex LLM workflows using frameworks like LangChain, LlamaIndex, or Semantic Kernel; implement multi-step reasoning, tool use, and structured output patternsDevelop conversational AI systems across text (chatbots) and speech (voicebots with STT/TTS integration); manage latency and streaming constraints in real-time voiceLeverage AWS Bedrock and Claude models directly; design effective system prompts, few-shot examples, and chain-of-thought reasoning; iterate based on real-world performanceBuild and scale AWS infrastructure — Lambda functions, API Gateway, Step Functions, DynamoDB, SQS, S3; implement CI/CD pipelines with Docker/ECS or EKS; write infrastructure-as-codeDefine quality and safety guardrails — implement content filtering, responsible AI practices, and evaluation frameworks to catch model drift and output degradation before production impactOwn API design (REST, GraphQL) and microservices architecture; build event-driven systems that integrate seamlessly with enterprise systemsMentor and unblock junior engineers; establish patterns and best practices for how your team ships AI features reliablyQualifications10+ years shipping production software in Python or Java (Python strongly preferred)2+ years hands-on building and deploying AI/ML applications in production — not coursework, not personal projects; real systems in a real environmentDeep RAG expertise — you've built chunking strategies, selected and tuned embedding models, chosen vector stores based on use case, and optimized retrieval for productionProficiency with AI orchestration frameworks (LangChain, LlamaIndex, Semantic Kernel, or CrewAI); you can explain the trade-offs between them and when to use eachHands-on AWS Bedrock and Claude experience — direct invocation, token counting, structured outputs, prompt optimization; familiarity with other LLM APIs (OpenAI, Azure) is transferable but AWS Bedrock depth is expectedAdvanced prompt engineering skills — system prompts, few-shot learning, chain-of-thought reasoning, tool use, and structured outputs aren't abstract concepts to you; you've debugged them in productionConversational AI systems — you've shipped text-based chatbots AND ideally have experience with voice systems (speech-to-text, text-to-speech, real-time streaming)Senior AWS proficiency — Lambda, API Gateway, Step Functions, DynamoDB, SQS, S3; you build without hand-holding and reason through cost and performance trade-offsCI/CD and containerization — solid grasp of Docker, ECS/EKS, infrastructure-as-code; shipping changes frequently and safely is table stakesAPI and systems design — you can articulate why you chose REST vs. GraphQL, how events flow through microservices, and where synchronous vs. asynchronous makes senseEvaluation and safety mindset — you've built evaluation frameworks for LLM outputs, implemented guardrails, and thought deeply about what can go wrong at scalePreferred SkillsExperience with model fine-tuning or retrieval optimization (RAG vs. in-context learning trade-offs)GraphQL implementation in productionFamiliarity with additional vector stores (Weaviate, Chroma, pgvector)Background in financial services or regulated industriesPublished writing or talks on AI systems architectureExperience with multi-modal models or agent frameworks beyond basic tool usePay range and compensation packageCompetitive salary, equity, and benefits package commensurate with experience.Equal Opportunity StatementWe're committed to building a diverse team and welcome applications from underrepresented groups in engineering. We do not discriminate based on race, color, religion, sex, national origin, age, disability, veteran status, sexual orientation, or any other protected characteristic.How to ApplyIf this role fits your background and you're ready to ship production AI systems at scale, please apply with:Your resume — include 2–3 specific examples of AI/ML production work (system, your role, outcome)Optional: Link to GitHub, blog, or publicly shipped workWe review applications on a rolling basis and move quickly with serious candidates.","datePosted":"2026-07-14T07:38:38.114Z","dateModified":"2026-07-14T07:38:38.114Z","hiringOrganization":{"@type":"Organization","name":"Operaxis","sameAs":"https://jobsearcher.com"},"jobLocation":{"@type":"Place","address":{"@type":"PostalAddress","addressLocality":"Garland","addressRegion":"TX","addressCountry":"US"}},"identifier":{"@type":"PropertyValue","name":"JobSearcher","value":"f375acbfd8bfcf1bb00106b7"},"url":"https://jobsearcher.com/jobs/f375acbfd8bfcf1bb00106b7"}}