{"schemaVersion":"jobsearcher.job.v1","id":"b3f79b607f121316e381ca7a","url":"https://jobsearcher.com/jobs/b3f79b607f121316e381ca7a","canonicalUrl":"https://jobsearcher.com/jobs/b3f79b607f121316e381ca7a","title":"Python Engineer - AI for Equities Technology","description":"Python Engineer - AI for Equities Technology\nWe are building a cutting-edge data transformation platform designed to convert unstructured data into structured, machine-readable formats for AI and analytics workflows. This platform processes vast volumes of heterogeneous data, applying advanced parsing methods, enrichment techniques, and LLM-powered extraction to produce high-quality datasets for downstream applications. A key component of this system is the Document Retrieval REST service, which serves as the interface for accessing the transformed data. This service ensures seamless integration with data pipelines and provides performant, scalable document retrieval capabilities.\nAs a cornerstone of the organization’s AI ecosystem, the system ensures reliable data ingestion, model-ready representations, and seamless information flow across multiple products and internal services. You will play a pivotal role in developing and optimizing next-generation data processing pipelines and maintaining the Document Retrieval REST service that powers intelligent applications and automated decision-making at scale.\nYour Role and Impact\nAs a Senior Python Engineer, you will be instrumental in designing, developing, and optimizing high-performance data transformation pipelines and the Document Retrieval REST service. You will leverage your expertise in software engineering and AI/NLP methodologies to deliver scalable, accurate, and robust solutions. This role offers the opportunity to apply advanced language-model technologies in regulated, data-intensive environments while contributing to the broader AI ecosystem.\nKey Responsibilities\nData Pipelines: Design, build, and maintain modular, high-throughput pipelines for ingesting and transforming diverse data types using both traditional parsing methods and AI-powered techniques.\nDocument Retrieval Service: Own the development and maintenance of the Document Retrieval REST service, ensuring seamless integration with data pipelines and high-performance document retrieval.\nDatabase Optimization: Manage and optimize the MongoDB database, ensuring scalability, performance, and efficient document retrieval through techniques like index optimization.\nStatistical Analysis & Governance: Perform statistical analysis on ingested and retrieved data, enforce governance policies, and manage entitlements for data pipelines.\nMetadata Normalization: Ensure metadata structures remain normalized while accommodating the breadth of diverse data sources.\nAI/NLP Integration: Engineer and optimize prompts, extraction logic, and workflows using modern LLM orchestration frameworks for accurate interpretation of complex data.\nAlgorithm Development: Implement algorithms for parsing, segmentation, vectorization, and structured output generation using standard data-processing libraries.\nCloud Deployment: Build and deploy solutions on major cloud platforms, ensuring scalability and reliability.\nProduction-Grade Code: Write resilient, production-ready code for Linux-based environments.\nCross-Functional Collaboration: Work within an Agile team alongside analysts, domain experts, and platform engineers to deliver high-quality solutions.\nCode Quality: Maintain high standards for code quality, testing, and documentation.\nRequired Qualifications (Must-Haves)\nPython Expertise: 5+ years of professional Python experience with strong OOP principles, design patterns, and large-scale system development.\nData Processing: Solid experience with data-processing and numerical libraries (e.g., Pandas, NumPy).\nAI/LLM Integration: Hands-on experience with LLM-integrated application frameworks and AI-powered workflows.\nData Transformation: Proven expertise in parsing and transforming data across multiple formats (structured, semi-structured, unstructured).\nCloud Platforms: Experience deploying solutions on major cloud providers (e.g., AWS, Azure, GCP).\nDatabase Skills: Strong DB proficiency and experience integrating with optimizing databases (NoSQL preferred).\nLinux Proficiency: Comfortable working in Linux/Unix environments with shell scripting.\nPreferred Qualifications (Nice-to-Haves)\nDomain Experience: Background in data-intensive or financial industries.\nBusiness Data Familiarity: Experience working with structured and semi-structured business data.\nNLP Expertise: Familiarity with NLP toolkits, embedding techniques, and text-processing pipelines.\nETL Systems: Experience building and maintaining large-scale ETL or data-processing systems.\nEnterprise Infrastructure: Exposure to enterprise tools like schedulers, monitoring systems, and authentication frameworks.\nTesting Frameworks: Experience with automated testing frameworks for data pipelines and APIs.\nSoft Skills\nCommunication: Clear and structured communication skills to collaborate effectively with cross-functional stakeholders.\nProblem-Solving: Strong analytical mindset with a focus on ownership and accountability.\nTeam Collaboration: Ability to work collaboratively in Agile teams with diverse stakeholders.\nWhy Join Us?\nThis is an exciting opportunity to work on a transformative AI-driven platform that directly impacts intelligent decision-making at scale. You will collaborate with a talented team, tackle complex technical challenges, and contribute to the future of AI-powered data processing.\nThe estimated base salary range for this position is $175,000 to $250,000, which is specific to New York and may change in the future. Millennium pays a total compensation package which includes a base salary, discretionary performance bonus, and a comprehensive benefits package. When finalizing an offer, we take into consideration an individual’s experience level and the qualifications they bring to the role to formulate a competitive total compensation package.","company":"Millenniummanagement","rawCompany":"millenniummanagement","city":"New York","state":"NY","isRemote":false,"isActive":false,"createdAt":"2026-04-14T10:34:03.355Z","occupations":[{"code":"15-1252.00","title":"Software Developers","slug":"software-developers"},{"code":"15-1243.01","title":"Data Warehousing Specialists","slug":"data-warehousing-specialists"},{"code":"15-2051.00","title":"Data Scientists","slug":"data-scientists"}],"industries":[{"code":"541512","title":"Computer Systems Design Services","slug":"computer-systems-design-services"},{"code":"541511","title":"Custom Computer Programming Services","slug":"custom-computer-programming-services"},{"code":"513210","title":"Software Publishers","slug":"software-publishers"}],"jobPosting":{"@context":"https://schema.org","@type":"JobPosting","title":"Python Engineer - AI for Equities Technology","description":"Python Engineer - AI for Equities Technology\nWe are building a cutting-edge data transformation platform designed to convert unstructured data into structured, machine-readable formats for AI and analytics workflows. This platform processes vast volumes of heterogeneous data, applying advanced parsing methods, enrichment techniques, and LLM-powered extraction to produce high-quality datasets for downstream applications. A key component of this system is the Document Retrieval REST service, which serves as the interface for accessing the transformed data. This service ensures seamless integration with data pipelines and provides performant, scalable document retrieval capabilities.\nAs a cornerstone of the organization’s AI ecosystem, the system ensures reliable data ingestion, model-ready representations, and seamless information flow across multiple products and internal services. You will play a pivotal role in developing and optimizing next-generation data processing pipelines and maintaining the Document Retrieval REST service that powers intelligent applications and automated decision-making at scale.\nYour Role and Impact\nAs a Senior Python Engineer, you will be instrumental in designing, developing, and optimizing high-performance data transformation pipelines and the Document Retrieval REST service. You will leverage your expertise in software engineering and AI/NLP methodologies to deliver scalable, accurate, and robust solutions. This role offers the opportunity to apply advanced language-model technologies in regulated, data-intensive environments while contributing to the broader AI ecosystem.\nKey Responsibilities\nData Pipelines: Design, build, and maintain modular, high-throughput pipelines for ingesting and transforming diverse data types using both traditional parsing methods and AI-powered techniques.\nDocument Retrieval Service: Own the development and maintenance of the Document Retrieval REST service, ensuring seamless integration with data pipelines and high-performance document retrieval.\nDatabase Optimization: Manage and optimize the MongoDB database, ensuring scalability, performance, and efficient document retrieval through techniques like index optimization.\nStatistical Analysis & Governance: Perform statistical analysis on ingested and retrieved data, enforce governance policies, and manage entitlements for data pipelines.\nMetadata Normalization: Ensure metadata structures remain normalized while accommodating the breadth of diverse data sources.\nAI/NLP Integration: Engineer and optimize prompts, extraction logic, and workflows using modern LLM orchestration frameworks for accurate interpretation of complex data.\nAlgorithm Development: Implement algorithms for parsing, segmentation, vectorization, and structured output generation using standard data-processing libraries.\nCloud Deployment: Build and deploy solutions on major cloud platforms, ensuring scalability and reliability.\nProduction-Grade Code: Write resilient, production-ready code for Linux-based environments.\nCross-Functional Collaboration: Work within an Agile team alongside analysts, domain experts, and platform engineers to deliver high-quality solutions.\nCode Quality: Maintain high standards for code quality, testing, and documentation.\nRequired Qualifications (Must-Haves)\nPython Expertise: 5+ years of professional Python experience with strong OOP principles, design patterns, and large-scale system development.\nData Processing: Solid experience with data-processing and numerical libraries (e.g., Pandas, NumPy).\nAI/LLM Integration: Hands-on experience with LLM-integrated application frameworks and AI-powered workflows.\nData Transformation: Proven expertise in parsing and transforming data across multiple formats (structured, semi-structured, unstructured).\nCloud Platforms: Experience deploying solutions on major cloud providers (e.g., AWS, Azure, GCP).\nDatabase Skills: Strong DB proficiency and experience integrating with optimizing databases (NoSQL preferred).\nLinux Proficiency: Comfortable working in Linux/Unix environments with shell scripting.\nPreferred Qualifications (Nice-to-Haves)\nDomain Experience: Background in data-intensive or financial industries.\nBusiness Data Familiarity: Experience working with structured and semi-structured business data.\nNLP Expertise: Familiarity with NLP toolkits, embedding techniques, and text-processing pipelines.\nETL Systems: Experience building and maintaining large-scale ETL or data-processing systems.\nEnterprise Infrastructure: Exposure to enterprise tools like schedulers, monitoring systems, and authentication frameworks.\nTesting Frameworks: Experience with automated testing frameworks for data pipelines and APIs.\nSoft Skills\nCommunication: Clear and structured communication skills to collaborate effectively with cross-functional stakeholders.\nProblem-Solving: Strong analytical mindset with a focus on ownership and accountability.\nTeam Collaboration: Ability to work collaboratively in Agile teams with diverse stakeholders.\nWhy Join Us?\nThis is an exciting opportunity to work on a transformative AI-driven platform that directly impacts intelligent decision-making at scale. You will collaborate with a talented team, tackle complex technical challenges, and contribute to the future of AI-powered data processing.\nThe estimated base salary range for this position is $175,000 to $250,000, which is specific to New York and may change in the future. Millennium pays a total compensation package which includes a base salary, discretionary performance bonus, and a comprehensive benefits package. When finalizing an offer, we take into consideration an individual’s experience level and the qualifications they bring to the role to formulate a competitive total compensation package.","datePosted":"2026-04-14T10:34:03.355Z","dateModified":"2026-04-14T10:34:03.355Z","hiringOrganization":{"@type":"Organization","name":"Millenniummanagement","sameAs":"https://jobsearcher.com"},"jobLocation":{"@type":"Place","address":{"@type":"PostalAddress","addressLocality":"New York","addressRegion":"NY","addressCountry":"US"}},"identifier":{"@type":"PropertyValue","name":"JobSearcher","value":"b3f79b607f121316e381ca7a"},"url":"https://jobsearcher.com/jobs/b3f79b607f121316e381ca7a"}}