{"schemaVersion":"jobsearcher.job.v1","id":"f1b1b0e462643d9060292603","url":"https://jobsearcher.com/jobs/f1b1b0e462643d9060292603","canonicalUrl":"https://jobsearcher.com/jobs/f1b1b0e462643d9060292603","title":"Principal Data Engineer","description":"“Be part of a company that is influential and the standard for a rapidly evolving industry!”\nWHO ARE WE?\nRealTime eClinical Solutions is a Global Leader and rapidly growing SaaS technology company that provides comprehensive Software Solutions to the clinical research industry.\nOur Vision is to reshape the global clinical research industry with innovative solutions that help advance medicine and save lives. Our cloud-based solutions are dedicated to solving complex problems and simplifying clinical research processes to be more organized, efficient, and cost-effective. We are based out of San Antonio, TX but are truly a remote and telecommuting company.\nWHAT ARE WE LOOKING FOR?\nThe Principal Data Engineer serves as the technical and analytical authority for the organization’s data science practice. This is a senior individual contributor and team lead role responsible for driving AI/ML strategy, delivering data-driven solutions, and elevating the analytical capability of the wider team. The role combines deep technical expertise in machine learning, NLP, and forecasting with strong product ownership and stakeholder communication skills, ensuring that data science investments translate into measurable business outcomes.\nWHAT WILL YOU BE DOING?\n1. Data Science & AI Leadership\nDefine and drive the data science and AI roadmap, aligning model development priorities with business objectives and product strategy.\nLead end-to-end delivery of ML and AI solutions — from problem framing, data discovery, and model design through validation, deployment, and performance monitoring.\nTranslate ambiguous business problems into well-scoped data science workstreams, identifying quick wins alongside longer-term strategic initiatives.\nChampion best practices in model development, including versioning, documentation, validation, and observability.\n2. NLP, Forecasting & Advanced Analytics\nDesign and implement NLP pipelines for use cases such as entity extraction, semantic mapping, classification, and retrieval-augmented generation (RAG).\nBuild and maintain forecasting and predictive models to support operational and strategic decision-making.\nApply statistical and machine learning methods to identify root causes of process inefficiencies and data quality issues.\nDevelop reusable data pipelines, crosswalk tables, and transformation workflows that support scalable, cross-functional data products.\n3. Data & Process Analysis\nConduct current-state assessments of data architecture, sources, and quality; define future-state data models and governance standards.\nDevelop and maintain KPI reporting frameworks and dashboards that enable performance monitoring and data-driven decision-making.\nApply process optimization methodologies (e.g., Lean Six Sigma) to identify bottlenecks, reduce cycle time, and improve data accuracy.\nEnsure analytical outputs are accurate, auditable, and aligned with regulatory and compliance requirements (e.g., HIPAA, GDPR).\n4. Stakeholder Collaboration & Communication\nPartner closely with Product, Engineering, and business stakeholders to clarify requirements, validate feasibility, and define measurable success criteria.\nCommunicate complex analytical findings and model outputs clearly to both technical and non-technical audiences, including executive stakeholders.\nDefine value-realization strategies for data and AI investments, ensuring ROI is tracked through improved search, reporting, and operational insight.\n5. Team Development & Knowledge Transfer\nMentor data analysts and junior data scientists through pairing, design reviews, and structured technical guidance.\nLead knowledge transfer of owned models, pipelines, and analytical frameworks to ensure team resilience and continuity.\nDrive a culture of continuous learning, analytical rigor, and responsible AI within the data science function.\nPerformance at this level is evaluated across four dimensions:\nTechnical Quality\nModels, pipelines, and analytical frameworks consistently meet peer review standards and produce reliable, reproducible results.\nData quality, model drift, and technical debt in areas of ownership trend downward over time.\nDelivery & Impact\nData science deliverables are completed on schedule with accurate effort estimation; risks and blockers are surfaced early.\nAnalytical insights are actionable, with measurable improvements in KPIs such as process efficiency, data accuracy, or cost reduction.\nOrganizational Impact\nData analysts and engineers who regularly collaborate with this role demonstrably improve analytical judgment and technical practice.\nData science recommendations are trusted by peers, product leadership, and executive stakeholders without requiring repeated validation.\nCommunication & Leadership\nAmbiguous analytical problems are framed and decomposed independently, without waiting for direction.\nFindings and model outputs are presented in a way that drives clear decisions by non-technical stakeholders.\nWHAT DO YOU NEED?\nBachelor’s degree in data science, Computer Science, Mathematics, Statistics, Economics, or a related quantitative field, or equivalent professional experience.\n5+ years of experience in data science, data analytics, or a related discipline, including production ML/AI deployments.\nStrong proficiency in Python for data science workflows, including pandas, scikit-learn, and NLP libraries (e.g., spaCy, Hugging Face Transformers).\nProven experience designing and delivering NLP pipelines and/or forecasting models in a business context.\nSolid command of SQL for data querying, transformation, and analysis across relational databases.\nExperience with BI and reporting tools, particularly Power BI, including data modeling and DAX.\nDemonstrated ability to communicate analytical findings clearly to non-technical stakeholders and drive decision-making.\nExperience working in regulated industries (healthcare, finance, or similar) with an understanding of compliance and data governance requirements.\nWHAT SETS YOU APART?\n7+ years of data science or analytics experience, including a team lead or principal contributor role.\nExperience with Large Language Models (LLMs), Retrieval-Augmented Generation (RAG), and prompt engineering for enterprise use cases.\nFamiliarity with cloud-based ML platforms (GCP, AWS, or Azure) and MLOps practices.\nExperience with process automation tools (e.g., UiPath or similar RPA platforms).\nWorking knowledge of process optimization frameworks such as Lean Six Sigma (Green Belt or higher).\nExposure to clinical data standards, health data interoperability, or cross-client data standardization projects.\nProficiency in data visualization and dashboard design; PL-300 Power BI Data Analyst certification is a plus.\nWHAT IS IN IT FOR YOU?\nThe company sponsors health insurance, long-term disability, and life insurance.\nUnlimited Paid Time Off.\n10 paid Holidays.\nPaid Parental Leave.\nWork Anniversary Bonus.\nParticipation in the Employee of the Quarter Program.\nMonthly $100 Connectivity Stipend Reimbursement.\nRealTime matches employee 401K contributions at 100% of the first 3% invested and 50% of the next 2% invested.\nAll successful candidates must complete and pass reference and background checks.\nThe desired salary must be indicated for the application to be considered.\nThe pay rate is commensurate with experience and is determined on an individual basis after an interview has occurred.\nEqual Opportunity Employer – RealTime eClinical Solutions strongly values diversity and is committed to equal opportunity and non-discrimination in all of its policies and practices, including employment.\nYour Right to Work – In compliance with federal law, all persons hired will be required to verify identity and eligibility.\nThank you for your interest in RealTime eClinical Solutions.","company":"Realtime Software Solutions","rawCompany":"realtime software solutions","city":"Remote","state":"OR","isRemote":false,"isActive":false,"createdAt":"2026-08-04T00:24:03.164Z","occupations":[{"code":"15-2051.00","title":"Data Scientists","slug":"data-scientists"},{"code":"15-1243.01","title":"Data Warehousing Specialists","slug":"data-warehousing-specialists"},{"code":"15-1243.00","title":"Database Architects","slug":"database-architects"}],"industries":[{"code":"541690","title":"Other Scientific and Technical Consulting Services","slug":"other-scientific-and-technical-consulting-services"},{"code":"541512","title":"Computer Systems Design Services","slug":"computer-systems-design-services"},{"code":"541511","title":"Custom Computer Programming Services","slug":"custom-computer-programming-services"}],"jobPosting":{"@context":"https://schema.org","@type":"JobPosting","title":"Principal Data Engineer","description":"“Be part of a company that is influential and the standard for a rapidly evolving industry!”\nWHO ARE WE?\nRealTime eClinical Solutions is a Global Leader and rapidly growing SaaS technology company that provides comprehensive Software Solutions to the clinical research industry.\nOur Vision is to reshape the global clinical research industry with innovative solutions that help advance medicine and save lives. Our cloud-based solutions are dedicated to solving complex problems and simplifying clinical research processes to be more organized, efficient, and cost-effective. We are based out of San Antonio, TX but are truly a remote and telecommuting company.\nWHAT ARE WE LOOKING FOR?\nThe Principal Data Engineer serves as the technical and analytical authority for the organization’s data science practice. This is a senior individual contributor and team lead role responsible for driving AI/ML strategy, delivering data-driven solutions, and elevating the analytical capability of the wider team. The role combines deep technical expertise in machine learning, NLP, and forecasting with strong product ownership and stakeholder communication skills, ensuring that data science investments translate into measurable business outcomes.\nWHAT WILL YOU BE DOING?\n1. Data Science & AI Leadership\nDefine and drive the data science and AI roadmap, aligning model development priorities with business objectives and product strategy.\nLead end-to-end delivery of ML and AI solutions — from problem framing, data discovery, and model design through validation, deployment, and performance monitoring.\nTranslate ambiguous business problems into well-scoped data science workstreams, identifying quick wins alongside longer-term strategic initiatives.\nChampion best practices in model development, including versioning, documentation, validation, and observability.\n2. NLP, Forecasting & Advanced Analytics\nDesign and implement NLP pipelines for use cases such as entity extraction, semantic mapping, classification, and retrieval-augmented generation (RAG).\nBuild and maintain forecasting and predictive models to support operational and strategic decision-making.\nApply statistical and machine learning methods to identify root causes of process inefficiencies and data quality issues.\nDevelop reusable data pipelines, crosswalk tables, and transformation workflows that support scalable, cross-functional data products.\n3. Data & Process Analysis\nConduct current-state assessments of data architecture, sources, and quality; define future-state data models and governance standards.\nDevelop and maintain KPI reporting frameworks and dashboards that enable performance monitoring and data-driven decision-making.\nApply process optimization methodologies (e.g., Lean Six Sigma) to identify bottlenecks, reduce cycle time, and improve data accuracy.\nEnsure analytical outputs are accurate, auditable, and aligned with regulatory and compliance requirements (e.g., HIPAA, GDPR).\n4. Stakeholder Collaboration & Communication\nPartner closely with Product, Engineering, and business stakeholders to clarify requirements, validate feasibility, and define measurable success criteria.\nCommunicate complex analytical findings and model outputs clearly to both technical and non-technical audiences, including executive stakeholders.\nDefine value-realization strategies for data and AI investments, ensuring ROI is tracked through improved search, reporting, and operational insight.\n5. Team Development & Knowledge Transfer\nMentor data analysts and junior data scientists through pairing, design reviews, and structured technical guidance.\nLead knowledge transfer of owned models, pipelines, and analytical frameworks to ensure team resilience and continuity.\nDrive a culture of continuous learning, analytical rigor, and responsible AI within the data science function.\nPerformance at this level is evaluated across four dimensions:\nTechnical Quality\nModels, pipelines, and analytical frameworks consistently meet peer review standards and produce reliable, reproducible results.\nData quality, model drift, and technical debt in areas of ownership trend downward over time.\nDelivery & Impact\nData science deliverables are completed on schedule with accurate effort estimation; risks and blockers are surfaced early.\nAnalytical insights are actionable, with measurable improvements in KPIs such as process efficiency, data accuracy, or cost reduction.\nOrganizational Impact\nData analysts and engineers who regularly collaborate with this role demonstrably improve analytical judgment and technical practice.\nData science recommendations are trusted by peers, product leadership, and executive stakeholders without requiring repeated validation.\nCommunication & Leadership\nAmbiguous analytical problems are framed and decomposed independently, without waiting for direction.\nFindings and model outputs are presented in a way that drives clear decisions by non-technical stakeholders.\nWHAT DO YOU NEED?\nBachelor’s degree in data science, Computer Science, Mathematics, Statistics, Economics, or a related quantitative field, or equivalent professional experience.\n5+ years of experience in data science, data analytics, or a related discipline, including production ML/AI deployments.\nStrong proficiency in Python for data science workflows, including pandas, scikit-learn, and NLP libraries (e.g., spaCy, Hugging Face Transformers).\nProven experience designing and delivering NLP pipelines and/or forecasting models in a business context.\nSolid command of SQL for data querying, transformation, and analysis across relational databases.\nExperience with BI and reporting tools, particularly Power BI, including data modeling and DAX.\nDemonstrated ability to communicate analytical findings clearly to non-technical stakeholders and drive decision-making.\nExperience working in regulated industries (healthcare, finance, or similar) with an understanding of compliance and data governance requirements.\nWHAT SETS YOU APART?\n7+ years of data science or analytics experience, including a team lead or principal contributor role.\nExperience with Large Language Models (LLMs), Retrieval-Augmented Generation (RAG), and prompt engineering for enterprise use cases.\nFamiliarity with cloud-based ML platforms (GCP, AWS, or Azure) and MLOps practices.\nExperience with process automation tools (e.g., UiPath or similar RPA platforms).\nWorking knowledge of process optimization frameworks such as Lean Six Sigma (Green Belt or higher).\nExposure to clinical data standards, health data interoperability, or cross-client data standardization projects.\nProficiency in data visualization and dashboard design; PL-300 Power BI Data Analyst certification is a plus.\nWHAT IS IN IT FOR YOU?\nThe company sponsors health insurance, long-term disability, and life insurance.\nUnlimited Paid Time Off.\n10 paid Holidays.\nPaid Parental Leave.\nWork Anniversary Bonus.\nParticipation in the Employee of the Quarter Program.\nMonthly $100 Connectivity Stipend Reimbursement.\nRealTime matches employee 401K contributions at 100% of the first 3% invested and 50% of the next 2% invested.\nAll successful candidates must complete and pass reference and background checks.\nThe desired salary must be indicated for the application to be considered.\nThe pay rate is commensurate with experience and is determined on an individual basis after an interview has occurred.\nEqual Opportunity Employer – RealTime eClinical Solutions strongly values diversity and is committed to equal opportunity and non-discrimination in all of its policies and practices, including employment.\nYour Right to Work – In compliance with federal law, all persons hired will be required to verify identity and eligibility.\nThank you for your interest in RealTime eClinical Solutions.","datePosted":"2026-08-04T00:24:03.164Z","dateModified":"2026-08-04T00:24:03.164Z","hiringOrganization":{"@type":"Organization","name":"Realtime Software Solutions","sameAs":"https://jobsearcher.com"},"jobLocation":{"@type":"Place","address":{"@type":"PostalAddress","addressLocality":"Remote","addressRegion":"OR","addressCountry":"US"}},"identifier":{"@type":"PropertyValue","name":"JobSearcher","value":"f1b1b0e462643d9060292603"},"url":"https://jobsearcher.com/jobs/f1b1b0e462643d9060292603"}}