{"schemaVersion":"jobsearcher.job.v1","id":"cd192e296d0e2f709673d6ee","url":"https://jobsearcher.com/jobs/cd192e296d0e2f709673d6ee","canonicalUrl":"https://jobsearcher.com/jobs/cd192e296d0e2f709673d6ee","title":"Data Scientist","description":"As a Data Scientist – Clinical NLP & AI, you will be part of an agile team focused on building intelligent healthcare solutions by developing advanced NLP modules, integrating LLMs and agentic workflows, and leveraging AWS big data technologies to enhance clinical data processing and usability.\nResponsibilities:\nProficient developer in multiple languages, Python is a must, with the ability to quickly learn new ones.\nExpertise in SQL (complex queries, relational databases preferably PostgreSQL, and NoSQL databases - Redis and Elasticsearch).\nExtensive big data experience, including EMR, Spark, Kafka/Kinesis, and optimizing data pipelines, architectures, and datasets.\nAWS expert with hands-on experience in Lambda, Glue, Athena, Kinesis, IAM, EMR/PySpark, Docker.\nProficient in CI/CD development using Git, Terraform, and agile methodologies.\nComfortable with stream-processing systems (Storm, Spark-Streaming) and workflow management tools (Airflow).\nExposure to knowledge graph technologies (Graph DB, OWL, SPARQL) is a plus.\nExperience in Machine Learning Frameworks: TensorFlow, PyTorch, Scikit-learn, XGBoost.\nExperience in model deployment - Flask, FastAPI, Docker, Kubernetes, TensorFlow Serving, TorchServe.\nSkills:Mandatory skills\nProficient developer in multiple languages, Python is a must, with the ability to quickly learn new ones.\nExpertise in SQL (complex queries, relational databases preferably PostgreSQL, and NoSQL databases - Redis and Elasticsearch).\nExtensive big data experience, including EMR, Spark, Kafka/Kinesis, and optimizing data pipelines, architectures, and datasets.\nAWS expert with hands-on experience in Lambda, Glue, Athena, Kinesis, IAM, EMR/PySpark, Docker.\nProficient in CI/CD development using Git, Terraform, and agile methodologies.\nComfortable with stream-processing systems (Storm, Spark-Streaming) and workflow management tools (Airflow).\nExposure to knowledge graph technologies (Graph DB, OWL, SPARQL) is a plus.\nExperience in Machine Learning Frameworks: TensorFlow, PyTorch, Scikit-learn, XGBoost.\nExperience in model deployment - Flask, FastAPI, Docker, Kubernetes, TensorFlow Serving, TorchServe.\nGood to have skills\nFamiliarity with generative AI applications in healthcare and related use cases.\nUnderstanding of healthcare data standards and terminologies such as HL7, FHIR, and CCDA.\nExperience in creating detailed documentation, user manuals, and technical specifications.\nBackground in automated testing and validation frameworks for NLP outputs.\nAbility to collaborate effectively with cross-functional teams including engineering and products.\nExposure to LangChain or similar frameworks for building intelligent agent workflows.\nEducational Qualifications:\nEngineering Degree – BE/ME/BTech/MTech/BSc/MSc.\nTechnical certification in multiple technologies is desirable.\nJob Type: Contract\nPay: $50.00 - $60.00 per hour\nLocation:\nHouston, TX 77002 (Preferred)\nAbility to Relocate:\nHouston, TX 77002: Relocate before starting work (Required)\nWork Location: In person","company":"Veerteq Solutions Dba Addsource","rawCompany":"veerteq solutions dba addsource","city":"Houston","state":"TX","isRemote":false,"isActive":false,"createdAt":"2026-08-06T12:58:33.846Z","occupations":[{"code":"15-2051.00","title":"Data Scientists","slug":"data-scientists"},{"code":"15-1243.01","title":"Data Warehousing Specialists","slug":"data-warehousing-specialists"},{"code":"15-1252.00","title":"Software Developers","slug":"software-developers"}],"industries":[{"code":"541512","title":"Computer Systems Design Services","slug":"computer-systems-design-services"},{"code":"541990","title":"All Other Professional, Scientific, and Technical Services","slug":"all-other-professional-scientific-and-technical-services"},{"code":"541690","title":"Other Scientific and Technical Consulting Services","slug":"other-scientific-and-technical-consulting-services"}],"jobPosting":{"@context":"https://schema.org","@type":"JobPosting","title":"Data Scientist","description":"As a Data Scientist – Clinical NLP & AI, you will be part of an agile team focused on building intelligent healthcare solutions by developing advanced NLP modules, integrating LLMs and agentic workflows, and leveraging AWS big data technologies to enhance clinical data processing and usability.\nResponsibilities:\nProficient developer in multiple languages, Python is a must, with the ability to quickly learn new ones.\nExpertise in SQL (complex queries, relational databases preferably PostgreSQL, and NoSQL databases - Redis and Elasticsearch).\nExtensive big data experience, including EMR, Spark, Kafka/Kinesis, and optimizing data pipelines, architectures, and datasets.\nAWS expert with hands-on experience in Lambda, Glue, Athena, Kinesis, IAM, EMR/PySpark, Docker.\nProficient in CI/CD development using Git, Terraform, and agile methodologies.\nComfortable with stream-processing systems (Storm, Spark-Streaming) and workflow management tools (Airflow).\nExposure to knowledge graph technologies (Graph DB, OWL, SPARQL) is a plus.\nExperience in Machine Learning Frameworks: TensorFlow, PyTorch, Scikit-learn, XGBoost.\nExperience in model deployment - Flask, FastAPI, Docker, Kubernetes, TensorFlow Serving, TorchServe.\nSkills:Mandatory skills\nProficient developer in multiple languages, Python is a must, with the ability to quickly learn new ones.\nExpertise in SQL (complex queries, relational databases preferably PostgreSQL, and NoSQL databases - Redis and Elasticsearch).\nExtensive big data experience, including EMR, Spark, Kafka/Kinesis, and optimizing data pipelines, architectures, and datasets.\nAWS expert with hands-on experience in Lambda, Glue, Athena, Kinesis, IAM, EMR/PySpark, Docker.\nProficient in CI/CD development using Git, Terraform, and agile methodologies.\nComfortable with stream-processing systems (Storm, Spark-Streaming) and workflow management tools (Airflow).\nExposure to knowledge graph technologies (Graph DB, OWL, SPARQL) is a plus.\nExperience in Machine Learning Frameworks: TensorFlow, PyTorch, Scikit-learn, XGBoost.\nExperience in model deployment - Flask, FastAPI, Docker, Kubernetes, TensorFlow Serving, TorchServe.\nGood to have skills\nFamiliarity with generative AI applications in healthcare and related use cases.\nUnderstanding of healthcare data standards and terminologies such as HL7, FHIR, and CCDA.\nExperience in creating detailed documentation, user manuals, and technical specifications.\nBackground in automated testing and validation frameworks for NLP outputs.\nAbility to collaborate effectively with cross-functional teams including engineering and products.\nExposure to LangChain or similar frameworks for building intelligent agent workflows.\nEducational Qualifications:\nEngineering Degree – BE/ME/BTech/MTech/BSc/MSc.\nTechnical certification in multiple technologies is desirable.\nJob Type: Contract\nPay: $50.00 - $60.00 per hour\nLocation:\nHouston, TX 77002 (Preferred)\nAbility to Relocate:\nHouston, TX 77002: Relocate before starting work (Required)\nWork Location: In person","datePosted":"2026-08-06T12:58:33.846Z","dateModified":"2026-08-06T12:58:33.846Z","hiringOrganization":{"@type":"Organization","name":"Veerteq Solutions Dba Addsource","sameAs":"https://jobsearcher.com"},"jobLocation":{"@type":"Place","address":{"@type":"PostalAddress","addressLocality":"Houston","addressRegion":"TX","addressCountry":"US"}},"identifier":{"@type":"PropertyValue","name":"JobSearcher","value":"cd192e296d0e2f709673d6ee"},"url":"https://jobsearcher.com/jobs/cd192e296d0e2f709673d6ee"}}