{"schemaVersion":"jobsearcher.job.v1","id":"f79ac65e2b2893eb8b2de3ec","url":"https://jobsearcher.com/jobs/f79ac65e2b2893eb8b2de3ec","canonicalUrl":"https://jobsearcher.com/jobs/f79ac65e2b2893eb8b2de3ec","title":"Data Scientist","description":"The Data Scientists' primary responsibility is to design, build, and maintain benchmarks for model testing and agentic workflow testing using InspectAI. They define evaluation datasets, test scenarios, metrics, scoring approaches, and repeatable evaluation workflows that measure model and agent performance in a consistent, auditable way. They also clean, transform, validate, and analyze supporting data; generate metrics and visualizations; and produce findings that support model evaluation, reporting, and decision-making. They partner with Researchers to ground benchmark design in sound methodology, and with Developers to operationalize evaluation workflows into stable, repeatable tooling. Data Scientists should also be able to document dataset lineage, scoring rationale, benchmark assumptions, and evaluation outputs so results are reproducible, traceable, and usable in constrained or airgapped environments.\nTechnologies / Skills Needed:\nInspectAI\nPython\nMongoDB\nJupyter notebooks\nData visualization tools\nStatistical analysis\nData cleaning and validation\nCloud or containerized environments\nGit-based version control\nCollaboration tools such as Jira or Confluence\nExperience operating in airgapped or constrained environments\n** Position Requires an Active TS/SCI with a Full-Scope Polygraph Clearance**\nWhat You’ll Love About Synergist:\nWe have an actual “award-winning” culture\nWe recognize that our employees are our most valuable asset, and we do everything in our power to give you the best experience every day\nWe’re big believers in transparency and you’ll always have a seat at the table for big decisions and can even attend key business meetings if you want\nWe offer a best-in-class benefits package that we can customize for your unique needs. Offerings include healthcare, dental, and vision coverage for you and your dependents, flexible Paid Time Off (PTO) accrual, 10% 401k employer contribution, annual education allowance, and many perks from a robust referral program, weekly catered lunches, monthly happy hours, access to company suite at Cap One Arena, financial advisory, and so much more\nWe have a deep catalog (and growing!) of potential project opportunities for you to choose from, and when you’re ready for new challenges, we’ll work to make that happen\nWe Want to Hear From You If:\nYou have a BS in Computer Science or similar technical field (4 years of relevant experience may be substituted in lieu of a BS degree)\nYou have experience with AI Evaluation & Frameworks such as InspectAI, LLM Benchmarking, Agentic Workflow Evaluation (tool-use, multi-step task performance)\nYou have experience with Statistical Analysis, Data Cleaning & Validation, Data Visualization Tools, Jupyter NotebooksData & Environment: MongoDB, Cloud/Containerized Environments, Airgapped & Constrained Operations\nYou're familiar with Git Version Control, Experimental Design, Dataset Lineage/Traceability, Jira, and Confluence\nHere’s What You Can Expect Out Of This Role:\nAI Model Benchmarks — Design, implement, and maintain benchmark suites that measure model performance against defined evaluation criteria, including dataset curation, scoring, and results reporting.\nAgentic AI Benchmarks — Develop benchmarks and test scenarios that evaluate agentic workflows and multi-step task performance, including tool-use, task completion, and reliability metrics.\nEvaluation Dataset Development — Curate, validate, and version evaluation datasets and test scenarios, including any adversarial or edge-case sets needed to stress agent behavior.\nMetrics, Analysis, and Reporting — Produce scoring frameworks, visualizations, and written analyses that translate raw evaluation results into actionable findings for stakeholders.\nJob Type: Full-time\nPay: $100,000.00 - $280,000.00 per year\nBenefits:\n401(k)\nDental insurance\nEmployee assistance program\nFlexible schedule\nFlexible spending account\nHealth insurance\nHealth savings account\nLife insurance\nPaid time off\nParental leave\nProfessional development assistance\nReferral program\nTuition reimbursement\nVision insurance\nSecurity clearance:\nTop Secret (Required)\nWork Location: In person","company":"Synergist Computing","rawCompany":"synergist computing","city":"Annapolis Junction","state":"MD","isRemote":false,"isActive":false,"createdAt":"2026-08-03T18:12:19.945Z","occupations":[{"code":"15-2051.00","title":"Data Scientists","slug":"data-scientists"},{"code":"15-1243.01","title":"Data Warehousing Specialists","slug":"data-warehousing-specialists"},{"code":"15-1221.00","title":"Computer and Information Research Scientists","slug":"computer-and-information-research-scientists"}],"industries":[{"code":"541690","title":"Other Scientific and Technical Consulting Services","slug":"other-scientific-and-technical-consulting-services"},{"code":"541512","title":"Computer Systems Design Services","slug":"computer-systems-design-services"},{"code":"541990","title":"All Other Professional, Scientific, and Technical Services","slug":"all-other-professional-scientific-and-technical-services"}],"jobPosting":{"@context":"https://schema.org","@type":"JobPosting","title":"Data Scientist","description":"The Data Scientists' primary responsibility is to design, build, and maintain benchmarks for model testing and agentic workflow testing using InspectAI. They define evaluation datasets, test scenarios, metrics, scoring approaches, and repeatable evaluation workflows that measure model and agent performance in a consistent, auditable way. They also clean, transform, validate, and analyze supporting data; generate metrics and visualizations; and produce findings that support model evaluation, reporting, and decision-making. They partner with Researchers to ground benchmark design in sound methodology, and with Developers to operationalize evaluation workflows into stable, repeatable tooling. Data Scientists should also be able to document dataset lineage, scoring rationale, benchmark assumptions, and evaluation outputs so results are reproducible, traceable, and usable in constrained or airgapped environments.\nTechnologies / Skills Needed:\nInspectAI\nPython\nMongoDB\nJupyter notebooks\nData visualization tools\nStatistical analysis\nData cleaning and validation\nCloud or containerized environments\nGit-based version control\nCollaboration tools such as Jira or Confluence\nExperience operating in airgapped or constrained environments\n** Position Requires an Active TS/SCI with a Full-Scope Polygraph Clearance**\nWhat You’ll Love About Synergist:\nWe have an actual “award-winning” culture\nWe recognize that our employees are our most valuable asset, and we do everything in our power to give you the best experience every day\nWe’re big believers in transparency and you’ll always have a seat at the table for big decisions and can even attend key business meetings if you want\nWe offer a best-in-class benefits package that we can customize for your unique needs. Offerings include healthcare, dental, and vision coverage for you and your dependents, flexible Paid Time Off (PTO) accrual, 10% 401k employer contribution, annual education allowance, and many perks from a robust referral program, weekly catered lunches, monthly happy hours, access to company suite at Cap One Arena, financial advisory, and so much more\nWe have a deep catalog (and growing!) of potential project opportunities for you to choose from, and when you’re ready for new challenges, we’ll work to make that happen\nWe Want to Hear From You If:\nYou have a BS in Computer Science or similar technical field (4 years of relevant experience may be substituted in lieu of a BS degree)\nYou have experience with AI Evaluation & Frameworks such as InspectAI, LLM Benchmarking, Agentic Workflow Evaluation (tool-use, multi-step task performance)\nYou have experience with Statistical Analysis, Data Cleaning & Validation, Data Visualization Tools, Jupyter NotebooksData & Environment: MongoDB, Cloud/Containerized Environments, Airgapped & Constrained Operations\nYou're familiar with Git Version Control, Experimental Design, Dataset Lineage/Traceability, Jira, and Confluence\nHere’s What You Can Expect Out Of This Role:\nAI Model Benchmarks — Design, implement, and maintain benchmark suites that measure model performance against defined evaluation criteria, including dataset curation, scoring, and results reporting.\nAgentic AI Benchmarks — Develop benchmarks and test scenarios that evaluate agentic workflows and multi-step task performance, including tool-use, task completion, and reliability metrics.\nEvaluation Dataset Development — Curate, validate, and version evaluation datasets and test scenarios, including any adversarial or edge-case sets needed to stress agent behavior.\nMetrics, Analysis, and Reporting — Produce scoring frameworks, visualizations, and written analyses that translate raw evaluation results into actionable findings for stakeholders.\nJob Type: Full-time\nPay: $100,000.00 - $280,000.00 per year\nBenefits:\n401(k)\nDental insurance\nEmployee assistance program\nFlexible schedule\nFlexible spending account\nHealth insurance\nHealth savings account\nLife insurance\nPaid time off\nParental leave\nProfessional development assistance\nReferral program\nTuition reimbursement\nVision insurance\nSecurity clearance:\nTop Secret (Required)\nWork Location: In person","datePosted":"2026-08-03T18:12:19.945Z","dateModified":"2026-08-03T18:12:19.945Z","hiringOrganization":{"@type":"Organization","name":"Synergist Computing","sameAs":"https://jobsearcher.com"},"jobLocation":{"@type":"Place","address":{"@type":"PostalAddress","addressLocality":"Annapolis Junction","addressRegion":"MD","addressCountry":"US"}},"identifier":{"@type":"PropertyValue","name":"JobSearcher","value":"f79ac65e2b2893eb8b2de3ec"},"url":"https://jobsearcher.com/jobs/f79ac65e2b2893eb8b2de3ec"}}