JOBSEARCHER

Remote | Research Engineer - Code Generation & Model Evaluation — $50–$100/hour

24MAGRemoteL5 SeniorOctober 4th, 2026
We are sharing a specialised part-time consulting opportunity for experienced Research Engineers with strong expertise in software engineering, competitive programming, code generation, debugging, and technical model evaluation to contribute to an advanced AI training and code-generation project.Selected professionals will work across diverse codebases and programming languages to debug issues, implement and evaluate technical solutions, improve code quality, and develop realistic coding tasks used to assess advanced AI systems. The work combines hands-on software engineering with model evaluation, technical annotation, open-source collaboration, and rigorous code review. No prior experience in AI is required.Key ResponsibilitiesCode Analysis, Debugging & Problem SolvingAnalyse, debug, and resolve technical issues across diverse software codebasesWork with languages including Python3, Java, Rust, C , Go, TypeScript, or comparable technologiesInvestigate bugs, implementation failures, incorrect behaviour, and edge casesApply strong algorithmic reasoning and data-structure knowledge to complex coding problemsEvaluate multiple possible solution paths and select appropriate technical approachesFeature Development & Codebase ImprovementContribute to the design and implementation of new features and technical enhancementsRefactor existing codebases to improve maintainability, clarity, and adaptabilityOptimise code for performance, efficiency, and reliabilityReview implementation decisions and identify opportunities for architectural or technical improvementProduce high-quality, well-structured code that can be tested and evaluated consistentlyCode Generation & Model EvaluationDevelop and review realistic coding tasks used to evaluate AI model capabilitiesAssess AI-generated code for correctness, completeness, performance, and technical qualityIdentify logic errors, implementation weaknesses, constraint violations, and incomplete solutionsValidate generated outputs against expected behaviour and objective technical requirementsHelp improve AI coding performance through rigorous technical assessment and feedbackTechnical Feedback & CollaborationAuthor clear, actionable feedback and annotations supporting AI model training and evaluationDocument technical decisions, best practices, implementation details, and problem-solving approachesCollaborate with technical stakeholders and open-source contributors to improve project qualityCommunicate complex engineering concepts clearly in written and verbal formContribute to transparent knowledge-sharing and consistent technical standards across project workflowsIdeal ProfileStrong expertise in competitive programming, coding problem analysis, or technically demanding software engineeringAdvanced proficiency in at least one of Python3, Java, Rust, C , Go, or TypeScriptSolid understanding of algorithms, data structures, complexity, and software-engineering fundamentalsProven track record of open-source contributions or participation in collaborative software projectsStrong analytical ability to interpret complex constraints and compare multiple technical solutionsExperience with debugging, feature implementation, code refactoring, or performance optimisationExcellent written and verbal communication skillsAbility to explain technical decisions and reasoning clearly and preciselyStrong attention to detail in code validation, correctness, and output consistencyComfortable working independently within a remote and collaborative environmentAbility to produce high-quality, well-documented work within demanding timelinesNo prior experience in AI research or model training is requiredEngagement DetailsPart-time independent contractor engagementFully remoteCompensation: $50--$100/hourCompensation is output-based, with payment made for tasks that meet project specificationsMinimum weekly submission requirements applyWork will involve debugging, feature implementation, codebase refactoring, performance optimisation, coding-task development, and AI model evaluationSelected professionals should be prepared to begin their first tasks within approximately 24--48 hours of completing onboardingRoles are typically filled within approximately 48 hoursProject workload, task complexity, and completion time may vary depending on experience and workflowWork must be completed without using confidential or proprietary information belonging to any employer, client, institution, or other third partyAbout the PlatformThis opportunity is available through 24-MAG LLC. We connect experienced professionals with remote consulting opportunities across technical, evaluation, and project-based workstreams.