SWE Coding Benchmarks & RL Environments
Posted 19 Sep 2026
Work
Remote · Task based
Experience
Mid level
Eligibility
No stated location restriction
Field
Software engineering
Skills
Software EngineeringProgrammingCoding Task EvaluationCoding Harness EvaluationAI Agent Evaluation
About this role
We're running a paid study to review and verify programming tasks used for testing AI agents. We need experienced software engineers to evaluate coding harnesses and assess the quality of realistic coding challenges.
Disclosure: Skillora is independent of Terac. The apply button carries our referral code, and if you are placed, Terac pays Skillora a fee. It does not change what you are paid, our ranking does not know which listings pay us, and selection is decided by Terac’s own vetting.