Please mention DailyRemote when applying
See how much of this job your resume covers, and what’s missing.
Want a recruiter to go through it line by line?
Get professional reviewQuestions interviewers often ask for this role, with sample answers.
Upload your resume and we draft a letter for this exact role, tailored to what it asks for.
You will review proposed coding tasks and programming environments to ensure they accurately reflect real-world engineering requirements. Additionally, you will provide technical feedback on the structure and difficulty of evaluation harnesses to improve test suite robustness.
We're running a paid study on the effectiveness of coding evaluation harnesses used to test AI agents. Our team is building a comprehensive suite of programming environments designed to measure agent performance accurately. This work ensures that AI coding assistants are evaluated against realistic, high-quality development scenarios.
You will walk through a series of proposed coding tasks and environments during an AI-moderated session. We will ask you to review the structure, difficulty, and realism of these programming challenges. You will provide technical feedback on the evaluation harnesses and suggest improvements to the test suites. The conversation will focus on ensuring these tasks accurately reflect real-world engineering requirements.
We are hiring experienced software engineers based in South Asia who have a strong background in building and reviewing complex codebases. We welcome backend developers, test automation engineers, and full-stack engineers who understand evaluation harnesses. Ideal candidates have hands-on experience verifying realistic programming tasks in professional environments.
Review proposed coding tasks and programming environments for realism
Provide technical feedback on the structure and difficulty of evaluation harnesses
Walk us through your approach to verifying complex code challenges
Suggest improvements to make the test suites more robust
Professional software engineer based in South Asia
Experience building or reviewing realistic programming tasks
Familiarity with coding evaluation harnesses or test automation
Comfortable discussing technical workflows in an interview setting
$35 per hour
Terac is building the world's largest pool of vetted human experts for AI. Researchers, AI labs, and product teams use Terac to recruit, screen, and pay study participants across industries, languages, and skill sets.
Learn more at terac.com or on YouTube at @jointerac.
Stop the endless job search. Our AI finds and applies to the best jobs for you.
Featuring 213,952+ Jobs in Software Engineer
Answer easy questions
213,952+ jobs across 15+ categories
Get your best job matches
Only hand-screened, legit jobs
Find a remote job faster
No ads, scams, or junk
“I was the first applicant for a remote marketing position that got listed on the company website the same day I applied. Had an interview within 48 hours!”