AI Benchmark Researcher
Location: Remote
Compensation: Piece Work
Reviewed: Fri, Sep 18, 2026
This job expires in: 30 days
Job Summary
Experienced professionals will find a full-time opportunity as an AI Benchmark Researcher, responsible for developing challenging tasks for AI benchmarks by identifying complex problems, translating them into Linux terminal tasks, and designing challenges that test the limitations of advanced AI agents.
Key responsibilities
- Identify difficult and meaningful problems in the researcher's area of expertise
- Translate identified problems into tasks solvable in a Linux terminal
- Design challenges that effectively expose limitations in advanced AI models through rigorous testing
Required qualifications
- A completed master's degree or higher; candidates with a bachelor's degree and 10 years of relevant experience may be considered
- At least 10 years of professional experience in a technical domain
- Expertise in coding, machine learning, systems, cybersecurity, or hardware
- At least one academic or professional publication
- Practical comfort working in a Linux terminal
Complete Job Description
The complete job description is available to members. Premium membership includes:
Full access to 41,641 remote jobs from human-vetted companies, updated daily
Resume Builder - AI-powered tool to craft, enhance, and tailor your resume to a specific job
Twice-monthly live group coaching and the full Remote Career Center
20% member discount on Career Services
Backed by a 30-day money-back guarantee