at Terac
WHAT WE'RE RESEARCHING
We're running a paid study to build a bench of people who are exceptionally good at designing tasks that expose AI model limitations. By turning real-world workflows into demanding requests, we can better evaluate where current models break down. This initial trial helps us identify individuals suited for ongoing prompt engineering and evaluation work.
HOW IT WORKS
You will spend about an hour translating a complex workflow from your job or personal life into a demanding prompt that requires reasoning and real-world lookup. After running it in ChatGPT to identify where the model fails, you will refine the prompt until it breaks the system. Finally, you will write a clear grading rubric that a stranger could use to evaluate any AI's attempt at your task. This entire process is screen-recorded, as we are assessing your thought process just as much as the final submitted files.
WHO THIS IS FOR
We welcome professionals, domain experts, and power users who have deep knowledge of specific workflows. You need to be capable of evaluating an AI's output within seconds and comfortable working on a laptop or desktop with a ChatGPT account. Candidates who excel at this trial will be considered for a long-term bench of evaluators.
WHAT YOU'LL DO
WHO SHOULD APPLY
COMPENSATION
$20 one-time
READY TO PARTICIPATE?
Start your paid interview now https://terac.com/interview/start/r/bac94240-9ba4-477a-bb8d-9b3b2af686d2?utm_source=ashby_listing_description
ABOUT TERAC
Terac is building the world's largest pool of vetted human experts for AI. Researchers, AI labs, and product teams use Terac to recruit, screen, and pay study participants across industries, languages, and skill sets.
Learn more at terac.com https://terac.com or on YouTube at @jointerac https://www.youtube.com/@jointerac.