Professionals: Designing Challenging AI Prompts
AI Summary
Designs demanding AI prompts by translating real-world workflows into tasks that expose model limitations, then writes grading rubrics and evaluates AI outputs.
About this role
What We're Researching
We're running a paid study to build a bench of people who are exceptionally good at designing tasks that expose AI model limitations. By turning real-world workflows into demanding requests, we can better evaluate where current models break down. This initial trial helps us identify individuals suited for ongoing prompt engineering and evaluation work.
How It Works
You will spend about an hour translating a complex workflow from your job or personal life into a demanding prompt that requires reasoning and real-world lookup. After running it in ChatGPT to identify where the model fails, you will refine the prompt until it breaks the system. Finally, you will write a clear grading rubric that a stranger could use to evaluate any AI's attempt at your task. This entire process is screen-recorded, as we are assessing your thought process just as much as the final submitted files.
Who This Is For
We welcome professionals, domain experts, and power users who have deep knowledge of specific workflows. You need to be capable of evaluating an AI's output within seconds and comfortable working on a laptop or desktop with a ChatGPT account. Candidates who excel at this trial will be considered for a long-term bench of evaluators.
What You'll Do
Pick a familiar workflow and convert it into a demanding AI prompt
Test your prompt in ChatGPT to find failure points, making it harder if the AI succeeds
Write a comprehensive rubric for grading the AI's performance
Share your screen, camera, and microphone while completing the task
Submit your prompt, failure notes, rubric, and the generated output file
Who Should Apply
Deep familiarity with a specific professional or personal workflow
Ability to quickly evaluate the accuracy and quality of AI outputs
Access to a laptop or desktop computer
An active ChatGPT account
Comfortable being screen-recorded while thinking through complex tasks
Compensation
$20 one-time
Ready to participate?
About Terac
Terac is building the world's largest pool of vetted human experts for AI. Researchers, AI labs, and product teams use Terac to recruit, screen, and pay study participants across industries, languages, and skill sets.
Learn more at terac.com or on YouTube at @jointerac.
Skills
Explore related jobs
More jobs at terac
Personal Trainers: Paid Study on Home Equipment WorkflowsUnited States
Software Engineers: Paid Interview on AI Evaluation TasksUnited States
Transaction Finance Professionals: AI Output Annotation and RankingUnited States
Management Consultants: Paid AI Output EvaluationUnited States
Consumers: 15-Minute Interview on Everyday ExperiencesUnited States
Dermatologists: Evaluation of AI Diagnostic OutputsUnited States
Browse these categories
Market data for this role
All reports →- SeriesRole reportsOne role family at a time: how many openings, what changed this week, who is hiring, what it pays.
- SeriesSalary reportsWhat employers publish in job postings, by level and workplace. Not self-reported pay.
- Market overviewState of tech hiring, September 2026: up 4.8%Tech hiring rose 4.8% month over month in September 2026, with 411,122 new listings. Customer support and account executive roles led the growth.