Professionals: Designing Challenging AI Prompts
我们正在研究
我们正在开展一项付费研究,以建立一个擅长设计能暴露AI模型局限性的任务的人才库。通过将现实工作流程转化为具有挑战性的请求,我们可以更好地评估当前模型的失效点。此次初步试验有助于我们找到适合长期进行提示工程和评估工作的人员。
如何运作
您将花费大约一个小时,将您工作中或个人生活中的复杂流程转化为需要推理和现实世界查询的具有挑战性的提示。在ChatGPT中运行后,您将识别模型的失败点,并不断优化提示,直到系统崩溃。最后,您将编写一个清晰的评分标准,让陌生人能够使用它来评估任何AI完成您任务的表现。整个过程将进行屏幕录制,因为我们不仅评估最终提交的文件,也评估您的思考过程。
适合谁
我们欢迎拥有特定工作流程深入知识的专业人士、领域专家和高级用户。您需要能够在几秒钟内评估AI的输出,并且熟悉使用配有ChatGPT账户的笔记本电脑或台式机。在此试验中表现出色的候选人将被考虑加入长期评估者人才库。
您将做什么
- 选择一个熟悉的流程并将其转化为具有挑战性的AI提示
- 在ChatGPT中测试您的提示,找出失败点,如果AI成功则使其更具挑战性
- 编写全面的评分标准来评估AI的表现
- 在完成任务时分享您的屏幕、摄像头和麦克风
- 提交您的提示、失败记录、评分标准和生成的输出文件
谁应该申请
- 对特定专业或个人流程有深入了解
- 能够快速评估AI输出的准确性和质量
- 拥有笔记本电脑或台式机
- 拥有活跃的ChatGPT账户
- 在思考复杂任务时舒适地被屏幕录制
报酬
20美元一次性报酬
准备好参与了吗?
立即开始您的付费面试 https://terac.com/interview/start/r/bac94240-9ba4-477a-bb8d-9b3b2af686d2?utm_source=ashby_listing_description
关于TERAC
Terac正在构建全球最大的经过验证的人类专家池,用于AI。研究人员、AI实验室和产品团队使用Terac在各个行业、语言和技能水平上招募、筛选和支付研究参与者。
了解更多信息请访问 terac.com https://terac.com 或在YouTube上关注 @jointerac https://www.youtube.com/
查看英文原文
WHAT WE'RE RESEARCHING
We're running a paid study to build a bench of people who are exceptionally good at designing tasks that expose AI model limitations. By turning real-world workflows into demanding requests, we can better evaluate where current models break down. This initial trial helps us identify individuals suited for ongoing prompt engineering and evaluation work.
HOW IT WORKS
You will spend about an hour translating a complex workflow from your job or personal life into a demanding prompt that requires reasoning and real-world lookup. After running it in ChatGPT to identify where the model fails, you will refine the prompt until it breaks the system. Finally, you will write a clear grading rubric that a stranger could use to evaluate any AI's attempt at your task. This entire process is screen-recorded, as we are assessing your thought process just as much as the final submitted files.
WHO THIS IS FOR
We welcome professionals, domain experts, and power users who have deep knowledge of specific workflows. You need to be capable of evaluating an AI's output within seconds and comfortable working on a laptop or desktop with a ChatGPT account. Candidates who excel at this trial will be considered for a long-term bench of evaluators.
WHAT YOU'LL DO
- Pick a familiar workflow and convert it into a demanding AI prompt
- Test your prompt in ChatGPT to find failure points, making it harder if the AI succeeds
- Write a comprehensive rubric for grading the AI's performance
- Share your screen, camera, and microphone while completing the task
- Submit your prompt, failure notes, rubric, and the generated output file
WHO SHOULD APPLY
- Deep familiarity with a specific professional or personal workflow
- Ability to quickly evaluate the accuracy and quality of AI outputs
- Access to a laptop or desktop computer
- An active ChatGPT account
- Comfortable being screen-recorded while thinking through complex tasks
COMPENSATION
$20 one-time
READY TO PARTICIPATE?
Start your paid interview now https://terac.com/interview/start/r/bac94240-9ba4-477a-bb8d-9b3b2af686d2?utm_source=ashby_listing_description
ABOUT TERAC
Terac is building the world's largest pool of vetted human experts for AI. Researchers, AI labs, and product teams use Terac to recruit, screen, and pay study participants across industries, languages, and skill sets.
Learn more at terac.com https://terac.com or on YouTube at @jointerac https://www.youtube.com/@jointerac.