Advancing AI Through Prompt Engineering

via Freelancer ·

Budget / SalaryHourly project
TypeFreelance project
LocationRemote
Posted1 hour ago
Project SEAL — Handshake AI
Platform: Handshake AI
Project: Project SEAL
Type of work: Advanced/adversarial AI prompt and web-research evaluation.
Main objective: Create difficult prompts that can expose weaknesses in AI models’ web search, research, reasoning, and source-grounding abilities. Public descriptions from people working on SEAL describe it as moving beyond simply rating an existing answer and instead designing challenging research questions.
Pay: Public reports currently indicate about $20 per completed task for the U.S./general SEAL project. One recent independent account specifically lists SEAL at $20/task.
Task structure: It is per-task rather than hourly, so the amount of time required for a task can vary substantially. Workers have reported that some tasks can take considerable time because the goal is to construct a prompt that actually causes a model failure.
Skills involved: Research, prompt engineering, fact checking, web searching, source evaluation, logical reasoning, and identifying subtle model failures.
Onboarding: There is an assessment before full participation. Public reports indicate an 80% cutoff has been used for at least some SEAL onboarding assessments.
Model-failure component: A major part of the work is getting an AI model to make a meaningful error rather than simply producing a normal correct response. Recent workers specifically discuss the challenge of creating valid model failures.
Confidentiality: Handshake participants have been warned not to publicly share specific project/task details because of confidentiality obligations.
What this means for you

Given your data-science background, SEAL is considerably closer to analytical/research work than a basic data-labeling project. The valuable skill is not just writing a clever prompt; it's constructing a research question with enough complexity that an AI system can plausibly fail, then documenting why the failure is actually meaningful.
research technical writing report writing research writing ai model development ai ethics ai quality assurance ai training data
Apply on Freelancer →

Project sourced from Freelancer.com. Applications happen directly on the original platform — we never collect your data.