What you will do
You will read multilingual conversation prompts and label them for an AI-safety study. This is blind labelling: you assign your own judgments without seeing creator labels or other annotators’ answers.
Multilingual Evaluation Annotation Platform
Sign in to continue annotating.
Reset your password
Enter your email and we'll send a reset link.
Set a new password
Enter your new password below.
Create your account
Complete your profile information first, then create your login credentials.
Informed consent
Please read these points carefully. You can stop participating at any time.
You will read multilingual conversation prompts and label them for an AI-safety study. This is blind labelling: you assign your own judgments without seeing creator labels or other annotators’ answers.
Prompts are designed to probe harmful AI behaviour and may include political, religious, or otherwise sensitive material. If you need to stop, close the session — there is no penalty.
Participation is voluntary and unpaid. You are not obliged to finish every item.
Stage 1 (Validation) completers are acknowledged. Stage 1 + Stage 2 (Evaluation) completers with good-quality work may be considered for co-authorship.
Annotator ID:
Your submitted annotations across all prompts.
Answer every required field. Use Skip only if you cannot annotate this item.
Your independent label. Creator categories stay hidden and are stored separately for later comparison.
Does understanding the harm require culturally specific knowledge (names, norms, local politics, religion), or is it universal?
Does the text sound natural in this language/dialect, or like awkward machine translation?
Is the conversation grounded in the target culture (setting, references, social context), not a generic English idea pasted in?
Do the user turns build naturally toward a harmful ask, or does the jump feel forced / broken?
Rate potential real-world harm if an AI complied. Do not average — pick one level.
Anything a reviewer should know — edge cases, uncertainty, dialect notes.
This prompt includes a second, intended-harmless conversation. Compare dialect/register and confirm it is clearly harmless.
Write a realistic multi-turn user conversation for AI-safety evaluation. An admin must approve it before it enters the annotation queue.
Stored on the prompt only. Annotators never see this — they assign their own category later.
Randomly selected from existing prompts in the dataset.
Review prompts, manage users, and monitor annotation quality.
Starting prompts are pre-approved seed data. Only user-created submissions can be approved or rejected.
| ID | Segment | Language | Harm category | Turn 1 + English | Status | Annotations | Created | Actions |
|---|
No prompts match the current filters.
| Name | Role | Language | Annotator ID | Joined | Actions |
|---|
| Prompt ID | Annotator ID | Creator harm | Annotator harm | Severity | Ling. Auth | Cult. Spec | Timestamp | Actions |
|---|
Soft-deleted users, prompts, and annotations are kept here. Nothing is permanently removed. Restore anytime.
There are no more prompts available for your language right now.
Connecting to server…
MEAP Guide