AnmeldenApp laden

Jobs › Stellenangebot

AI Response Evaluator

iMerit Technology · France, Japan, Turkey, Vietnam, Mexico, Norway

Freie MitarbeitRemoteEnglisch

Über die Stelle

Aus der Anzeige des Arbeitgebers · iMerit Technology · veröffentlicht am 11. September 2026

iMerit is looking for detail oriented analysts to evaluate and rank AI generated responses to image based prompts. You will judge answers on accuracy, relevance, clarity, conciseness, safety, localization, and how well they follow the user's instructions, then explain your reasoning in writing. Much of the job comes down to this: look at the image, look at what the model said about it, and decide whether the two actually match.

What you will do

-

Interpret conversational context and identify what the user really wanted

-

Rate and rank responses against defined quality criteria

-

Compare multiple answers and explain in writing why one wins

-

Verify factual claims using approved research sources

-

Flag tasks that cannot be reliably assessed rather than guessing

What you bring

-

Strong critical thinking and sound judgment in ambiguous cases

-

Solid research skills and attention to detail

-

Excellent reading comprehension

-

Self direction and the discipline to hit deadlines without supervision

Good to know

-

Independent contractor engagement for the length of the project.

-

Fully remote and flexible.

-

You choose your hours as long as volume and deadlines are met.

-

Task volume varies with project demand.

Genannte Fähigkeiten

AI/MLSOLID

Sieh deinen Match-Score für jede Stelle

BabZituna bewertet jede Stelle nach sechs echten Kriterien gegen dein Profil und zeigt dir, WARUM sie so abschneidet, auf Fairness geprüft (das öffentliche Bias-Audit lesen).

App laden → ✓ 100 % kostenlos für Jobsuchende
Wie eine Stelle abschnitt Beispiel
Fähigkeiten96Erfahrung90Standort84Arbeitsweise74Anstellungsart61GehaltKeine Daten

Beispielwerte, kein echter Kandidat. Jedes Kriterium wird aus deinem eigenen Profil mit bis zu 100 Punkten bewertet, und eines, das wir nicht messen können, sagt das, statt zu raten.