Skip to content

Choose the judgment before configuring the job

An Experiment type defines what one reviewer sees and submits. GenMedia separates this protocol from the Dataset and from the Job that staffs or generates work. Launchable types are validated end to end; planned catalog entries remain visibly unavailable until their evaluator, result method, and controls are implemented.

Experiment type catalog
TypeReviewer actionCurrent state
Pairwise image, video, or audioChoose between two blind conditions, optionally with reference mediaAvailable
ACR image, video, or audioRate one condition on an absolute category scaleCatalog only
Rubric image, video, or audioRate one condition across explicit criteriaCatalog only
MUSHRAGrade several audio conditions against reference and anchorsCatalog only
Rank audioOrder several audio conditionsCatalog only
Real or generatedMake one binary authenticity judgmentCatalog only
SurveyAnswer structured questions after inspecting mediaCatalog only

Type, Data, Experiment, and Job

  1. 1

    Type: choose the media modality and judgment protocol.

  2. 2

    Data: bind the main immutable Dataset version and optional Reviewer Practice and Golden Control versions.

  3. 3

    Experiment: set the reviewer-facing title, question, instructions, criteria, reference behavior, and playback controls.

  4. 4

    Job: choose the saved rater cohort, session target, comparisons per Session, sampling plan, budget, and launch state.

Pairwise video controls

  • Presentation

    Flip between candidates or show them side by side; sequential or synchronized playback is explicit.

  • Decision

    Allow or disallow ties and skips without changing the blind candidate order.

  • Reference

    Show a separate reference when fidelity matters, and decide whether a reference condition is excluded from candidate sampling.

  • Playback

    Set minimum display and play duration, audio state, loop, seeking, fit or crop behavior, and an optional transition mask.

  • Markers

    Render Dataset marker annotations on supported audio and video controls.

Job settings do not redefine the Experiment

A Job controls how many Sessions to collect and how comparisons are sampled. It does not change the Dataset, question, or scoring semantics. Starting another Job under the same frozen task preserves the Experiment contract and records a separate operational run.

Generation jobs are provider requests that produce candidates. Evaluation staffing jobs allocate reviewer Sessions. They are different operational records even when the interface presents both under one Evaluation Task.

Why unsupported cards stay disabled

A familiar name is not enough to ship a protocol. MUSHRA, ranking, authenticity, and survey tasks need specific assignment rules, reviewer controls, exports, and statistical interpretation. GenMedia exposes their intended place in the catalog without allowing an operator to launch a task whose evidence pipeline is incomplete.