TL;DR
Improve Large Language Models' performance by labelling, ranking, auditing, and correcting model output.
Key Responsibilities Evaluate and rank model outputs Stress-test and break models Create datasets Build and apply rubrics and taxonomies Annotate and correct multimodal data Calibrate and maintain consistency Adapt to experimental work Report on model performance Requirements 1+ years of experience in AI data annotation, LLM evaluation, content moderation, research, or a related analytical role Experience applying detailed guidelines to complex and often ambiguous content Comfort with ambiguity and a sharp eye for inconsistencies and model failure modes Excellent command of written English and strong reading comprehension Strong attention to detail and commitment to accuracy Comfort working with annotation platforms and structured formats Strong execution in a remote environment
Independent Contractor Agreement