Browse categories
Explore
Fiverr Pro
English
$
USD
I help AI companies and ML teams produce high-quality labeled data and structured evaluations for model training and benchmarking.
My experience includes building evaluation rubrics and scoring criteria for LLM outputs, annotating data across specialized domains (healthcare, finance, real estate, education), and conducting side-by-side response comparisons with clear pass/fail reasoning.
What I offer:
I work methodically and communicate clearly if requirements are ambiguous accuracy over speed, always.
Technique:
Manual
Tagging type:
Text
•
Image
•
Video