Freelance AI Data Annotation & LLM Evaluation (Remote)
Evaluated and ranked AI-generated responses by assessing accuracy, clarity, and relevance for real-world user needs. Identified logical errors, inconsistencies, and weak reasoning, then provided concise structured justifications to improve model outputs. Maintained quality standards and complied with project guidelines while meeting deadlines. • Assessed response quality for correctness and usefulness. • Ranked alternatives based on clarity and relevance. • Detected reasoning flaws, contradictions, and factual issues. • Wrote structured improvement feedback to guide better outputs.