LLM Response Evaluation & AI Training Data Annotation
Worked on AI training data and LLM evaluation projects involving prompt-response assessment, ranking, reasoning validation, instruction-following checks, and factual accuracy review. Evaluated model outputs using detailed rubrics for coherence, safety, formatting compliance, usefulness, and technical correctness. Specialized in coding and reasoning-heavy tasks leveraging 6+ years of software engineering experience in JavaScript, TypeScript, React, Node.js, APIs, and system design. Maintained high annotation quality through consistency checks, edge-case analysis, calibration adherence, and structured feedback across large-scale datasets.