LLM Evaluator / AI Training Intern — Ethara AI (Remote)
This role involved evaluating and ranking AI-generated text responses for quality, relevance, factual accuracy, and adherence to instructions. Tasks included prompt analysis to identify hallucinations, inconsistencies, and low-quality outputs and participating in RLHF-based evaluation for Large Language Model (LLM) training. Structured guidelines were followed to maintain annotation consistency and provide quality feedback for model improvement. • Compared multiple AI responses and provided detailed, actionable feedback. • Performed structured data annotation and review of LLM-generated outputs. • Maintained high attention to detail and evaluation consistency throughout projects. • Worked remotely and coordinated with team members to execute evaluations efficiently.