AI Training & LLM Evaluation & RLHF / Prompt Engineering (Independent AI Workflow Practice)
Provided hands-on LLM evaluation with an emphasis on RLHF-style quality checks and iterative prompt improvements. Focused on identifying output issues including hallucinations, logical inconsistencies, factual inaccuracies, and formatting problems. Used structured evaluation guidelines to score and validate responses against expected behavior. • Performed response quality evaluation aligned to instruction-following requirements • Checked factuality and internal logic for mathematical/physics reasoning tasks • Validated formatting and adherence to specified output structures • Repeated quality-focused review cycles to ensure consistency