Artificial Intelligence (AI) Engineer, Tencent (AI Language Model Evaluator)
Performed evaluation of AI language model outputs across tasks including text generation, translation, summarization, and question answering. Assessed coherence, fluency, overall quality, and bias presence in generated text. Provided structured judgments to help ensure outputs align with human expectations and high accuracy standards. • Rated/assessed generated text quality (coherence and fluency) • Reviewed outputs for bias and potential ethical issues • Evaluated performance across multiple NLP task types • Supported iterative improvements through quality evaluation