AI Content Evaluation & Localization (independent/ongoing evaluation)
Evaluated large volumes of AI-generated English and Chinese responses for quality and compliance using structured criteria. Assessed fluency, accuracy, factual consistency, cultural appropriateness, instruction following, usefulness, and safety to produce quality outcomes. Compared alternative AI outputs and identified linguistic, translation, and reasoning problems to support improvements. • Fluency and accuracy checks for bilingual responses • Factual consistency and cultural relevance evaluation • Safety and instruction-following assessment • Prompt and response quality comparison using AI tools