Independent AI Evaluation & Prompt Engineering Projects
No description provided.
Hire this AI Trainer
Sign in or create an account to invite AI Trainers to your job.
No software listed
Independent AI Evaluation & Prompt Engineering Projects (AI trainer/evaluator for LLM responses). Brings 4+ years of professional experience across complex professional workflows, research, and quality-focused execution. AI-training focus includes data types such as Text and labeling workflows including Evaluation, Rating, and Prompt + Response Writing (SFT).
No description provided.
Delivered independent AI evaluation and prompt engineering projects focused on assessing and refining AI-generated outputs. Evaluated responses for accuracy, relevance, completeness, and safety while applying critical thinking and quality standards. Refined prompts and created structured instructions and test scenarios to improve task completion and reliability. • Evaluated AI-generated responses for accuracy, relevance, completeness, and safety • Refined prompts to improve output quality and task completion • Compared multiple AI responses and identified strengths and weaknesses • Performed research-based verification and created structured test scenarios
Organized and classified information using consistent evaluation criteria for downstream AI use. Applied categorization to support content analysis and structured information organization. Maintained taxonomy-style organization to ensure uniform labeling of items and content fragments. • Information organization and categorization • Consistent application of evaluation criteria • Content analysis support through structured labels • Structured classification for AI quality workflows
Refined prompts and created structured instructions and test scenarios to improve model output effectiveness and task completion. Developed prompt optimization workflows for content generation and research-related business use cases. Produced consistent evaluation-oriented prompt variations to assess and enhance response quality. • Prompt refinement and iteration • Creation of structured instructions and test scenarios • Design of prompts for research and content generation • Using evaluations to guide prompt improvements
Evaluated generative AI outputs by rating quality, factual accuracy, relevance, completeness, and safety against provided criteria. Performed comparative response analysis by assessing multiple answers to identify strengths, weaknesses, and instruction-following behavior. Conducted research-based verification to validate factual claims and reduce misinformation in responses. • Accuracy and relevance scoring • Safety and guideline compliance checks • Comparative evaluation across multiple model responses • Fact checking using internet research
Degree not specified
Independent AI Evaluator & Prompt Engineer
Freelance Digital Marketing & Research Specialist