For employers

Hire this AI Trainer

Sign in or create an account to invite AI Trainers to your job.

Invite to Job
K
Kyle

Kyle

AI Model Evaluation Analyst

USA flagFarmington, Usa

Key Skills

Software

Don't disclose

Top Subject Matter

Autonomous systems safety validation and visual model evaluation
Model Alignment
Safety/Reliability evaluation

Top Data Types

ImageImage
TextText

Top Task Types

Question AnsweringQuestion Answering
ClassificationClassification
RLHFRLHF

Freelancer Overview

AI Model Evaluation Analyst at SparkAI. Brings 5 years of professional experience across complex professional workflows, research, and quality-focused execution. Core strengths include Data Annotation, Image Annotation, and RLHF. AI-training focus includes data types such as Image and Text and labeling workflows including Evaluation, Rating, and Video Annotation.

Labeling Experience

AI Model Evaluation Analyst at SparkAI

ImageImageBounding BoxBounding Box

Evaluated model accuracy for live visual data streams from autonomous hardware to ensure correct obstacle identification and safe behavior in real-world operations. Trained and assessed multiple AI models under different logic and safety rules, adapting evaluation to each model’s project requirements. Maintained high quality targets and provided audit-focused insights to improve overall model performance and reliability. • Measured and validated model performance against accuracy and safety standards for autonomous obstacle detection. • Trained six separate AI models with distinct behavior rules, ensuring no cross-contamination of safety logic. • Sustained a 99% accuracy rate over approximately three and a half years while supporting ongoing reliability improvements. • Contributed insights that supported implementation of an internal audit system to increase team accuracy.

2022 - Present

Model Alignment Specialist

Don't discloseTextTextRLHFRLHF

Performed model evaluation and alignment work for large-scale language models, focusing on reasoning quality, safety, and reliability for training use. Conducted head-to-head comparisons of advanced model outputs to select and refine optimal responses for downstream training. Implemented test cases and entity tagging to identify weaknesses and improve model contextual understanding of visual and dialogue content. • Compared generated responses across advanced AI models to select and refine the best output for training. • Developed complex test cases to expose model weak points and drive more reliable answers. • Performed precise entity tagging, including annotating reference images and visual elements for improved context and recognition. • Developed refined training dialogues to better connect visual context with natural language descriptions of video sequences.

2024 - 2026

Education

H

High School Diploma

Degree not specified

Not specified
Not specified

Work History

S

SparkAI

AI Model Evaluation Analyst

Farmington
2022 - Present
N

Not Disclosed at This Time

Model Alignment Specialist

Farmington
2024 - 2026