AI Engineer / LLM Training Specialist (Expert data labeling, prompt/response & code evaluation, task artifact development)
Provided expert data labeling and evaluation for LLM training focused on software engineering, code generation, debugging, reasoning, tool use, and agentic workflows. Assessed model-generated prompts, responses, code diffs, unit tests, logs, and execution results using strict technical rubrics. • Evaluated frontier LLM behavior across coding and reasoning scenarios • Identified failure modes including hallucinations, missing edge cases, and architectural weaknesses • Performed senior-level checks on correctness, maintainability, security, performance, and debugging strategy • Developed task artifacts such as programming tasks and validation materials