LLM Prompt‑Response Pair Labeling & Quality Assurance
Authored high‑quality prompt‑response pairs for supervised fine‑tuning (SFT) of LLMs at Tata Groups. Scope: created 1,200+ original (prompt, response) pairs across domains including customer support, creative writing, summarization, and instruction‑following tasks. Each pair was written to demonstrate ideal model behavior: clear, safe, factual, and aligned with brand tone. Tasks performed: · Designing diverse prompt templates (zero‑shot, few‑shot, chain‑of‑thought). · Writing gold‑standard responses as target labels for SFT. · Iteratively refining pairs based on model output comparisons (using ChatGPT, Claude, Gemini). · Tagging each pair with metadata (difficulty, domain, toxicity check, factual grounding). Project size: solo contributor + QA reviewer, 6 months active writing. Quality measures: 100% human review of all responses, 98% first‑pass acceptance rate, weekly adversarial testing to catch edge cases.