Self Project / AI Training (AI Training Specialist)
Conducted AI training by fine-tuning open-source large language models on custom datasets for improved instruction following and chatbot quality. Applied QLoRA and LoRA to perform parameter-efficient supervised fine-tuning and preference optimization. Evaluated outputs using both automated benchmarks and human judgment to refine model behavior. • Fine-tuned Llama-3 8B and Mistral-7B using QLoRA/LoRA • Performed SFT and Direct Preference Optimization (DPO) • Built domain-specific multilingual chatbots (Hindi + English) • Created and validated high-quality instruction datasets using MT-Bench and human evaluation