In my recent project, Fixora (a RAG-powered Biomedical AI Assistant), I gained extensive experience in data preparation
In my recent project, Fixora (a RAG-powered Biomedical AI Assistant), I gained extensive experience in data preparation and curation for Large Language Models. I processed, cleaned, and structured raw, unstructured biomedical equipment OEM manuals into 13,105 semantically meaningful chunks. This involved carefully categorizing the data to ensure 100% isolation purity across 32+ device models, and embedding it into a FAISS vector knowledge base to accurately train and ground a local Qwen 2.5 LLM for diagnostic troubleshooting.