AI Data Specialist
Model Training & Alignment: Actively train and refine next-generation large language models (LLMs) using advanced techniques, including Reinforcement Learning from Human Feedback (RLHF) and Supervised Fine-Tuning (SFT). Prompt Engineering & Evaluation: Design, test, and iterate on complex prompts to assess and improve model capabilities in areas like instruction-following, reasoning, and creativity. AI Safety & Ethics: Evaluate model outputs for safety, truthfulness, and harmlessness, ensuring alignment with legal, ethical, and policy guidelines. Quality Assurance: Perform rigorous quality assurance on data and feedback produced by other analysts, upholding high standards for model training datasets.