Independent AI Technical Evaluator (Contract) — Advanced Tech Tasking Platforms (Remote)
Independently evaluates AI-generated Python scripts by performing structured code review to verify correctness, edge-case constraints, and algorithmic complexity. Applies highly structured grading rubrics and scientific criteria to multi-step logic workflows to maintain consistent audit accuracy. Uses Python-based validation methods with numerical and simulation checks to stress-test LLM outputs and ensure logical alignment with scientific integrity. • Performs deep-dive evaluations of AI-generated code for absolute structural correctness • Assesses edge cases and Big O complexity as part of quality scoring • Builds and applies validation scripts using NumPy and Pandas for scientific calculations • Collaborates asynchronously with a global engineering and academic community to refine evaluations