For employers

Hire this AI Trainer

Sign in or create an account to invite AI Trainers to your job.

Invite to Job
S
Shuo Y.

Shuo Y.

Multilingual Data Annotation Expert

Malaysia flag吉隆坡, Malaysia

Key Skills

Software

No software listed

Top Subject Matter

Technology / SaaS — 12 years PM experience building software products
AI / Machine Learning — speech corpus annotation
familiar with AI training data workflows
E-commerce / Internet Platforms — product management in e-commerce merchant platforms
content categorization systems

Top Data Types

AudioAudio

Top Task Types

SegmentationSegmentation
ClassificationClassification
Data CollectionData Collection
Audio RecordingAudio Recording
TranscriptionTranscription

Freelancer Overview

While I do not have direct data labeling experience, 12 years in product management have honed my abilities in information classification, attention to detail, and standards-driven thinking — for instance, I designed product categorization systems and content moderation rules for an e-commerce platform, which closely parallels taxonomy building and quality consistency in data annotation. My EE background from Tsinghua enables me to quickly grasp data requirements from algorithm teams, and my proficiency in Chinese, English, Japanese, and German makes me well-suited for multilingual corpus tasks.

Labeling Experience

I have hands-on experience in speech data annotation for AI training

I have hands-on experience in speech data annotation for AI training. I independently produced a bilingual Chinese-English short sentence comparison corpus as a portfolio project for an xAI audio trainer position. The work involved end-to-end speech data pipeline tasks: recording 10 phonetically diverse short sentences (covering declarative, interrogative, exclamatory, imperative, and negative sentence types in both languages) at professional-grade specifications (48kHz, 16-bit WAV), and performing dual-tier time-aligned annotation using Praat — a phoneme tier with 50 precisely segmented intervals (pinyin/English phonemes with boundary timestamps at millisecond resolution) and a word tier with 21 intervals mapping Chinese characters and English words to their corresponding audio segments. All silence intervals were systematically labeled. I documented the full recording environment parameters (noise floor ~25dB, mic distance 15-20cm, recording level -6dB to -3dB), identified and root-caused a Bluetooth audio latency issue (~500ms-1s delay) that interfered with annotation playback accuracy, and established a best-practice guideline to use wired headphones for phonetic annotation workflows. This project demonstrates practical proficiency in speech corpus construction, phoneme-level segmentation, bilingual alignment, annotation tooling (Praat/TextGrid), and quality control documentation.

Not specified

Education

T

Tsinghua University, Department of Electronic Engineering: M.S

Tsinghua University, Department of Electronic Engineering: M.S. 2010–2013 B.S. 2006–2010

Tsinghua University, Department of Electronic Engineering: M.S. 2010–2013 B.S. 2006–2010
Not specified

Work History

C

Company not specified

Led B2B SaaS platform product planning and data-driven optimization, built product metrics frameworks with SQL, and auth

Location not specified
Not specified