Record natural, expressive Southern American English for text-to-speech and speech model training. This flexible U.S. contract role pays $50-$100 per hour and suits experienced voice performers with quality recording equipment.
Audio & Speech
Remote Hourly · $50–$100/hr
$50–$100/hr
Compensation
1 country
Eligibility
Entry
Experience
Jul 10, 2026
Posted
Open to applicants in
United States
Interested in this role?
Create a free OpenTrain account and apply in minutes.
OpenTrain AI is the hiring and contracting organization for this role. OpenTrain is the #1 platform for finding and building careers in AI training and data labeling, helping people discover projects, build a professional profile, and apply in minutes. Creating an OpenTrain account is free.
About AI Speech Training
AI training is the human side of building artificial intelligence. For speech and text-to-speech systems, contributors provide carefully recorded examples that help models learn pronunciation, pacing, tone, emotion, and natural conversational delivery. This work gives voice professionals a direct role in shaping how advanced AI systems communicate.
The Role
OpenTrain is recruiting a Southern American English Voice Recorder to create high-quality voice samples for text-to-speech and speech model training and evaluation. You will record conversational, narrative, and instructional scripts while maintaining consistent accent, pronunciation, tone, pacing, and delivery across sessions.
This is an entry-level contract opportunity with a specialized voice-recording focus. The role is part time, requires less than 20 hours per week, and is available to candidates currently located in the United States.
Contractor and part-time engagement
Less than 20 hours per week
United States location required
Pay range: $50-$100 per hour
Work focused on audio recording for AI speech systems
What You'll Do
You will follow detailed recording guidelines and produce clear, natural, expressive speech across varied scripts. Consistency and precision matter, while some assignments may require different emotional styles or delivery variations.
Record clear, natural, expressive speech for AI speech model training and evaluation
Deliver consistent Southern American English vocal performance across recording sessions
Follow instructions for the recording environment, microphone setup, and file formatting
Record multiple takes when variations in emotion, emphasis, or style are requested
Perform conversational, narrative, and instructional scripts naturally and accurately
Requirements
Applicants must have a native Southern American English accent with strong clarity and be able to follow scripts precisely while sounding natural. Professional or near-professional recording quality is essential.
Native Southern American English accent
Current location in the United States
Experience in voice acting, dubbing, narration, podcasting, or broadcasting
Professional or near-professional recording equipment
A quiet recording space suitable for high-quality audio capture
Strong intonation, diction, and emotional range
Ability to deliver natural, expressive reads across varied scripts
Helpful Background
Experience with speech technology or audio production can help you contribute effectively, although the core requirements are strong vocal performance, clear delivery, and reliable recording quality.
Prior work on text-to-speech projects or AI voice datasets
Experience recording audiobooks or IVR systems
Familiarity with Audacity, Adobe Audition, Reaper, or similar audio editing tools
Comfort delivering conversational, corporate, energetic, and calm vocal styles
Why Join AI Training Work
AI training and data-labeling work is a fast-growing part of the technology industry. Contributors help prepare and evaluate the examples modern AI models learn from, and many projects offer flexible, remote-friendly ways to build experience around other commitments.
Contribute directly to the development of cutting-edge speech AI
Use your voice and performance expertise in a growing technology field
Choose a part-time workload of less than 20 hours per week
Build experience in text-to-speech and AI voice data production
Apply Through OpenTrain
Create a free OpenTrain account to build your AI training profile and apply in minutes. If selected for this contract role, you will use your voice-recording experience and equipment to help produce reliable speech data for AI systems.
Review the role and submit your application through OpenTrain
Highlight your voice, narration, broadcasting, or audio production experience
Show that you meet the accent, location, equipment, and recording-space requirements
Earn $42 per hour recording 100–150 short English sentences on an Android phone. This one-time project is open to genuine native UK or US English speakers in Great Britain and the United States.
Create clear, natural US English voice recordings that help train next-generation AI systems. Work remotely as a part-time contractor for $10-$30 per hour while applying your audio and performance skills.
Record natural, expressive Australian English for text-to-speech training and evaluation. This flexible contract role offers fewer than 20 hours per week and a listed rate of $50-$100 per hour.