Innodata Inc.
AI Voice Evaluation Specialist
Remote - Alabama · Remote - Alaska · Remote - Arkansas · Remote - Delaware · Remote - Florida · Remote - Georgia · Remote - Indiana · Remote - Maryland · Remote - Massachusetts · Remote - New Hampshire · Remote - New Jersey · Remote - New York · Remote - South Carolina · Remote - Texas
Get past the screening software and onto a recruiter's desk
hirly rewrites your resume for this job — matching the keywords and skills in the posting, moving your most relevant experience to the top, and writing a cover letter to fit. About 30 seconds.
- Keywords matched to this posting
- Fit score before you apply
- Cover letter included
Matched against 2.3M live jobs from 200,000+ employers in 200+ countries.
Tailor my resume for this job →hirly's read of this role
- Seniority
- Mid level
- Country
- US
- Work mode
- Remote-friendly
- First seen by hirly
- 28 Sept 2026
Derived automatically from the posting. Upload your resume above to see how the role scores against it.
the posting
Innodata (Nasdaq: INOD) is a global data engineering company. We believe that data and Artificial Intelligence (AI) are inextricably linked. Our mission is to enable the responsible advancement of artificial intelligence by providing the data, evaluation frameworks, and human expertise required to build AI systems that can be trusted at scale. We provide a range of transferable solutions, platforms, and services for Generative AI / AI builders and adopters. In every relationship, we honor our 36+ year legacy delivering the highest quality data and outstanding outcomes for our customers.
Scope of the Role:
We are looking for detail-oriented Voice Specialists to evaluate the performance of AI models through real-time, voice-based conversations. In this role, you will interact with two different AI models using the same assigned scenario, carefully compare their responses, and provide structured evaluations based on defined quality criteria.
The goal of the evaluation is to create a fair, consistent comparison between models and identify which model delivers the stronger overall conversational experience.
What You’ll Own:
Review an assigned scenario and roleplay a natural conversation with two different AI models .
Conduct comparable conversations with Model A and Model B, keeping the scenario, conversational approach, and overall interaction as consistent as possible.
Maintain approximately the same number of conversational turns with each model to support a fair comparison.
Record your voice during each interaction and pay close attention to the quality and behavior of the model's responses.
Evaluate each model across five defined evaluation dimensions .
Identify and categorize relevant error clusters that may occur in the audio or conversation.
Compare the performance of Model A and Model B based on the evaluation criteria and your observations.
Select the model that demonstrated the stronger overall performance according to the defined evaluation framework.
Write a detailed rationale explaining your final preference, referencing specific examples and observations from both conversations.
Apply evaluation guidelines consistently across different scenarios and model interactions
You’ll Thrive in This Role If You Have
Bachelors Degree
Strong attention to detail and ability to notice subtle differences in conversational quality.
Excellent listening and comprehension skills.
Strong written communication skills, particularly the ability to explain observations clearly and objectively.
Ability to follow detailed evaluation guidelines consistently.
Comfort speaking naturally and roleplaying different conversational scenarios.
Ability to compare two interactions fairly without allowing personal preferences to influence the evaluation.
Strong critical-thinking and analytical skills.
Reliability and consistency when completing structured evaluation tasks.
Familiarity with AI assistants, voice-based AI, or conversational systems
The expected hourly salary range for this position is $20-$27 p/hour, based on experience, skills, and qualifications.
Please be aware of recruitment scams involving individuals or organizations falsely claiming to represent employers. Innodata will never ask for payment, banking details, or sensitive personal information during the application process. To learn more on how to recognize job scams, please visit the Federal Trade Commission’s guide at https://consumer.ftc.gov/articles/job-scams.
If you believe you’ve been targeted by a recruitment scam, please report it to Innodata at [email protected] and consider reporting it to the FTC at ReportFraud.ftc.gov .
Similar jobs
- Test & Evaluation SpecialistAgile Defense · Fort Huachuca, AZFirst seen today
- RN Professional Practice Evaluation Specialist - Quality Mgmt - Full TimeKingman Regional Medical Center · Kingman, AZ, United StatesFirst seen 2d ago
- Impact Evaluation SpecialistImpact Genome · Washington DCFirst seen 3d agoremote
- Merchant Mariner Evaluation Specialist (Deck)U.S. Coast Guard · Juneau, Alaska, United StatesFirst seen 4d ago
- Merchant Mariner Evaluation Specialist (Engine)U.S. Coast Guard · Juneau, Alaska, United StatesFirst seen 4d ago
Browse similar roles
Want this one?
Upload your resume and hirly rewrites it for this job and writes the cover letter — in about thirty seconds, before you sign up.
Tailor my resume for this job