Voice Specialist for AI Model Evaluation
Job Description
Key Skills Required
Master these to land this role
Want to know if you're a match for this job?
Scope of the Role:
We are looking for detail-oriented Voice Specialists to evaluate the performance of AI models through real-time, voice-based conversations. In this role, you will interact with two different AI models using the same assigned scenario, carefully compare their responses, and provide structured evaluations based on defined quality criteria.
The goal of the evaluation is to create a fair, consistent comparison between models and identify which model delivers the stronger overall conversational experience.
What You’ll Own:
- Review an assigned scenario and roleplay a natural conversation with two different AI models.
- Conduct comparable conversations with Model A and Model B, keeping the scenario, conversational approach, and overall interaction as consistent as possible.
- Maintain approximately the same number of conversational turns with each model to support a fair comparison.
- Record your voice during each interaction and pay close attention to the quality and behavior of the model’s responses.
- Evaluate each model across five defined evaluation dimensions.
- Identify and categorize relevant error clusters that may occur in the audio or conversation.
- Compare the performance of Model A and Model B based on the evaluation criteria and your observations.
- Select the model that demonstrated the stronger overall performance according to the defined evaluation framework.
- Write a detailed rationale explaining your final preference, referencing specific examples and observations from both conversations.
- Apply evaluation guidelines consistently across different scenarios and model interactions.
You’ll Thrive in This Role If You Have:
- A Bachelor’s Degree
- Strong attention to detail and ability to notice subtle differences in conversational quality.
- Excellent listening and comprehension skills.
- Strong written communication skills, particularly the ability to explain observations clearly and objectively.
- Ability to follow detailed evaluation guidelines consistently.
- Comfort speaking naturally and roleplaying different conversational scenarios.
- Ability to compare two interactions fairly without allowing personal preferences to influence the evaluation.
- Strong critical-thinking and analytical skills.
- Reliability and consistency when completing structured evaluation tasks.
- Familiarity with AI assistants, voice-based AI, or conversational systems.
How would you rate this job post?
See what other professionals think about this role.
Similar Opportunities
AI Adoption Manager
Neurons Lab
Poland
PortugalAccount Executive (AI Sales Stress-Testing) - QuickBooks Expert
Huzzle
Albania🌍Bosnia and HerzegovinaAI-for-Security Domain Lead (Offensive Security & LLM Expert)
Neurons Lab
Romania
SloveniaSenior AI Platform / DevTool Engineer
Progress Partners
Georgia
PolandMore Openings at Innodata
Technical Training & Quality Manager – AI Data
Innodata
United StatesResearch Data Scientist – GenAI/LLM
Innodata
United StatesTechnical Delivery Manager (TDM) for AI/ML Engagements
Innodata
United StatesBusiness Development Representative - Federal Practice
Innodata
United StatesExplore Top Companies in this Space
Redhorse Corporation
Government Technology (GovTech) & National Security / Defense AI & Data Engineering / Management Consulting / IT Modernization
SunnyData
IT Consulting / Data Engineering / AI / Enterprise Software
itD Tech
Digital Transformation Consulting & IT Architecture / Enterprise AI Strategy & Implementation / Advanced Data Engineering & Cloud-Native Development
Bounteous
AI Services & Digital Transformation Consulting / Product Engineering & Experience Design / Enterprise Data, Cloud & Marketing Technology Integration
Innodata
View Company ProfileInnodata is an elite, global data engineering and AI enablement powerhouse engineered to orchestrate massive-scale high-quality data operations, algorithmic training datasets, and digital transformation workflows for the world’s largest technology companies and enterprises. Operating as a critical "intelligence infrastructure layer" for the modern AI economy, the company eliminates the operational friction of deploying complex LLMs and generative AI applications—which frequently suffer from low-quality training data, biased outputs, and fragmented annotation pipelines—by seamlessly deploying a combination of advanced proprietary data annotation platforms, automated synthetic data generation, and an elite global network of subject matter experts. Moving beyond basic crowdsourced data validation paradigms, Innodata empowers Fortune 500 enterprises, global legal publishers, and leading medical institutions to dynamically scale their core AI foundation models, custom machine learning pipelines, and multi-modal semantic data processing with elite, scalable, and audit-ready precision. Under the hood, their sophisticated operational framework—bolstered by strict data security compliance, persistent programmatic data quality controls, and deep domain expertise across vertical domains like legal, healthcare, and finance—natively manages high-velocity data curation, complex enterprise knowledge graph construction, and high-stakes model evaluation. What sets Innodata apart is its uncompromising dedication to purpose-driven data craftsmanship; by bridging the gap between raw, unstructured enterprise information and high-performance, fine-tuned AI execution, the firm enables global commercial organizations to radically accelerate their AI time-to-market, eliminate systemic data engineering bottlenecks, and build an unassailable foundation for continuous commercial growth in the modern, AI-transformed global marketplace.
Safety First
- Never pay for a job application.
- Do not share sensitive bank info.
- Verify the client before starting work.