Content Evaluator Customer Support & LLM Benchmarking

🏢 innodata inc.
📍 QatarFull-timeRemote
📅 Posted: 1w ago🔄 Updated: 1w ago
CV%
✨ AI Summary
Innodata is seeking detail-oriented evaluators to conduct human quality evaluations for an enterprise AI customer support product on different social media platforms. You will evaluate and benchmark AI model responses against complex evaluation rubrics using provided business knowledge bases. This role requires native proficiency in Arabic and excellent English proficiency, with strong grammar, tone, and comprehension. The ideal candidate will have experience in customer service, call centers, retail, or customer communications, strong analytical and attention-to-detail skills, and be comfortable working with web-based evaluation and labeling tools.
Required Skills
Business, Sales & Management
Customer Success
Finance, Legal & Governance
Valuation
Soft Skills & Professional Competencies
Attention to DetailCommunication
Requirements
Required Skills: Experience in customer service, call centers, retail, or customer communications. Excellent English proficiency with strong grammar, tone, and comprehension. Strong analytical and attention-to-detail skills. Ability to follow complex guidelines consistently. Comfortable working with web-based evaluation and labeling tools.
Description
Innodata (NASDAQ: INOD) is a leading data engineering company. With more than 2,000 customers and operations in 13 cities around the world, we are an AI technology solutions provider-of-choice for 4 out of 5 of the world's biggest technology companies, as well as leading companies across financial services, insurance, technology, law, and medicine.By combining advanced machine learning and artificial intelligence (ML/AI) technologies, a global workforce of subject matter experts, and a high-security infrastructure, we're helping usher in the promise of AI. Innodata offers a powerful combination of both digital data solutions and easy-to-use, high-quality platforms.Our global workforce includes over 5,000 employees in the United States, Canada, United Kingdom, the Philippines, India, Sri Lanka, Israel and Germany.About the role:We are seeking detail-oriented evaluators to conduct human quality evaluations for an enterprise AI customer support product on different social media platform. You will evaluate and benchmark AI model responses against complex evaluation rubrics using provided business knowledge bases.Job role: Content Evaluator – Customer Support & LLM BenchmarkingLanguage : Arabic(Native proficiency)Hourly commitment: 6 hours per day Key ResponsibilitiesEvaluate AI-generated customer interactions using defined quality parameters.Verify AI responses against FAQs, product catalogs, SOPs, and other approved sources.Participate in dual reviews and daily calibration to maintain evaluation consistency.Follow detailed rubrics and logic trees while meeting productivity targets.Maintain an average handling time of approximately 45 minutes per job.Required SkillsExperience in customer service, call centers, retail, or customer communications.Excellent English proficiency with strong grammar, tone, and comprehension.Strong analytical and attention-to-detail skills.Ability to follow complex guidelines consistently.Comfortable working with web-based evaluation and labeling tools.Engagement DetailsDuration: 2 months (60 working days), extendableSchedule: Availability for 6 hours per dayWork Type: FreelanceIf you're ready to contribute to the future of AI while working remotely on an exciting music-focused project, we'd love to hear from you!
✨ Premium Match Details
Deep-dive CV analysis, customized Cover Letters, and Interview prep!
📊 Match Analysis
Insights against your active CV
📊
Personalized Match Analysis
Upload your CV to see exact matching percentages, detailed skills mapping, and gap analysis for this role.
🎯 Overalli74%
⚡ Skillsi85%
View Breakdown
Ontology Match: 85.0
Matched:✓ Requirements Matching✓ Ontology Skills Mapping
📜 Eligibilityi49%
View Breakdown
Local: 19600%
🏗️ Career Fiti91%
View Breakdown
Seniority: 91.0
📋 Requirementsi67%
View Breakdown
Domain: 67.0
🔥 Motivationi78%
View Breakdown
Title Fit: 78.00