Remote - United States
3 weeks ago

Job Overview

Pay
Not disclosed

Job description

Location:
Remote - United States
Work arrangement:
Remote

Role Summary

Innodata (Nasdaq: INOD) is a global data engineering company. We believe that data and Artificial Intelligence (AI) are inextricably linked. Our mission is to enable the responsible advancement of artificial intelligence by providing the data, evaluation frameworks, and human expertise required to build AI systems that can be trusted at scale. We provide a range of transferable solutions, platforms, and services for Generative AI / AI builders and adopters. In every relationship, we honor our 36+ year legacy delivering the highest quality data and outstanding outcomes for our customers.

We are seeking detail-oriented evaluators to conduct human quality evaluations for an enterprise AI customer support product on Instagram, WhatsApp, and Messenger. You will evaluate and benchmark AI model responses against complex evaluation rubrics using provided business knowledge bases.

Responsibilities

  • Model Evaluation: Review and score AI-generated customer interactions across Foundational, Experiential, and Operational dimensions (e.g., Action Fidelity, Faithfulness, Hallucination, Compliance, Tone, and Handoff).
  • Intent & Fact Verification: Benchmark both informational (R1) and transactional (R2) customer queries against authoritative business sources (FAQs, product catalogs, SOPs) within the task UI.
  • Quality Assurance: Participate in dual-review processes and daily calibration audits to ensure inter-rater agreement and establish ground-truth performance targets.
  • Performance Targets: Deliver precise evaluation

Requirements

You’ll Thrive in This Role If You Have

  • Customer Service Background: Prior experience in customer service, call centers, retail, or handling customer communications via email, chat, or phone (highly prioritized).
  • English Proficiency: Exceptional written English skills with a strong command of tone, brand voice, grammar, and nuance.
  • French Proficiency: Exceptional written French skills with a strong command of tone, brand voice, grammar, and nuance.
  • Analytical Precision: Ability to strictly follow multi-tier evaluation guidelines, complex logic trees, and technical rubrics without deviation.
  • Tech Adaptability: Comfort using dedicated web-based tools and labeling interfaces.

Please be aware of recruitment scams involving individuals or organizations falsely claiming to represent employers. Innodata will never ask for payment, banking details, or sensitive personal information during the application process. To learn more on how to recognize job scams, please visit the Federal Trade Commission’s guide at https://consumer.ftc.gov/articles/job-scams.

If you believe you’ve been targeted by a recruitment scam, please report it to Innodata at verifyjoboffer@innodata.com and consider reporting it to the FTC at ReportFraud.ftc.gov .

Role:
Content Evaluator – Bilingual (French and English)- Flexible Hours

Company profile

Innodata Inc.

innodata.com

Innodata is a data and AI company that helps organizations train and deploy generative and traditional AI. It provides supervised fine-tuning, RLHF, model safety evaluation and red teaming, data collection and creation, and a GenAI platform, along with data annotation for image, video, LiDAR sensor, text, document, audio and speech data. The company brings more than 36 years of data and domain expertise, says seven of the world's largest tech companies trust it, and operates across more than 20 delivery locations worldwide.

Headquarters
Hackensack, New Jersey, United States
Founded
1988
Funding Stage
Public

More jobs at Innodata Inc.

Similar ISL - Meta - MEPMS - (Project Delivery) jobs at other companies