LLM Model Response Evaluation
UpworkUSA3d ago
LLMremotecurated-jd
Job description
LLM Model Response Evaluation at Upwork.
About the role
We are seeking subject matter experts to assess AI-generated outputs based on specific quality standards. You will analyze model performance across various formats to ensure accuracy, visual quality, and logical consistency.
Key facts
What you'll do
- Perform side-by-side comparisons of AI responses to determine quality.
- Review diverse content types including text, images, audio, video, PDFs, and HTML widgets.
- Evaluate infographics and UI elements for factual precision.
- Conduct independent research on complex topics ranging from science and engineering to arts and history.
- Apply project-specific guidelines provided within the evaluation interface to every task.
Requirements
- Minimum of 3 years of professional experience in GenAI or LLM data evaluation.
- Master degree or PhD (PhD candidates are preferred).
- Proficiency in researching unknown subjects using reliable sources to form evidence-based conclusions.
- Capability to critique content across multiple media modalities.
Practical notes
- This role offers a flexible, remote schedule where you choose which tasks to accept.
- There are no guaranteed weekly hours, as the volume of available work fluctuates.