LLM Model Response Evaluation

Maryland City, MDContractPosted Jul 31, 2026

LLM Model Response Evaluation

  • Contract

Company Description

Our client, a global technology company that helps businesses build, train, and manage AI systems is looking for experts to evaluate model-generated content against defined quality rubrics such as factuality, consistency, aesthetics, and other evaluation criteria.

Job Description

  • Evaluating UI widgets, infographics, image factuality, side-by-side comparisons, and similar AI evaluation activities.
  • The work may involve text, images, audio, video, HTML widgets, PDFs, or combinations of these modalities.
  • The work is domain-agnostic and may cover topics across arts, culture, history, science, engineering, and more.
  • Resources will be expected to independently research unfamiliar topics using trusted sources before making evaluation decisions.
  • Each task will include detailed project guidelines within the evaluation platform.

Qualifications

  • 3+ years of hands-on experience in LLM / GenAI data evaluation.
  • Master's or PhD required (PhD candidates strongly preferred).
  • Ability to research unfamiliar topics using trusted sources and make well-supported judgments.
  • Comfortable evaluating content across multiple modalities 

Additional Information

Flexible and remote work
Variable workload: Accept or decline tasks based on your availability
No guaranteed hours: Workload may vary weekly

By clicking the link above or any third-party link within this posting, you are leaving this site and going to a third-party website where the third-party website's terms and privacy policy apply

I'm interestedI'm interestedPrivacy Notice

Want jobs like this matched to you?

SimpleCareer scores fresh postings against your résumé so you only see the matches that matter.

Get started free