Model Evaluation

Explore Model Evaluation through related topics and the articles other pages reference most.

Explore articles

Reset filters
Browse subtopics: Meta AI

Articles that also belong to these categories. Counts cover all of Model Evaluation.

Showing 1-2 of 2 articles

HalluLens

HalluLens is a large language model hallucination benchmark introduced by researchers at Meta AI's Fundamental AI Research (FAIR) lab, together with collaborators at the Hong Kong University of Science and…

AI BenchmarksMeta AI