LocalLLaMA

3 readers

1 users here now

Community to discuss about Llama, the family of large language models created by Meta AI.

founded 1 year ago

MODERATORS

communick@poweruser.forum

RAG - Vectara's Hallucination leaderboard (alien.top)

submitted 1 year ago by AdamDhahabi@alien.top to c/localllama@poweruser.forum

5 comments fedilink hide all child comments

Vectara's Hallucination Evaluation Model and leaderboard was launched last week. I notice Mistral having a hallucination rate of 9.4% compared to 5.6% for Llama2. Any thoughts?

https://preview.redd.it/sj0akn15tszb1.png?width=1118&format=png&auto=webp&s=ca9ec766f592a8748bf95a8ad2ef81483c2270bd

Source: https://github.com/vectara/hallucination-leaderboard

you are viewing a single comment's thread
view the rest of the comments

[–] Distinct-Target7503@alien.top 1 points 1 year ago

How is possible that Llama2 13B and 7B have lower hallucination rate than Claude?

permalink
fedilink
source