this post was submitted on 18 Nov 2023
1 points (100.0% liked)
LocalLLaMA
3 readers
1 users here now
Community to discuss about Llama, the family of large language models created by Meta AI.
founded 1 year ago
MODERATORS
you are viewing a single comment's thread
view the rest of the comments
view the rest of the comments
As usual, "the beauty is in the eye of the beholder".
I think part of the point for these tests is to be able to solve these logical puzzles given all of the richness and ambiguity of NLs. We've had deterministic theorem solvers capable of solving these problems expressed as a closed set for decades.
That said, please see the capstone version of the prompt in the second update, which removes most of the ambiguity per the points you raised. It also removes the 'singles' aspect of tennis, which consistently trips up in-context reasoning, making the weaker LLMs think its a solo activity (despite an explicit following clarification).