A large language model (LLM) agent can follow more graph paths without acquiring more independent evidence. GraphEcho tests whether agents mistake these repeated encounters for additional corroboration. The benchmark varies path counts and evidential origins while holding…
Read the original source — arxiv.org
paper · Shared by tscosj
0 comments
No comments yet.