[COLM 24] MANGO: A Benchmark for Evaluating Mapping and Navigation Abilities of Large Language Models
By chatting or signing in you agree to the Terms and chat-message logging (revocable in History).