arXiv cs.AI / cs.LG / cs.CL·10d agoMUSE: Benchmarking Large Vision-Language Models on Multi-Modal Understanding in Situated Education#benchmark#cultural-understanding#educationAI research1