Lead story
Models & availability
Latest
Lead story
Models & availability
Latest
Researchers from UCLA introduce ConTextual, a dataset and leaderboard for evaluating multimodal models on context-sensitive text-rich visual reasoning tasks. Initial experiments show that both proprietary and open-source models struggle on this benchmark compared to humans.
From the source
That’s why we (researchers from University of California Los Angeles) created ConTextual, a Context-sensitive Text-rich visuaL reasoning dataset for evaluating LMMs. We also released a leaderboard, so that the community can see for themselves which models are the best at this task.
huggingface.co