MEGA-Bench: A Comprehensive AI Benchmark that Scales Multimodal Evaluation to Over 500 Real-World Tasks at a Manageable Inference Cost
A major challenge in the evaluation of vision-language models (VLMs) lies in understanding their diverse capabilities across a wide range...
