Submitted by nielsr 36 How Well Does GPT-4o Understand Vision? Evaluating Multimodal Foundation Models on Standard Computer Vision Tasks · 6 authors 72 2
Submitted by akhaliq 22 Lost in Latent Space: An Empirical Study of Latent Diffusion Models for Physics Emulation · 6 authors 58 3
Submitted by RajveeSheth 14 Eka-Eval : A Comprehensive Evaluation Framework for Large Language Models in Indian Languages Lingo Research Group 30 2
Submitted by violetxi 5 LitBench: A Benchmark and Dataset for Reliable Evaluation of Creative Writing · 6 authors 2