Res-Bench: Benchmarking the Robustness of Multimodal Large Language Models to Dynamic Resolution InputChenxu Li, Wang Xiang, Z. H. Wang et al.|arXiv (Cornell University)|2025Cited by 0
SeViCES: Unifying Semantic-Visual Evidence Consensus for Long Video UnderstandingSheng Yuan, Xiangnan He, Chenxu Li et al.|arXiv (Cornell University)|2025Cited by 0