Video-Text as Game Players: Hierarchical Banzhaf Interaction for Cross-Modal Representation Learning
Peng Jin, Jie Chen(Peking University), Pengfei Xiong, Chang Liu(Nanjing University of Science and Technology), Jinfa Huang(Peking University), Shangxuan Tian, Yuan Li(Peking University), Xiangyang Ji(Tsinghua University)
Cited by 69
Related Papers
High-Performance FPGA-Based CNN Accelerator With Block-Floating-Point Arithmetic
|IEEE Transactions on Very Large Scale Integration (VLSI) Systems|2019|160
High-speed low-light in vivo two-photon voltage imaging of large neuronal populations
|Nature Methods|2023|112
In Situ Synthesis of Ultrathin ZIF-8 Film-Coated MSNs for Codelivering Bcl 2 siRNA and Doxorubicin to Enhance Chemotherapeutic Efficacy in Drug-Resistant Cancer Cells
|ACS Applied Materials & Interfaces|2018|111
Transcranial Focused Ultrasound Enhances Sensory Discrimination Capability through Somatosensory Cortical Excitation
|Ultrasound in Medicine & Biology|2021|68
UATVR: Uncertainty-Adaptive Text-Video Retrieval
|Unknown|2023|66