User-Regulation Deconfounded Conversational Recommender System with Bandit FeedbackYu Xia, Shuai Li, Ryan A. Rossi et al.|Unknown|2023Cited by 11
Which LLM to Play? Convergence-Aware Online Model Selection with Time-Increasing BanditsYu Xia, Shuai Li, Fang Kong et al.|Unknown|2024Cited by 6