Okapi: Instruction-tuned Large Language Models in Multiple Languages with Reinforcement Learning from Human Feedback

Viet Dac Lai(Adobe Systems (United States)), Thien Huu Nguyen(University of Oregon), Thuat Nguyen-Tran, Chien Van Nguyen, Nghia Trung Ngo, Franck Dernoncourt(Adobe Systems (United States)), Ryan A. Rossi(Adobe Systems (United States))
arXiv (Cornell University)
July 29, 2023
Cited by 3


Related Papers

The Network Data Repository with Interactive Graph Analytics and Visualization
|Proceedings of the AAAI Conference on Artificial Intelligence|2015|2.5k
Bias and Fairness in Large Language Models: A Survey
|Computational Linguistics|2024|583
Attention Models in Graphs
|ACM Transactions on Knowledge Discovery from Data|2019|242