Not-Just-Scaling Laws: Towards a Better Understanding of the Downstream Impact of Language Model Design Decisions
Emmy Liu(Carnegie Mellon University), Graham Neubig(Carnegie Mellon University), Aditi Raghunathan, Kiril Gashteovski, Patrick Fernandes(Instituto de Telecomunicações), Carolin Lawrence, Lindia Tjuatja(Carnegie Mellon University), Michael Chen, Lintang Sutawika, Lara Marinov, Amanda Bertsch(Carnegie Mellon University)
Cited by 0
Related Papers
Pre-train, Prompt, and Predict: A Systematic Survey of Prompting Methods in Natural Language Processing
|ACM Computing Surveys|2022|3.5k
Language Models of Code are Few-Shot Commonsense Learners
|Unknown|2022|106
PAL: Program-aided Language Models
|arXiv (Cornell University)|2022|104
Do LLMs Exhibit Human-like Response Biases? A Case Study in Survey Design
|Transactions of the Association for Computational Linguistics|2024|61