Limits for learning with language models. Asher, N., Bhar, S., Chaturvedi, A., Hunter, J., & Paul, S. In Palmer, A. & Camacho-collados, J., editors, Proceedings of the 12th Joint Conference on Lexical and Computational Semantics (*SEM 2023), pages 236–248, Toronto, Canada, July, 2023. Association for Computational Linguistics.
Limits for learning with language models [link]Paper  doi  abstract   bibtex   
With the advent of large language models (LLMs), the trend in NLP has been to train LLMs on vast amounts of data to solve diverse language understanding and generation tasks. The list of LLM successes is long and varied. Nevertheless, several recent papers provide empirical evidence that LLMs fail to capture important aspects of linguistic meaning. Focusing on universal quantification, we provide a theoretical foundation for these empirical findings by proving that LLMs cannot learn certain fundamental semantic properties including semantic entailment and consistency as they are defined in formal semantics. More generally, we show that LLMs are unable to learn concepts beyond the first level of the Borel Hierarchy, which imposes severe limits on the ability of LMs, both large and small, to capture many aspects of linguistic meaning. This means that LLMs will operate without formal guarantees on tasks that require entailments and deep linguistic understanding.
@inproceedings{asher_limits_2023,
	address = {Toronto, Canada},
	title = {Limits for learning with language models},
	url = {https://aclanthology.org/2023.starsem-1.22},
	doi = {10.18653/v1/2023.starsem-1.22},
	abstract = {With the advent of large language models (LLMs), the trend in NLP has been to train LLMs on vast amounts of data to solve diverse language understanding and generation tasks. The list of LLM successes is long and varied. Nevertheless, several recent papers provide empirical evidence that LLMs fail to capture important aspects of linguistic meaning. Focusing on universal quantification, we provide a theoretical foundation for these empirical findings by proving that LLMs cannot learn certain fundamental semantic properties including semantic entailment and consistency as they are defined in formal semantics. More generally, we show that LLMs are unable to learn concepts beyond the first level of the Borel Hierarchy, which imposes severe limits on the ability of LMs, both large and small, to capture many aspects of linguistic meaning. This means that LLMs will operate without formal guarantees on tasks that require entailments and deep linguistic understanding.},
	urldate = {2024-11-02},
	booktitle = {Proceedings of the 12th {Joint} {Conference} on {Lexical} and {Computational} {Semantics} (*{SEM} 2023)},
	publisher = {Association for Computational Linguistics},
	author = {Asher, Nicholas and Bhar, Swarnadeep and Chaturvedi, Akshay and Hunter, Julie and Paul, Soumya},
	editor = {Palmer, Alexis and Camacho-collados, Jose},
	month = jul,
	year = {2023},
	pages = {236--248},
}

Downloads: 0