CRIS Current Research Information System

Prior research has explored the ability of computational models to predict a word semantic fit with a given predicate. While much work has been devoted to modeling the typicality relation between verbs and arguments in isolation, in this paper we take a broader perspective by assessing whether and to what extent computational approaches have access to the information about the typicality of entire events and situations described in language (Generalized Event Knowledge). Given the recent success of Transformers Language Models (TLMs), we decided to test them on a benchmark for the dynamic estimation of thematic fit. The evaluation of these models was performed in comparison with SDM, a framework specifically designed to integrate events in sentence meaning representations, and we conducted a detailed error analysis to investigate which factors affect their behavior. Our results show that TLMs can reach performances that are comparable to those achieved by SDM. However, additional analysis consistently suggests that TLMs do not capture important aspects of event knowledge, and their predictions often depend on surface linguistic features, such as frequent words, collocations and syntactic patterns, thereby showing sub-optimal generalization abilities.

Paolo Pedinotti, Giulia Rambelli, Emmanuele Chersoni, Enrico Santus, Alessandro Lenci, Philippe Blache (2021). Did the Cat Drink the Coffee? Challenging Transformers with Generalized Event Knowledge. Stroudsburg : Association for Computational Linguistics [10.18653/v1/2021.starsem-1.1].

Did the Cat Drink the Coffee? Challenging Transformers with Generalized Event Knowledge

Paolo Pedinotti;Giulia Rambelli;Emmanuele Chersoni;Enrico Santus;Alessandro Lenci;Philippe Blache

2021

Abstract

Prior research has explored the ability of computational models to predict a word semantic fit with a given predicate. While much work has been devoted to modeling the typicality relation between verbs and arguments in isolation, in this paper we take a broader perspective by assessing whether and to what extent computational approaches have access to the information about the typicality of entire events and situations described in language (Generalized Event Knowledge). Given the recent success of Transformers Language Models (TLMs), we decided to test them on a benchmark for the dynamic estimation of thematic fit. The evaluation of these models was performed in comparison with SDM, a framework specifically designed to integrate events in sentence meaning representations, and we conducted a detailed error analysis to investigate which factors affect their behavior. Our results show that TLMs can reach performances that are comparable to those achieved by SDM. However, additional analysis consistently suggests that TLMs do not capture important aspects of event knowledge, and their predictions often depend on surface linguistic features, such as frequent words, collocations and syntactic patterns, thereby showing sub-optimal generalization abilities.

Scheda breve

Scheda completa

Scheda completa (DC)

	Anno
	
				2021
			
	Titolo del volume
	
				Proceedings of *SEM 2021: The Tenth Joint Conference on Lexical and Computational Semantics
			
	Pagina iniziale
	
				1
			
	Pagina finale
	
				11
			
	Codice DOI
	
				https://dx.doi.org/10.18653/v1/2021.starsem-1.1
			
	Citazione
	
				Paolo Pedinotti,  Giulia Rambelli,  Emmanuele Chersoni,  Enrico Santus,  Alessandro Lenci,  Philippe Blache (2021). Did the Cat Drink the Coffee? Challenging Transformers with Generalized Event Knowledge. Stroudsburg : Association for Computational Linguistics [10.18653/v1/2021.starsem-1.1].
			
	Tutti gli autori
	
						Paolo Pedinotti; Giulia Rambelli; Emmanuele Chersoni; Enrico Santus; Alessandro Lenci; Philippe Blache
					
	Appare nelle tipologie:
	
				4.01 Contributo in Atti di convegno

File in questo prodotto:

File	Dimensione	Formato
2021.starsem-1.1.pdf accesso aperto Tipo: Versione (PDF) editoriale / Version Of Record Licenza: Creative commons Dimensione 453.45 kB Formato Adobe PDF Visualizza/Apri	453.45 kB	Adobe PDF	Visualizza/Apri

I documenti in IRIS sono protetti da copyright e tutti i diritti sono riservati, salvo diversa indicazione.

Utilizza questo identificativo per citare o creare un link a questo documento: https://hdl.handle.net/11585/938096

Citazioni

ND

ND

6

social impact