Generics are statements that express generalizations and are used to communicate generalizable knowledge. While generics convey general truths (e.g., Birds can fly), they often allow for exceptions (e.g., penguins do not fly). Nonetheless, generics form the basis of how we communicate our commonsense about the world. We explored the interpretation of generics in Masked Language Models (MLMs), building on psycholinguistic experimental designs. As this interpretation requires a comparison with overtly quantified sentences, we investigated i) the probability of quantifiers, ii) the internal representation of nouns in generic vs. quantified sentences, and iii) whether the presence of a generic sentence as context influences quantifiers’ probabilities. The outcomes confirm that MLMs are insensitive to quantification; nevertheless, they appear to encode a meaning associated with the generic form, which leads them to reshape the probability associated with various quantifiers when the generic sentence is provided as context.

Collacciani, C., Rambelli, G., Interpretation of Generalization in Masked Language Models: An Investigation Straddling Quantifiers and Generics, Comunicazione, in Proceedings of the 9th Italian Conference on Computational Linguistics - CLiC-it 2023, (Venice, 30-November 02-December 2023), CEUR-WS, Aachen 2023:3596 143-153 [https://hdl.handle.net/10807/339577]

Interpretation of Generalization in Masked Language Models: An Investigation Straddling Quantifiers and Generics

Rambelli, Giulia
Supervision
2023

Abstract

Generics are statements that express generalizations and are used to communicate generalizable knowledge. While generics convey general truths (e.g., Birds can fly), they often allow for exceptions (e.g., penguins do not fly). Nonetheless, generics form the basis of how we communicate our commonsense about the world. We explored the interpretation of generics in Masked Language Models (MLMs), building on psycholinguistic experimental designs. As this interpretation requires a comparison with overtly quantified sentences, we investigated i) the probability of quantifiers, ii) the internal representation of nouns in generic vs. quantified sentences, and iii) whether the presence of a generic sentence as context influences quantifiers’ probabilities. The outcomes confirm that MLMs are insensitive to quantification; nevertheless, they appear to encode a meaning associated with the generic form, which leads them to reshape the probability associated with various quantifiers when the generic sentence is provided as context.
2023
Inglese
Proceedings of the 9th Italian Conference on Computational Linguistics - CLiC-it 2023
CLiC-it 2023 - 9th Italian Conference on Computational Linguistics
Venice
Comunicazione
30-nov-2023
2-dic-2023
1613-0073
CEUR-WS
Collacciani, C., Rambelli, G., Interpretation of Generalization in Masked Language Models: An Investigation Straddling Quantifiers and Generics, Comunicazione, in Proceedings of the 9th Italian Conference on Computational Linguistics - CLiC-it 2023, (Venice, 30-November 02-December 2023), CEUR-WS, Aachen 2023:3596 143-153 [https://hdl.handle.net/10807/339577]
File in questo prodotto:
File Dimensione Formato  
paper17.pdf

accesso aperto

Tipologia file ?: Versione Editoriale (PDF)
Licenza: Creative commons
Dimensione 1.93 MB
Formato Adobe PDF
1.93 MB Adobe PDF Visualizza/Apri

I documenti in IRIS sono protetti da copyright e tutti i diritti sono riservati, salvo diversa indicazione.

Utilizza questo identificativo per citare o creare un link a questo documento: https://hdl.handle.net/10807/339577
Citazioni
  • ???jsp.display-item.citation.pmc??? ND
  • Scopus 0
  • ???jsp.display-item.citation.isi??? ND
social impact