How Do BERT Embeddings Organize Linguistic Knowledge

Giovanni Puccetti,Alessio Miaschi,Felice DellOrletta

How Do BERT Embeddings Organize Linguistic Knowledge

2021

Giovanni Puccetti
Alessio Miaschi
Felice DellOrletta

Several studies investigated the linguistic information implicitly encoded in Neural Language Models. Most of these works focused on quantifying the amount and type of information available within their internal representations and across their layers. In line with this scenario, we proposed a different study, based on Lasso regression, aimed at understanding how the information encoded by BERT sentence-level representations is arrange within its hidden units. Using a suite of several probing tasks, we showed the existence of a relationship between the implicit knowledge learned by the model and the number of individual units involved in the encodings of this competence. Moreover, we found that it is possible to identify groups of hidden units more relevant for specific linguistic properties.

Keywords:

Rule-based machine translation
Type (model theory)
lasso regression
Competence (human resources)
Linguistics
Computer science
Line (text file)
Language model
implicit knowledge
Suite

Correction
Source
Cite
Save
Machine Reading By IdeaReader

References

Citations