Permutation invariant matrix statistics and computational language tasks


Kartsaklis、Rangoolam、Sadrzadeh によって導入された言語行列理論プログラムは、重要な統計をエンコードする重要な観測値とみなされる順列不変多項式関数に基づいて、型駆動型の分布セマンティクスで生成される行列の統計へのアプローチです。


The Linguistic Matrix Theory programme introduced by Kartsaklis, Ramgoolam and Sadrzadeh is an approach to the statistics of matrices that are generated in type-driven distributional semantics, based on permutation invariant polynomial functions which are regarded as the key observables encoding the significant statistics. In this paper we generalize the previous results on the approximate Gaussianity of matrix distributions arising from compositional distributional semantics. We also introduce a geometry of observable vectors for words, defined by exploiting the graph-theoretic basis for the permutation invariants and the statistical characteristics of the ensemble of matrices associated with the words. We describe successful applications of this unified framework to a number of tasks in computational linguistics, associated with the distinctions between synonyms, antonyms, hypernyms and hyponyms.


著者 Manuel Accettulli Huber,Adriana Correia,Sanjaye Ramgoolam,Mehrnoosh Sadrzadeh
発行日 2023-09-26 17:29:38+00:00
arxivサイト arxiv_id(pdf)

提供元, 利用サービス, Google

カテゴリー: cond-mat.stat-mech, cs.CL, hep-th パーマリンク