A Note on Local Ultrametricity in Text

Murtagh, Fionn

(2007)

Murtagh, Fionn (2007) A Note on Local Ultrametricity in Text.

Our Full Text Deposits

Full text access: Open

Full text file - 160.68 KB

Abstract

High dimensional, sparsely populated data spaces have been characterized in terms of ultrametric topology. This implies that there are natural, no tnecessarily unique, tree or hierarchy structures defined by the ultrametrictopology. In this note we study the extent of local ultrametric topology in texts, with the aim of finding unique ``fingerprints'' for a text or corpus, discriminating between texts from different domains, and opening up the possibility of exploiting hierarchical structures in the data. We use coherent and meaningful collections of over 1000 texts, comprising over 1.3 millionwords.

Information about this Version

This is a Submitted version
This version's date is: 27/1/2007
This item is not peer reviewed

Link to this Version

https://repository.royalholloway.ac.uk/items/fe65be48-3ddd-f4a8-c900-7f3c5ed3f33b/2/

Item TypeMonograph (Working Paper)
TitleA Note on Local Ultrametricity in Text
AuthorsMurtagh, Fionn
Uncontrolled Keywordscs.CL, I.5.3; I.7.2; H.3
DepartmentsFaculty of Science\Computer Science

Identifiers

Deposited by Research Information System (atira) on 24-Jul-2012 in Royal Holloway Research Online.Last modified on 24-Jul-2012

Notes

18 pp


Details