Skip to main navigation Skip to search Skip to main content

Some Further Dialectometrical Steps

John Nerbonne, Jelena Prokić, Martijn Wieling, Charlotte Gooskens

Research output: Chapter in Book/Report/Conference proceedingChapterResearchpeer-review

Abstract

This article surveys recent developments furthering dialectometric research which the authors have been involved in, in particular techniques for measuring large numbers of pronunciations (in phonetic transcription) of comparable words at various sites. Edit distance (also known as Levenshtein distance) has been deployed for this purpose, for which refinements and analytic techniques continue to be developed. The focus here is on (i) an empirical approach, using an information-theoretical measure of mutual information, for deriving the appropriate segment distances to serve within measures of sequence distance; (ii) a heuristic technique for simultaneously aligning large sets of comparable pronunciations, a necessary step in applying phylogenetic analysis to sound segment data; (iii) spectral clustering, a technique borrowed from bio-informatics, for identifying the (linguistic) features responsible for (dialect) divisions among sites; (iv) techniques for studying the (mutual) comprehensibility of closely related varieties; and (v) Séguy’s law, or the generality of sub-linear diffusion of aggregate linguistic variation.
Original languageEnglish
Title of host publicationTools for Linguistic Variation
EditorsGotzon Aurrekoetxea, Jose Luis Ormaetxea
Place of PublicationBilbao, Spain
PublisherUniversidad del País Vasco
Pages41-56
Edition1
ISBN (Print)9788498604290
Publication statusPublished - 31 Dec 2010

Fingerprint

Dive into the research topics of 'Some Further Dialectometrical Steps'. Together they form a unique fingerprint.

Cite this