The following investigation presents a constructional procedure segmenting an utterance in a way which correlates well with word and morpheme boundaries. The procedure requires a large set of utterances, elicited in a certain manner from an informant (or found in a very large corpus); and it requires that all the utterances be written in the same phonemic representation, determined without reference to morphemes. It then investigates a particular distributional relation among the phonemes in the utterances thus collected; and on the basis of this relation among the phonemes, it indicates particular points of segmentation within one utterance at a time. For example, in the utterance /hiyzkwikər/ He's quicker it will indicate segmentation at the points marked by dots: /hiy.z.kwik.ər/; and it will do so purely by comparing this phonemic sequence with the phonemic sequences of other utterances.
No takes yet. Share an insight, caveat, or question.
Zellig S. Harris (1955) studied this question.