Differing strategies in English and Japanese word segmentation: A computational-psycholinguistic approach to bootstrapping the lexicon

作者
Hagen Peukert
出处
期刊:Yearbook of the German Cognitive Linguistics Association [De Gruyter]
卷期号:1 (1)
标识
DOI:10.1515/gcla-2013-0006
摘要

Abstract How can six- to eight-month-olds find out where a word begins and where it ends in a continuous speech stream? A computer simulation reveals that the necessary information for segmenting word-like units is present though hidden in English Child-Directed-Speech. This holds even if the cognitive abilities of eight-month-olds constrain the range of possible segmentation algorithms. Applying transitional probability calculations to the incoming speech stream results in segmented chunks, most of which correspond to nonce formations. The key finding is, however, that the most frequent chunks from these formations are indeed words or phrases. Provided that infants prefer and recognize frequent items, it can be shown that a list of sound chains gradually augments a pseudo-lexicon of some eighty to one hundred entries. Now it can be further assumed that these entries are mapped to some new speech material. These mapping locate previously undetected word boundaries. In addition to that, from the pseudo-lexicon, other cues useful for segmentation - phonotactic constructions, prosody, or allophonic variants - could be unambiguously derived and used for complete segmentation before meanings are allocated to these chains. This segmentation mechanism does not seem to be universally true. A second line of computer simulations on Japanese reveals some indirect evidence against a universal segmentation mechanism based on transitional probabilities. To cope with the lack of representative samples of Japanese CDS, the segmentation algorithm was applied to Japanese adult speech (CSJ). For reasons of valid comparison, the simulation was also run on English adult speech (ICE-GB). While English adult speech still generates a viable lexicon, although at a significantly lower performance than English CDS, Japanese adult speech produces an error function that is, although above chance segmentation, insufficient for producing a lexicon

科研通智能强力驱动
Strongly Powered by AbleSci AI
科研通是完全免费的文献互助平台,具备全网最快的应助速度,最高的求助完成率。 对每一个文献求助,科研通都将尽心尽力,给求助人一个满意的交代。
实时播报
酷波er应助危机的百褶裙采纳,获得10
1秒前
GUYIMI完成签到,获得积分10
1秒前
1秒前
回家放羊发布了新的文献求助10
1秒前
3秒前
willow完成签到,获得积分10
3秒前
可爱的函函应助yan采纳,获得10
4秒前
在水一方应助从容的白风采纳,获得10
5秒前
yyyyyy完成签到,获得积分10
5秒前
123发布了新的文献求助10
6秒前
无花果应助Horizon采纳,获得10
7秒前
lixiao完成签到,获得积分10
8秒前
8秒前
一叶知秋完成签到,获得积分10
8秒前
zz完成签到,获得积分10
9秒前
9秒前
满意静丹完成签到,获得积分10
10秒前
霸气向秋完成签到,获得积分10
11秒前
吃的发布了新的文献求助30
11秒前
12秒前
12秒前
完美世界发布了新的文献求助10
12秒前
无花果应助MHR采纳,获得10
13秒前
善良紫南发布了新的文献求助10
14秒前
烟花应助任伟超采纳,获得10
14秒前
lixs应助hyw采纳,获得10
14秒前
不吃辣椒发布了新的文献求助10
14秒前
fafa完成签到,获得积分10
15秒前
黄陈涛完成签到 ,获得积分10
16秒前
17秒前
英姑应助追梦人采纳,获得10
17秒前
hl51完成签到,获得积分10
17秒前
yan发布了新的文献求助10
17秒前
优雅的皮卡丘完成签到,获得积分10
17秒前
今后应助poiu采纳,获得10
17秒前
毕业比耶完成签到,获得积分10
18秒前
19秒前
19秒前
20秒前
21秒前
高分求助中
(应助此贴封号)【重要!!请各用户(尤其是新用户)详细阅读】【科研通的精品贴汇总】 10000
An Introduction to Foreign Language Learning and Teaching 750
China Pluperfect I: Epistemology of Past and Outside in Chinese Art 520
Matrix Methods in Data Mining and Pattern Recognition Second Edition 510
What is the Future of Psychotherapy in Digital Age? Technology, AI Bots, and Psychotherapy after Covid 444
Synthesis of P-Chiral Phosphine Ligands and Their Applications in Asymmetric Catalysis 400
Management and the Arts 310
热门求助领域 (近24小时)
化学 材料科学 医学 生物 纳米技术 工程类 有机化学 化学工程 生物化学 计算机科学 内科学 物理 复合材料 催化作用 细胞生物学 无机化学 光电子学 物理化学 电极 基因
热门帖子
关注 科研通微信公众号,转发送积分 7629705
求助须知:如何正确求助?哪些是违规求助? 9204069
关于积分的说明 19736982
捐赠科研通 7199182
什么是DOI,文献DOI怎么找? 3274314
关于科研通互助平台的介绍 2436445
邀请新用户注册赠送积分活动 2270480