白丝美女被狂躁免费视频网站,500av导航大全精品,yw.193.cnc爆乳尤物未满,97se亚洲综合色区,аⅴ天堂中文在线网官网

Systems and methods for generating recitation items

專利號
US10867525B1
公開日期
2020-12-15
申請人
Educational Testing Service(US NJ Princeton)
發(fā)明人
Su-Youn Yoon; Lei Chen; Keelan Evanini; Klaus Zechner
IPC分類
G09B19/04; G10L13/08; G06F40/211
技術(shù)領(lǐng)域
text,phoneme,prosodic,metric,recitation,phonetic,texts,syntactic,native,language
地域: NJ NJ Princeton

摘要

Computer-implemented systems and methods are provided for automatically generating recitation items. For example, a computer performing the recitation item generation can receive one or more text sets that each includes one or more texts. The computer can determine a value for each text set using one or more metrics, such as a vocabulary difficulty metric, a syntactic complexity metric, a phoneme distribution metric, a phonetic difficulty metric, and a prosody distribution metric. Then the computer can select a final text set based on the value associated with each text set. The selected final text set can be used as the recitation items for a speaking assessment test.

說明書

1 2 3 4 5 6 7 8 9 10 11 12 13 14 15 16

At 220, a vocabulary difficulty metric analyzes the text to determine a level of difficulty of its words. “Difficult words,” for example, may be foreign words or uncommon names that are inappropriate for a particular recitation task. Since difficult words tend to appear less frequently than easy words, the difficulty of a word can be estimated based on the frequency of the word appearing in a reference corpus (i.e., a word that rarely appears in the reference corpus may be deemed difficult). Based on this assumption, a vocabulary difficulty value can be estimated based on, for example, the proportion of the text's low frequency words and average word frequency. At 225, the vocabulary difficulty value is compared to a pre-determined vocabulary difficulty range, which in one embodiment is determined based on a set of training texts that have been deemed suitable for the recitation task. Then at 250, the text filter module determines whether to filter out the text from the text set based on the result of the comparison and, optionally, the results of other metrics.

At 230, a syntactic complexity metric is employed to identify texts with overly complicated syntactic structure, which may not be appropriate for a recitation task. A syntactic complexity value may be calculated using any conventional means for estimating syntax complexity (such as those used in automated essay scoring systems). At 235, the text filter module compares the syntactic complexity value to a pre-determined syntactic complexity range, which in one embodiment is determined based on a set of training texts that have been deemed suitable for the recitation task. Then at 250, the text filter module determines whether to filter out the text from the text set based on the result of the comparison and, optionally, the results of other metrics.

權(quán)利要求

1
微信群二維碼
意見反饋