中文文本可讀性分析:指標選取、模型建立與效度驗證

dc.contributor國立臺灣師範大學教育心理與輔導學系zh_tw
dc.contributor.author宋曜廷zh_tw
dc.contributor.author陳茹玲zh_tw
dc.contributor.author李宜憲zh_tw
dc.contributor.author查日龢zh_tw
dc.contributor.author曾厚強zh_tw
dc.contributor.author林維駿zh_tw
dc.contributor.author張道行zh_tw
dc.contributor.author張國恩zh_tw
dc.date.accessioned2014-12-02T06:38:49Z
dc.date.available2014-12-02T06:38:49Z
dc.date.issued2013-03-01zh_TW
dc.description.abstract本研究根據中文特性發展可讀性指標,接著建立中文文本可讀性數學模型,並進行模型效度驗證。本研究以所發展24個可讀性指標為預測變項,386篇教科書文章之年級值為效標變項,建立逐步迴歸(stepwise regression)與SVM可讀性數學模型,再以96篇新文章為測試資料進行模型驗證。研究結果顯示:在逐步迴歸模型中,難詞數、單句數比率、實詞頻對數平均與人稱代名詞數為重要的預測變項;以SVM模型F-score方法所得的重要預測變項則為難詞數、二字詞數、字數與中筆畫字元數等。逐步迴歸模型與SVM模型對新文章的預測正確性分別為55.21%及72.92%,兩種模型預測低年級文章之正確性均高於高年級文章。zh_tw
dc.description.abstractThis study aims to (a) develop readability indicators based on the textual factors that influence reading comprehension; (b) construct the readability model for Chinese text; and (c) validate the proposed readability models. This study constructs readability models employing step regression and SVM, using 24 readability indicators as its predictive variable and the grade level of 386 textbook articles as the criteria. The proposed models are then validated according to an additional 96 texts. The results show that in step regression, the critical predictors are the number of complex words, proportion of simple sentences, average logarithm of content word frequency, and number of personal pronouns. In the SVM model, the critical predictors selected by using the F-score include the number of complex words, number of two-character words, number of characters, and number of intermediate-stroke characters. The accuracy rates of step regression and SVM are 55.21% and 72.92%, respectively. Both models predict the texts more accurately at the lower grade levels than at the higher grade levels.en_US
dc.description.urihttp://www.airitilibrary.com/Publication/alDetailedMesh?DocID=10139656-201303-201303280010-201303280010-75-106zh_TW
dc.identifierntnulib_tp_A0201_01_068zh_TW
dc.identifier.issn1013-9656zh_TW
dc.identifier.urihttp://rportal.lib.ntnu.edu.tw/handle/20.500.12235/40735
dc.languagezh_TWzh_TW
dc.publisher臺灣心理學會zh_tw
dc.relation中華心理學刊,55(1),75-106。zh_tw
dc.relation.urihttp://dx.doi.org/10.6129/CJP.20120621zh_TW
dc.subject.other可讀性zh_tw
dc.subject.other正確性zh_tw
dc.subject.other逐步迴歸zh_tw
dc.subject.otherSVM數學模型zh_tw
dc.subject.otheraccuracyen_US
dc.subject.otherreadabilityen_US
dc.subject.otherstepwise regressionen_US
dc.subject.othersupport vector machineen_US
dc.title中文文本可讀性分析:指標選取、模型建立與效度驗證zh_tw

Files

Collections