Walter Jb van Heuven, Joshua S Payne, Manon W Jones
{"title":"SUBTLEX-CY: A new word frequency database for Welsh.","authors":"Walter Jb van Heuven, Joshua S Payne, Manon W Jones","doi":"10.1177/17470218231190315","DOIUrl":null,"url":null,"abstract":"<p><p>We present SUBTLEX-CY, a new word frequency database created from a 32-million-word corpus of Welsh television subtitles. An experiment comprising a lexical decision task examined SUBTLEX-CY frequency estimates against words with inconsistent frequencies in a much smaller Welsh corpus that is often used by researchers, the <i>Cronfa Electroneg o'r Gymraeg</i> (CEG), and three other Welsh word frequency databases. Words were selected that were classified as low frequency (LF) in SUBTLEX-CY and high frequency (HF) in CEG and compared with words that were classified as medium frequency (MF) in both SUBTLEX-CY and CEG. Reaction time analyses showed that HF words in CEG were responded to more slowly compared to MF words, suggesting that SUBTLEX-CY corpus provides a more reliable estimate of Welsh word frequencies. The new Welsh word frequency database that also includes part-of-speech, contextual diversity, and other lexical information is freely available for research purposes on the Open Science Framework repository at https://osf.io/9gkqm/.</p>","PeriodicalId":20869,"journal":{"name":"Quarterly Journal of Experimental Psychology","volume":" ","pages":"1052-1067"},"PeriodicalIF":1.5000,"publicationDate":"2024-05-01","publicationTypes":"Journal Article","fieldsOfStudy":null,"isOpenAccess":false,"openAccessPdf":"https://www.ncbi.nlm.nih.gov/pmc/articles/PMC11032624/pdf/","citationCount":"0","resultStr":null,"platform":"Semanticscholar","paperid":null,"PeriodicalName":"Quarterly Journal of Experimental Psychology","FirstCategoryId":"102","ListUrlMain":"https://doi.org/10.1177/17470218231190315","RegionNum":3,"RegionCategory":"心理学","ArticlePicture":[],"TitleCN":null,"AbstractTextCN":null,"PMCID":null,"EPubDate":"2023/8/30 0:00:00","PubModel":"Epub","JCR":"Q4","JCRName":"PHYSIOLOGY","Score":null,"Total":0}
引用次数: 0
Abstract
We present SUBTLEX-CY, a new word frequency database created from a 32-million-word corpus of Welsh television subtitles. An experiment comprising a lexical decision task examined SUBTLEX-CY frequency estimates against words with inconsistent frequencies in a much smaller Welsh corpus that is often used by researchers, the Cronfa Electroneg o'r Gymraeg (CEG), and three other Welsh word frequency databases. Words were selected that were classified as low frequency (LF) in SUBTLEX-CY and high frequency (HF) in CEG and compared with words that were classified as medium frequency (MF) in both SUBTLEX-CY and CEG. Reaction time analyses showed that HF words in CEG were responded to more slowly compared to MF words, suggesting that SUBTLEX-CY corpus provides a more reliable estimate of Welsh word frequencies. The new Welsh word frequency database that also includes part-of-speech, contextual diversity, and other lexical information is freely available for research purposes on the Open Science Framework repository at https://osf.io/9gkqm/.
期刊介绍:
Promoting the interests of scientific psychology and its researchers, QJEP, the journal of the Experimental Psychology Society, is a leading journal with a long-standing tradition of publishing cutting-edge research. Several articles have become classic papers in the fields of attention, perception, learning, memory, language, and reasoning. The journal publishes original articles on any topic within the field of experimental psychology (including comparative research). These include substantial experimental reports, review papers, rapid communications (reporting novel techniques or ground breaking results), comments (on articles previously published in QJEP or on issues of general interest to experimental psychologists), and book reviews. Experimental results are welcomed from all relevant techniques, including behavioural testing, brain imaging and computational modelling.
QJEP offers a competitive publication time-scale. Accepted Rapid Communications have priority in the publication cycle and usually appear in print within three months. We aim to publish all accepted (but uncorrected) articles online within seven days. Our Latest Articles page offers immediate publication of articles upon reaching their final form.
The journal offers an open access option called Open Select, enabling authors to meet funder requirements to make their article free to read online for all in perpetuity. Authors also benefit from a broad and diverse subscription base that delivers the journal contents to a world-wide readership. Together these features ensure that the journal offers authors the opportunity to raise the visibility of their work to a global audience.