A high-frequency sense list.

Lei Liu,Tongxi Gong,Jianjun Shi,Yi Guo

doi:10.3389/fpsyg.2024.1430060

Abstract

A number of high-frequency word lists have been created to help foreign language learners master English vocabulary. These word lists, despite their widespread use, did not take word meaning into consideration. Foreign language learners are unclear on which meanings they should focus on first. To address this issue, we semantically annotated the Corpus of Contemporary American English (COCA) and the British National Corpus (BNC) with high accuracy using a BERT model. From these annotated corpora, we calculated the semantic frequency of different senses and filtered out 5000 senses to create a High-frequency Sense List. Subsequently, we checked the validity of this list and compared it with established influential word lists. This list exhibits three notable characteristics. First, it achieves stable coverage in different corpora. Second, it identifies high-frequency items with greater accuracy. It achieves comparable coverage with lists like GSL, NGSL, and New-GSL but with significantly fewer items. Especially, it includes everyday words that used to fall off high-frequency lists without requiring manual adjustments. Third, it describes clearly which senses are most frequently used and therefore should be focused on by beginning learners. This study represents a pioneering effort in semantic annotation of large corpora and the creation of a word list based on semantic frequency.

Talk to us

Join us for a 30 min session where you can share your feedback and ask us any queries you have

Schedule a call

R Discovery Prime

R Discovery Prime

A high-frequency sense list.

Abstract

Talk to us

Similar Papers

More From: Frontiers in psychology

Lead the way for us

Similar Papers

I. S. Paul Nation: Making and Using Word Lists for Language Learning and Testing
Chen Ding ... Barry Lee Reynolds
Applied Linguistics | VOL. 39
Chen Ding, et. al.Chen Ding ... Barry Lee Reynolds
13 Dec 2017
Applied Linguistics | VOL. 39

Mining Word Meanings
Harvey S Wiener
-
Harvey S WienerHarvey S Wiener
29 Aug 1996
29 Aug 1996

Corpus-Based Frequency Profiling: Migration To A Word List Based On The British National Corpus
Leah Gilner ... Frank Morales
The Buckingham Journal of Language and Linguistics | VOL. 1
Leah Gilner, et. al.Leah Gilner ... Frank Morales
22 Jun 2010
The Buckingham Journal of Language and Linguistics | VOL. 1

Author response: Neural tracking of phrases in spoken language comprehension is automatic and task-dependent
Sanne ten Oever ... Greta Kaufeld
-
Sanne ten Oever, et. al.Sanne ten Oever ... Greta Kaufeld
22 Jun 2022
22 Jun 2022

Editage

Paperpal

R Discovery

Mind the Graph

R Discovery Prime

R Discovery Prime

A high-frequency sense list.

Abstract

Talk to us

Similar Papers

More From: Frontiers in psychology