• Skip to primary navigation
  • Skip to main content

site logo
The Electronic Journal for English as a Second Language
search
  • Home
  • About TESL-EJ
  • Vols. 1-15 (1994-2012)
    • Volume 1
      • Volume 1, Number 1
      • Volume 1, Number 2
      • Volume 1, Number 3
      • Volume 1, Number 4
    • Volume 2
      • Volume 2, Number 1 — March 1996
      • Volume 2, Number 2 — September 1996
      • Volume 2, Number 3 — January 1997
      • Volume 2, Number 4 — June 1997
    • Volume 3
      • Volume 3, Number 1 — November 1997
      • Volume 3, Number 2 — March 1998
      • Volume 3, Number 3 — September 1998
      • Volume 3, Number 4 — January 1999
    • Volume 4
      • Volume 4, Number 1 — July 1999
      • Volume 4, Number 2 — November 1999
      • Volume 4, Number 3 — May 2000
      • Volume 4, Number 4 — December 2000
    • Volume 5
      • Volume 5, Number 1 — April 2001
      • Volume 5, Number 2 — September 2001
      • Volume 5, Number 3 — December 2001
      • Volume 5, Number 4 — March 2002
    • Volume 6
      • Volume 6, Number 1 — June 2002
      • Volume 6, Number 2 — September 2002
      • Volume 6, Number 3 — December 2002
      • Volume 6, Number 4 — March 2003
    • Volume 7
      • Volume 7, Number 1 — June 2003
      • Volume 7, Number 2 — September 2003
      • Volume 7, Number 3 — December 2003
      • Volume 7, Number 4 — March 2004
    • Volume 8
      • Volume 8, Number 1 — June 2004
      • Volume 8, Number 2 — September 2004
      • Volume 8, Number 3 — December 2004
      • Volume 8, Number 4 — March 2005
    • Volume 9
      • Volume 9, Number 1 — June 2005
      • Volume 9, Number 2 — September 2005
      • Volume 9, Number 3 — December 2005
      • Volume 9, Number 4 — March 2006
    • Volume 10
      • Volume 10, Number 1 — June 2006
      • Volume 10, Number 2 — September 2006
      • Volume 10, Number 3 — December 2006
      • Volume 10, Number 4 — March 2007
    • Volume 11
      • Volume 11, Number 1 — June 2007
      • Volume 11, Number 2 — September 2007
      • Volume 11, Number 3 — December 2007
      • Volume 11, Number 4 — March 2008
    • Volume 12
      • Volume 12, Number 1 — June 2008
      • Volume 12, Number 2 — September 2008
      • Volume 12, Number 3 — December 2008
      • Volume 12, Number 4 — March 2009
    • Volume 13
      • Volume 13, Number 1 — June 2009
      • Volume 13, Number 2 — September 2009
      • Volume 13, Number 3 — December 2009
      • Volume 13, Number 4 — March 2010
    • Volume 14
      • Volume 14, Number 1 — June 2010
      • Volume 14, Number 2 – September 2010
      • Volume 14, Number 3 – December 2010
      • Volume 14, Number 4 – March 2011
    • Volume 15
      • Volume 15, Number 1 — June 2011
      • Volume 15, Number 2 — September 2011
      • Volume 15, Number 3 — December 2011
      • Volume 15, Number 4 — March 2012
  • Vols. 16-Current
    • Volume 16
      • Volume 16, Number 1 — June 2012
      • Volume 16, Number 2 — September 2012
      • Volume 16, Number 3 — December 2012
      • Volume 16, Number 4 – March 2013
    • Volume 17
      • Volume 17, Number 1 – May 2013
      • Volume 17, Number 2 – August 2013
      • Volume 17, Number 3 – November 2013
      • Volume 17, Number 4 – February 2014
    • Volume 18
      • Volume 18, Number 1 – May 2014
      • Volume 18, Number 2 – August 2014
      • Volume 18, Number 3 – November 2014
      • Volume 18, Number 4 – February 2015
    • Volume 19
      • Volume 19, Number 1 – May 2015
      • Volume 19, Number 2 – August 2015
      • Volume 19, Number 3 – November 2015
      • Volume 19, Number 4 – February 2016
    • Volume 20
      • Volume 20, Number 1 – May 2016
      • Volume 20, Number 2 – August 2016
      • Volume 20, Number 3 – November 2016
      • Volume 20, Number 4 – February 2017
    • Volume 21
      • Volume 21, Number 1 – May 2017
      • Volume 21, Number 2 – August 2017
      • Volume 21, Number 3 – November 2017
      • Volume 21, Number 4 – February 2018
    • Volume 22
      • Volume 22, Number 1 – May 2018
      • Volume 22, Number 2 – August 2018
      • Volume 22, Number 3 – November 2018
      • Volume 22, Number 4 – February 2019
    • Volume 23
      • Volume 23, Number 1 – May 2019
      • Volume 23, Number 2 – August 2019
      • Volume 23, Number 3 – November 2019
      • Volume 23, Number 4 – February 2020
    • Volume 24
      • Volume 24, Number 1 – May 2020
      • Volume 24, Number 2 – August 2020
      • Volume 24, Number 3 – November 2020
      • Volume 24, Number 4 – February 2021
    • Volume 25
      • Volume 25, Number 1 – May 2021
      • Volume 25, Number 2 – August 2021
      • Volume 25, Number 3 – November 2021
      • Volume 25, Number 4 – February 2022
    • Volume 26
      • Volume 26, Number 1 – May 2022
      • Volume 26, Number 2 – August 2022
      • Volume 26, Number 3 – November 2022
      • Volume 26, Number 4 – February 2023
    • Volume 27
      • Volume 27, Number 1 – May 2023
      • Volume 27, Number 2 – August 2023
      • Volume 27, Number 3 – November 2023
      • Volume 27, Number 4 – February 2024
    • Volume 28
      • Volume 28, Number 1 – May 2024
      • Volume 28, Number 2 – August 2024
      • Volume 28, Number 3 – November 2024
      • Volume 28, Number 4 – February 2025
    • Volume 29
      • Volume 29, Number 1 – May 2025
      • Volume 29, Number 2 – August 2025
      • Volume 29, Number 3 – November 2025
      • Volume 29, Number 4 – February 2026
    • Volume 30
      • Volume 30, Number 1 – May 2026
      • Volume 30, Number 2 – August 2026
  • Books
  • How to Submit
    • Submission Info
    • Ethical Standards for Authors and Reviewers
    • TESL-EJ Style Sheet for Authors
    • TESL-EJ Tips for Authors
    • Book Review Policy
    • Media Review Policy
    • TESL-EJ Special issues
    • APA Style Guide
  • Editorial Board
  • Support

Corpus Linguistics for Vocabulary: A Guide for Research

November 2020 – Volume 24, Number 3

Corpus Linguistics for Vocabulary: A Guide for Research

Author: Pawel Szudarski (2018) book cover
Publisher: London, New York: Routledge
Pages ISBN Price
pp. X+228 978-1-138-18722-1 (paper) $148.00 U.S.

Corpus linguistics is typically defined as the analysis of language using corpora (McEnery & Hardie, 2012). Corpora, in turn, refer to large organized collections of authentic data (spoken, written, or both) stored electronically (Tognini-Bonelli, 2001).  Corpus linguistics plays a key role in researching, teaching, and learning vocabulary because it demonstrates how language is used authentically and in different contexts. This, in turn, enables instructors to teach language more efficiently, material developers to develop more authentic materials, and students to learn and use authentic words, collocations, and patterns. Corpus Linguistics for Vocabulary: A Guide for Research by Pawel Szudarski is an introductory book written to familiarize readers with the interrelationship between corpus linguistics and vocabulary learning. This book is a helpful resource for students and teachers who wish to conduct corpus-based research on vocabulary-related questions.

Corpus Linguistics for Vocabulary contains ten chapters and begins by discussing general points concerning corpus linguistics (Chapters 1-2) before moving on to elucidate more specific uses of corpus linguistics for vocabulary purposes (Chapters 3-10). Most chapters deal with common issues of interest to researchers and teachers, such as spoken/written differences, formulaic language, academic and general vocabulary, and so on. The chapters are largely unstructured (except for a “Summary” section that exists in all of them).

Chapter 1 provides clear definitions and explanations of fundamental terms (e.g., corpus linguistics, corpora) and concepts including, but not limited to, common features of corpora, merits and demerits of corpus linguistics, and different types of corpora (e.g., technical corpora, written corpora, spoken corpora). In the second chapter, more specific corpora-related tools, such as frequency analysis, n-gram analysis, and word combinations (e.g., collocations) which might be utilized for searches in corpus linguistics, are explained. Chapter 3 introduces the concept of vocabulary learning and pertinent terms, such as “vocabulary,” “lexis,” “lexeme,” “lemma,” and “word family.” The author intends to make readers familiar with these terms, as they are necessary in running a corpus analysis and using corpora to analyze various aspects of vocabulary (e.g., synonymy, polysemy, etc.).

The rest of the book (4-10) depicts how different facets of vocabulary, such as formulaic sequences, academic lexical items, etc. can be studied via corpus linguistics. Specifically, Chapter 4 discusses how corpora can be used for finding frequencies of lexical items. It explains how frequency lists might be used by researchers (e.g., adopting stimuli lists appropriate to participants’ proficiency levels) and teachers (e.g., selecting vocabulary items appropriate to participants’ proficiency levels). Chapter 5 explains how corpus linguistics might be employed to explore formulaic language. It discusses lexical bundles, collocations, and colligations, which are the most commonly-used types of formulaic language. It further explains that corpus analysis is useful to recognize and categorize different types of formulaic sequences as integral parts of natural language use. The use of corpus linguistics in teaching vocabulary is presented in Chapter 6 via two approaches: the indirect teaching approach and the direct teaching approach. An indirect teaching approach entails consideration of corpus analysis findings (such as frequency results) in selecting and teaching lexical items. A direct teaching approach, in contrast, foregrounds the direct manipulation of corpora by learners to inductively discover and learn word meanings.

Moving on in the book, Chapter 7 raises the issue of using learner corpora to examine learner language. It depicts the ways that learner corpora can facilitate research questions about learners’ vocabulary learning and development. For example, the chapter describes the English Profile Project (EPP), a research program based on the Cambridge learner corpus which brings together a number of language institutes. It aims to obtain corpus evidence that will serve as a basis for the description of the English learner.

Chapter 8 elaborates on the use of corpora for English for specific and academic purposes. It contains discussion on specialized kinds of corpora, such as the Corpus of Contemporary American English (COCA) and Corpus Linguistics in Cheshire (CLiC) by first illuminating the characteristic features of such corpora and then demonstrating their usefulness. The chapter presents examples of their applications in different areas, including English for Specific Purposes and English for Academic Purposes. Chapter 9 highlights how pragmatics and discourse can be used to examine vocabulary learning. The chapter first discusses the advantages of combining corpus and discourse approaches to analyze lexical items (e.g., reaching more robust and contextualized findings). It then elucidates the process of how corpus techniques can help researchers with pragmatic functions of vocabulary, such as speech acts. Finally, the tenth chapter provides a condensed but comprehensive summary of the book. It also offers numerous ideas and recommendations for researchers, including how to run a corpus-based study to explore the use of language in spoken and written contexts, or how to look into the use of phrasal and single-word verbs in spoken and written contexts.

Corpus Linguistics for Vocabulary presents a novel view of corpus linguistics for vocabulary practice, introducing an array of corpus tools and procedures that are increasingly important to examine intricacies and patterns of vocabulary in natural contexts. It is a practical guide for teachers, language learners, and vocabulary researchers who have little or no experience in corpus linguistics. Another advantage of the book is that it provides authentic examples and tasks (with answers and commentary) for almost every topic raised in the chapters. It deals with real-life issues, such as authentic spoken and written conversations, speech acts, ambiguity in conversations, collocations in speech, etc.

The main shortcoming of the book is its rudimentary explanations and topics. The author could have devoted more chapters to advanced and professional issues regarding vocabulary and corpus linguistics, such as the use of different research methods (quantitative, qualitative, and mixed-methods) in running corpus-based studies to appeal to more experienced teachers and language professionals. Also, the author could have dedicated a chapter to testing and assessment-related issues of vocabulary that can be addressed through corpus analyses. For example, a chapter examining the way corpus analyses can be performed to select items for developing vocabulary tests could have been included.

Taken together, Corpus Linguistics for Vocabulary: A Guide for Research is strongly recommended to language learners, teachers, materials developers, and syllabus designers in that it is an up-to-date and insightful collection of information regarding corpus linguistics in vocabulary learning.

References

McEnery, T., & Hardie, A. (2012). Corpus Linguistics: Method, theory and practice. Cambridge University Press.

Tognini-Bonelli, E. (2001). Corpus linguistics at work. John Benjamins.

Reviewed by
Kamal Heidari
Shiraz University, Iran
<K_86_teflatmarkyahoo.com>

© Copyright rests with authors. Please cite TESL-EJ appropriately.
Editor’s Note: The HTML version contains no page numbers. Please use the PDF version of this article for citations.

© 1994–2026 TESL-EJ, ISSN 1072-4303
Copyright of articles rests with the authors.