{"id":21603,"date":"2024-07-18T17:22:34","date_gmt":"2024-07-19T01:22:34","guid":{"rendered":"https:\/\/tesl-ej.org\/wordpress\/?page_id=21603"},"modified":"2024-07-24T04:40:50","modified_gmt":"2024-07-24T12:40:50","slug":"ej110a4","status":"publish","type":"page","link":"https:\/\/tesl-ej.org\/wordpress\/issues\/volume28\/ej110\/ej110a4\/","title":{"rendered":"Text Complexity of Cambridge-delivered IELTS Academic Reading Tests: Comparability with IELTS Academic Reading Practice Tests from Other Publishers"},"content":{"rendered":"<h2 class=\"doi\">August 2024 &#8211; Volume 28, Number 2<\/h2>\n<p><strong>https:\/\/doi.org\/10.55593\/ej.28110a4<\/strong><\/p>\n<p><strong>Huu Thanh Minh Nguyen<\/strong><br \/>\nUniversity of Foreign Language Studies, The University of Danang<br \/>\n&lt;nhtminh<img loading=\"lazy\" decoding=\"async\" class=\"atmark\" src=\"http:\/\/www.tesl-ej.org\/atmark.png\" alt=\"atmark\" width=\"12\" height=\"12\" \/>ufl.udn.vn&gt;<\/p>\n<p><strong>Nguyen Van Anh Le<\/strong><br \/>\nUniversity of Foreign Language Studies, The University of Danang<br \/>\n&lt;lnvanh<img loading=\"lazy\" decoding=\"async\" class=\"atmark\" src=\"http:\/\/www.tesl-ej.org\/atmark.png\" alt=\"atmark\" width=\"12\" height=\"12\" \/>ufl.udn.vn&gt;<\/p>\n<h3 class=\"abstractnew\">Abstract<\/h3>\n<p class=\"abstractnew\">Comparing language tests and test preparation materials holds important implications for the latter\u2019s validity and reliability. However, not enough studies compare such materials across a wide range of indices. Therefore, this study investigated the text complexity of IELTS academic reading tests (IRT) and IELTS reading practice tests (IRPrT). Fine-grained quantitative analyses were undertaken to delineate measures of lexical, syntactic, and discourse complexity across a corpus of 108 IRT and 108 IRPrT published by Pearson, Macmillan, and Cengage Learning. The results suggest little difference between IRT and IRPrT at the lexical level; however, there were significant differences in some measures of syntactic and discourse level complexity. The findings bear implications for stakeholders including learners as test takers, instructors, material developers, and language testing researchers. We interpret this to mean that while IRPrT materials are lexically conducive to the practice for IRT and can provide a similar experience, the IRPrT do show some differences in the amount of subordination and idea repetition at the discourse level. Therefore, instructors and learners may seek to supplement practice with these structures when preparing for IRT , and the designers of such practice materials should consider aligning these factors in the future.<\/p>\n<p class=\"abstractnew\"><strong><em>Keywords<\/em><\/strong>: Language testing, Reading tests, Reading practice test materials, Lexical text complexity, Syntactic text complexity, Discourse text complexity<\/p>\n<p>Readability research and text-leveling schemes have emphasized the need to delineate text complexity to interpret the interaction between the reader and the text (Mesmer et al., 2012). Text complexity refers to the \u201ctext elements that can be analyzed, studied or manipulated\u201d (Mesmer et al., 2012, p. 236). According to Snow\u2019s (2002) RAND model of reading comprehension, text elements including vocabulary, syntax, and discourse affect how readers construct different representations of a text, including the surface code (i.e., the exact wording), the text base (i.e., idea units representing textual meaning), and the mental models (i.e., the way of processing textual information). For this reason, understanding text complexity at lexical, syntactic and discourse levels is necessary for research on L2 reading to better understand the relationship between texts and readers (Alderson, 2000; Guthrie et al., 2013).<\/p>\n<p>High-stakes English language proficiency tests (ELPT) are important to test takers as gate-keeping tools. One such test is the International English Language Testing System (IELTS), which is often used for selection in academic contexts (Pearson, 2019). Because of its importance, preparation for the IELTS has become important, and increasing attention has been given to practice test materials, which are generally considered helpful to test takers (Kirby, 2016; O\u2019Sullivan et al., 2019). This has given rise to an international industry of practice materials for test preparation, but since such materials are not created by the test designers themselves, there are questions about how comparable the IELTS academic reading practice tests (IRPrT) are to the actual IELTS academic reading tests (IRT) (Bachman et al., 1996; Kunnan &amp; Carr, 2017).<\/p>\n<p>However, to our knowledge, no studies have been conducted to examine whether IRPrT have similar text complexity to IRT preparation. This is an important potential gap in the literature because incongruencies between practice test materials and the actual test could potentially negatively affect learners (Green, 2007). Therefore, this study seeks to compare the text complexity of IRT with IRPrT at the lexical, syntactic and discourse levels to uncover any significant differences so that educators can make informed decisions about the use of practice materials and so that the designers of such materials can potentially revise them if necessary.<\/p>\n<h3>Literature Review<\/h3>\n<h4>Text Complexity and L2 Reading Comprehension<\/h4>\n<p>L2 readability research has been viewed from two main strands. One strand is situated in traditional readability formulas that generally measure the number of words per sentence (i.e., sentence difficulty index), the number of syllables per words (i.e., word difficulty index), and word frequency (Brown, 1998; Greenfield, 1999). However, traditional readability formulas are criticized for being restrained to measures at the word level, given that L2 readability formulas are sensitive to other elements such as syntax and rhetorical organization (Carrell, 1987). Because of this limitation, the other strand has been motivated by more recent formulas that transcend the word level to demonstrate the relationships between textual elements. According to Crossley et al. (2008), the readability formula is not only constructed of word frequency but comprises such additional measures as syntax across sentences, and cohesion at the discourse level.<\/p>\n<p>Text complexity shows a negative correlation with reading comprehension (Yang et al., 2021), although it is potential to support learner engagement with a text (Fulmer et al., 2015). Several empirical studies have revealed the relationships between lexical, syntactic, and discourse features of texts and reading comprehension. In terms of the lexical level, lexical sophistication, diversity and density are the three most widely examined properties (Read, 2000). Lexical sophistication is often considered a measure of what percentage of words the learner is likely to know (Laufer &amp; Nation, 1995; Nation, 2006). Several works have suggested that readers need to be able to understand somewhere between 95 and 98% of the words or word families in a text in order to understand it (Laufer &amp; Ravenhorst-Kalovski, 2010; Nation, 2006). Aside from lexical sophistication, higher lexical diversity, i.e., the amount of different words in the text, and density, i.e., the amount of content words, contributes to increasing difficulty in reading comprehension (Read, 2000).<\/p>\n<p>Other studies have shown that the syntactic features of a text also impact its complexity (e.g., Perfetti &amp; Stafura, 2014). For example, Shiotsu and Weir (2017) suggests that readers\u2019 syntactic knowledge accounts for more variance in the L2 reading comprehension level than lexical knowledge. This is because limited syntactic knowledge may impede the ability to understand a text despite the ability to understand meaning of single words within it (Tong et al., 2024). In addition to this, more complex sentences and grammatical structures make the sentences more difficult to be parsed, thereby possibly increasing text complexity beyond comprehension (Mesmer et al., 2012; Kyle, 2016).<\/p>\n<p>Cohesion has also been suggested as the most prominent discourse level feature to impact text complexity in L2 readability research. In general, more cohesive texts are easier to be comprehended (Gernsbacher, 2013; Graesser et al., 2004), because complex mental representations and recall are not required (Ehrlich, 1991). The use of cohesive devices is conducive to establishing textual coherence (Goldman &amp; Rakestraw, 2000);therefore, texts that are coherent at both sentence and global levels aid readers\u2019 memory of text information, increasing comprehension (Koda, 2005). However, increased text cohesion is generally accompanied by the inclusion of more information (Beck et al., 1991), which is associated with \u201cincreased text length, density, and complexity,\u2019 requiring readers to \u2018process larger amounts of text-based information&#8221; (Ozuru et al., 2009, p. 229).<\/p>\n<h4>Text Comparability between IRT and IRPrT<\/h4>\n<p>Examining the comparability between reading tests and practice test materials is significant due to washback, i.e., \u201cthe effects of tests on the teaching and learning directed toward them\u201d (Green, 2006, p. 334). According to Green\u2019s (2007) model of washback direction, washback is positive if there is a consistency \u201cbetween test design and skills developed by a curriculum or required in a target language use domain\u201d (Green, 2006, p. 339). However, if there is a large gap between the test design in terms of format, content, and complexity and what is learned during test preparation, negative washback, such as construct-irrelevant variance, may occur (Messick, 1989). Therefore, examining the comparability of texts in authentic and practice tests can provide insights into how accurately they can reflect future test performance, and whether or not they are providing sufficiently realistic reading materials in terms of text complexity.<\/p>\n<p>To our knowledge, there has been no investigation into IRPrT, except for Everett and Colman (2003). They examined the appropriateness of content in the listening and reading components of commercially produced IELTS practice tests dating from 1996 to 1998 and simply found that the vocabulary used in the IRPrT in their study varied from being unfamiliar and difficult to familiar while sentence structures ranged from complicated to uncomplicated. However, Everett and Colman (2003) only examined lexical and syntactic levels and did not look at cohesion or other discourse level factors. Furthermore, their study has become somewhat dated as it was conducted on much older materials and was limited to the technology available at the time, i.e., since such time many advances have been made in natural language processing that allow for more text features to be calculated automatically.<\/p>\n<p>Given the lack of recent research in this area, we believe it is important to reexamine the text complexity of the reading passages of test preparation materials and actual reading test passages to investigate how valid more current materials are when viewed from a wider lense of text complexity. Accordingly, this study addresses the following research questions:<\/p>\n<ol>\n<li>To what extent are IRT comparable to IRPrT in terms of lexical text complexity?<\/li>\n<li>To what extent are IRT comparable to IRPrT in terms of syntactic text complexity?<\/li>\n<li>To what extent are IRT comparable to IRPrT in terms of discourse text complexity?<\/li>\n<\/ol>\n<h3>Methodology<\/h3>\n<h4>Research Design<\/h4>\n<p>This study employed two corpora of <a href=\"https:\/\/docs.google.com\/document\/d\/12IINIpz9XBq48XakRFYgr6v2pgBtc_JK\/edit?usp=sharing&amp;ouid=107176873123998493362&amp;rtpof=true&amp;sd=true\" target=\"_blank\" rel=\"noopener\">IRT<\/a> and <a href=\"https:\/\/docs.google.com\/document\/d\/1dhyLnQcPiVa2oXXaEuPmVG6PMak-Empg\/edit?usp=sharing&amp;ouid=107176873123998493362&amp;rtpof=true&amp;sd=true\" target=\"_blank\" rel=\"noopener\">IRPrT<\/a>. The IRT were obtained from Cambridge IELTS series 9-17 published by Cambridge English Assessment. The IRPrT were extracted from the test preparation materials of three publishers:<\/p>\n<ul>\n<li>Pearson \u2013 IELTS Practice test Plus (Jakeman &amp; McDowell, 2001), IELTS Practice test Plus 2 (Terry &amp; Wilson, 2005), IELTS Practice test Plus 3 (Matthews &amp; Salisbury, 2011)<\/li>\n<li>Macmillan \u2013 IELTS Test Builder 1 (McCarter &amp; Ash, 2008), IELTS Test Builder 2 (McCarter, 2008)<\/li>\n<li>Cengage Learning \u2013 Exam Essential Practice Tests: IELTS 1 (Harrison &amp; Whitehead, 2015), Exam Essential Practice Tests: IELTS 2 (Gough &amp; Hutchison, 2015)<\/li>\n<\/ul>\n<p>The reading texts in both IRT and IRPrT vary in genres. Each corpus comprises 108 texts, with the average number of tokens in each text and the total number of tokens in total being relatively comparable (Table 1). It can therefore be argued that the compilation of both corpora was balanced and representative (McEnery, 2006).<\/p>\n<p><strong>Table 1. IRT and IRPrT corpora<\/strong><\/p>\n<table>\n<tbody>\n<tr>\n<td width=\"102\"><strong>Corpus<\/strong><\/td>\n<td width=\"108\"><strong>No. <\/strong><strong>\u00a0texts<\/strong><\/td>\n<td width=\"228\"><strong>Average tokens<br \/>\nper passage<\/strong><\/td>\n<td width=\"162\"><strong>Total tokens<\/strong><\/td>\n<\/tr>\n<tr>\n<td width=\"102\">IRT<\/td>\n<td width=\"108\">108<\/td>\n<td width=\"228\">872<\/td>\n<td width=\"162\">94,247<\/td>\n<\/tr>\n<tr>\n<td width=\"102\">IRPrT<\/td>\n<td width=\"108\">108<\/td>\n<td width=\"228\">853<\/td>\n<td width=\"162\">92,230<\/td>\n<\/tr>\n<\/tbody>\n<\/table>\n<h4>Data Collection Instruments<\/h4>\n<p>Lexical text complexity<strong>. <\/strong>Measures of lexical text complexity encompass lexical sophistication, diversity, and density (Michel, 2017). Lexical sophistication is traditionally measured by VocabProfilers to reveal the percentage of words in different frequency bands in a text (Cobb, 2009). Kim et al. (2018) extended the analysis of lexical sophistication beyond word frequency by employing the Tool for the Automatic Analysis of Lexical Sophistication (TAALES; Kyle &amp; Crossley, 2015). Through TAALES, Kim et al. (2018) included additional domains of lexical sophistication such as word range, contextual distinctiveness, word neighborhood, academic language, and so forth. Lexical diversity is traditionally measured by the type-token ratio (TTR) as the ratio of the number of different words (types) to the total number of words (tokens) (Read, 2000). As text length may influence the TTR, more robust indices were developed to measure lexical diversity. Later, Kyle et al. (2021) devised the Tool for the Automatic Analysis of Lexical Diversity (TAALED) that incorporates a wide range of measures of lexical diversity, namely the classic TTR, MTLD, MATTR, HD-D, and so forth. Lexical density, i.e., the percentage of content words in a text (Fang &amp; Pace, 2013), is also calculated by TAALED for both types and tokens. Table 2 and 3 detail lexical sophistication, diversity and density measures used in the study and adapted from Yu (2021).<\/p>\n<p><strong>Table 2. Lexical sophistication measures from TAALES<\/strong><\/p>\n<table width=\"624\">\n<tbody>\n<tr>\n<td width=\"180\"><strong>Category<\/strong><\/td>\n<td width=\"168\"><strong>Index Name<\/strong><\/td>\n<td width=\"276\"><strong>Description<\/strong><\/td>\n<\/tr>\n<tr>\n<td rowspan=\"2\" width=\"180\">Text coverage<\/td>\n<td width=\"168\">3000_level<\/td>\n<td width=\"276\">The percentage of words in the most frequent 3000-word level<\/td>\n<\/tr>\n<tr>\n<td width=\"168\">5000_level<\/td>\n<td width=\"276\">The percentage of words in the most frequent 5000-word level<\/td>\n<\/tr>\n<tr>\n<td rowspan=\"3\" width=\"180\">Word frequency \u201cBNC_Written_Freq_\u201d<\/td>\n<td width=\"168\">AW_Log<\/td>\n<td width=\"276\">BNC Written Frequency AW Logarithm<\/td>\n<\/tr>\n<tr>\n<td width=\"168\">CW_Log<\/td>\n<td width=\"276\">BNC Written Frequency CW Logarithm<\/td>\n<\/tr>\n<tr>\n<td width=\"168\">FW_Log<\/td>\n<td width=\"276\">BNC Written Frequency FW Logarithm<\/td>\n<\/tr>\n<tr>\n<td rowspan=\"3\" width=\"180\">Word range<br \/>\n\u201cBNC_Written_Range_\u201d<\/td>\n<td width=\"168\">AW<\/td>\n<td width=\"276\">BNC Written Range AW<\/td>\n<\/tr>\n<tr>\n<td width=\"168\">CW<\/td>\n<td width=\"276\">BNC Written Range CW<\/td>\n<\/tr>\n<tr>\n<td width=\"168\">FW<\/td>\n<td width=\"276\">BNC Written Range FW<\/td>\n<\/tr>\n<tr>\n<td width=\"180\">Academic language<\/td>\n<td width=\"168\">All_AWL_Normed<\/td>\n<td width=\"276\">Academic Word List All<\/td>\n<\/tr>\n<tr>\n<td rowspan=\"4\" width=\"180\">Word recognition norms<\/td>\n<td width=\"168\">LD_Mean_RT_Zscore<\/td>\n<td width=\"276\">Lexical Decision Time (z-score)<\/td>\n<\/tr>\n<tr>\n<td width=\"168\">LD_Mean_Accuracy<\/td>\n<td width=\"276\">Lexical Decision Accuracy<\/td>\n<\/tr>\n<tr>\n<td width=\"168\">WN_Zscore<\/td>\n<td width=\"276\">Word Naming Response Time (z-score)<\/td>\n<\/tr>\n<tr>\n<td width=\"168\">WN_Mean_Accuracy<\/td>\n<td width=\"276\">Word Naming Response Accuracy<\/td>\n<\/tr>\n<tr>\n<td width=\"180\">Contextual distinctiveness<\/td>\n<td width=\"168\">lsa_average_all_cosine<\/td>\n<td width=\"276\">LSA Contextual Distinctiveness<br \/>\n(all cosine)<\/td>\n<\/tr>\n<tr>\n<td width=\"180\">Age of exposure<\/td>\n<td width=\"168\">aoe_inverse_average<\/td>\n<td width=\"276\">LDA Age of Exposure (inverse average)<\/td>\n<\/tr>\n<\/tbody>\n<\/table>\n<p><strong>Table 3. Lexical density and diversity measures from TAALED<\/strong><\/p>\n<table width=\"606\">\n<tbody>\n<tr>\n<td><strong>Category<\/strong><\/td>\n<td><strong>Index Name<\/strong><\/td>\n<td><strong>Description<\/strong><\/td>\n<\/tr>\n<tr>\n<td rowspan=\"2\" style=\"vertical-align: middle;\">Lexical density<br \/>\n\u201clexical_density_\u201d<\/td>\n<td>types<\/td>\n<td>Content word types (<em>N<\/em>) divided by word types (<em>N<\/em>)<\/td>\n<\/tr>\n<tr>\n<td>tokens<\/td>\n<td>Content word tokens divided by word tokens<\/td>\n<\/tr>\n<tr>\n<td rowspan=\"3\" style=\"vertical-align: middle;\">TTR<br \/>\n\u201csimple_ttr_\u201d<\/td>\n<td>aw<\/td>\n<td>TTR for all word types<\/td>\n<\/tr>\n<tr>\n<td>cw<\/td>\n<td>TTR for content words<\/td>\n<\/tr>\n<tr>\n<td>fw<\/td>\n<td>TTR for function words<\/td>\n<\/tr>\n<tr>\n<td rowspan=\"3\" style=\"vertical-align: middle;\">MATTR<br \/>\n\u201cmattr50_\u201d<\/td>\n<td>aw<\/td>\n<td>Moving 50-word Average TTR of all words<\/td>\n<\/tr>\n<tr>\n<td>cw<\/td>\n<td>Moving 50-word Average TTR of content words<\/td>\n<\/tr>\n<tr>\n<td>fw<\/td>\n<td>Moving 50-word Average TTR of function words<\/td>\n<\/tr>\n<tr>\n<td rowspan=\"3\" style=\"vertical-align: middle;\">MTLD original<br \/>\n\u201cmtld_original_\u201d<\/td>\n<td>aw<\/td>\n<td>Average number of all tokens required to reach<br \/>\nTTR &gt;= .720<\/td>\n<\/tr>\n<tr>\n<td>cw<\/td>\n<td>Average number of all content word tokens to reach TTR &gt;= .720<\/td>\n<\/tr>\n<tr>\n<td>fw<\/td>\n<td>Average number of all function word tokens to reach TTR &gt;= .720<\/td>\n<\/tr>\n<tr>\n<td rowspan=\"3\" style=\"vertical-align: middle;\">MTLD-MA-Wrap<br \/>\n\u201cmtld_ma_wrap_\u201d<\/td>\n<td>aw<\/td>\n<td>Moving Average of MTLD (all words)<\/td>\n<\/tr>\n<tr>\n<td>cw<\/td>\n<td>Moving Average of MTLD (content words)<\/td>\n<\/tr>\n<tr>\n<td>fw<\/td>\n<td>Moving Average of MTLD (function words)<\/td>\n<\/tr>\n<\/tbody>\n<\/table>\n<p id=\"back1\">Syntactic text complexity<strong>. <\/strong>Common measures of syntactic complexity include the mean length of clause (MLC), T-unit [<a href=\"#note\">1<\/a>] (MLTU), and sentence (MLS) (Vajjala &amp; Meurers, 2013). The number of dependent clauses, complex T-units and elaborated phrasal structures (e.g., complex nominals, verb phrases) per clause, T-unit and sentence has also been employed to measure syntactic complexity (Lu, 2010; Ortega, 2003; Yu, 2021). Lu (2010) created the L2 Syntactic Complexity Analyzer (L2SCA) that incorporates 14 syntactic complexity measures subsumed into five categories: (a) length of production unit, (b) sentence complexity ratio, (c) the amount of subordination, (d) the amount of coordination, and (e) particular syntactic structures. Table 4 details measures of syntactic complexity in this study.<\/p>\n<p><strong>Table 4. Syntactic complexity measures from L2SCA<\/strong><\/p>\n<table width=\"600\">\n<tbody>\n<tr>\n<td width=\"210\"><strong>Category<\/strong><\/td>\n<td width=\"228\"><strong>Measure<\/strong><\/td>\n<td width=\"162\"><strong>Index Name<\/strong><\/td>\n<\/tr>\n<tr>\n<td rowspan=\"3\" width=\"210\" style=\"vertical-align: middle;\">Length of the<br \/>\nproduction unit<\/td>\n<td width=\"228\">Mean length of clause<\/td>\n<td width=\"162\">MLC<\/td>\n<\/tr>\n<tr>\n<td width=\"228\">Mean length of sentence<\/td>\n<td width=\"162\">MLS<\/td>\n<\/tr>\n<tr>\n<td width=\"228\">Mean length of T-unit<\/td>\n<td width=\"162\">MLT<\/td>\n<\/tr>\n<tr>\n<td width=\"210\">Sentence complexity<\/td>\n<td width=\"228\">Sentence complexity ratio<\/td>\n<td width=\"162\">C\/S<\/td>\n<\/tr>\n<tr>\n<td rowspan=\"4\" width=\"210\" style=\"vertical-align: middle;\">Subordination<\/td>\n<td width=\"228\">T-unit complexity ratio<\/td>\n<td width=\"162\">C\/T<\/td>\n<\/tr>\n<tr>\n<td width=\"228\">Complex T-unit ratio<\/td>\n<td width=\"162\">CT\/T<\/td>\n<\/tr>\n<tr>\n<td width=\"228\">Dependent clause ratio<\/td>\n<td width=\"162\">DC\/C<\/td>\n<\/tr>\n<tr>\n<td width=\"228\">Dependent clause per T-unit<\/td>\n<td width=\"162\">DC\/T<\/td>\n<\/tr>\n<tr>\n<td rowspan=\"3\" width=\"210\" style=\"vertical-align: middle;\">Coordination<\/td>\n<td width=\"228\">Coordinate phrases per clause<\/td>\n<td width=\"162\">CP\/C<\/td>\n<\/tr>\n<tr>\n<td width=\"228\">Coordinate phrases per T-unit<\/td>\n<td width=\"162\">CP\/T<\/td>\n<\/tr>\n<tr>\n<td width=\"228\">Sentence coordination ratio<\/td>\n<td width=\"162\">T\/S<\/td>\n<\/tr>\n<tr>\n<td rowspan=\"3\" width=\"210\" style=\"vertical-align: middle;\">Particular structures<\/td>\n<td width=\"228\">Complex nominals per clause<\/td>\n<td width=\"162\">CN\/C<\/td>\n<\/tr>\n<tr>\n<td width=\"228\">Complex nominals per T-unit<\/td>\n<td width=\"162\">CN\/T<\/td>\n<\/tr>\n<tr>\n<td width=\"228\">Verb phrases per T-unit<\/td>\n<td width=\"162\">VP\/T<\/td>\n<\/tr>\n<\/tbody>\n<\/table>\n<p>Kyle (2016) extended the examination of syntactic complexity to include phrasal complexity (i.e., absolute complexity; Bult\u00e9 &amp; Housen, 2012) to fully unfold the syntactic features of a text, and created the Tool for the Automatic Analysis of Syntactic Sophistication and Complexity (TAASC) to automatically calculate 190 indices of fine-grained clausal and phrasal complexity. In this study, we adopted 15 indices that calculate the average number of particular structures per clause. Auxiliary verbs, bare noun phrase temporal modifiers, negation and discourse markers, existential &#8220;there&#8221;, parataxis, modals, agents, passive auxiliaries, passive clausal and nominal subjects, phrasal verb particles, and undefined dependents are all structures removed from analyses because they contribute less to clausal complexity than other structures measured by TAASSC, as justified below:<\/p>\n<ul>\n<li>Auxiliary verbs are functional elements that only marginally increase clausal complexity because they are features of obligatory inflection which reflect grammatical rather than lexical meanings (Biber et al., 2014).<\/li>\n<li>Bare noun phrase temporal modifiers often behave similarly to functional elements signifying time rather than demonstrating syntactic patterns.<\/li>\n<li>Negation markers and discourse markers are better classified as features of textual cohesion rather than clausal complexity (Halliday &amp; Hasan, 2014).<\/li>\n<li>The existential &#8220;there&#8221; is a fixed construction that does not demonstrate syntactic flexibility and instead exemplify an ideational metafunction of text rather than interpersonal adaptations (Biber et al., 2013).<\/li>\n<li>Parataxis is ambiguous, \u201csometimes competing with coordination, sometimes with subordination, for the same semantic niche in language\u201d (Hoeksema &amp; Napoli, 1993, p. 291), and therefore is rather limited to additive but disjointed structures.<\/li>\n<li>Modals indicate mood, which is more of a morphological structure that expresses viewpoint than a structure indicating syntactic sophistication (Bardovi-Harlig, 2000).<\/li>\n<li>Agents and passive auxiliaries are also more morphological than syntactic. Agents demonstrate relational processes rather than clausal transformations according to Halliday&#8217;s transitivity system (Thompson, 2014). Passive auxiliaries accompany the passive construction rather than exemplifying complexity in their own right.<\/li>\n<li>Passive clausal and nominal subjects are byproducts of passivization &#8211; the passive itself is the sophisticated structure and already represented elsewhere.<\/li>\n<li>Phrasal verb particles are minimal functional elements that add meaning and therefore are better understood as increasing semantic complexity rather than syntactic complexity (Garniner &amp; Schmitt, 2015; Spring, 2019). Halliday (2004) characterizes particles as instantiating circumstantial features of verb group rather than increasing complexity.<\/li>\n<li>Undefined dependents are an artifact of parsing errors rather than sophisticated syntax and such parsing inaccuracies do not genuinely reflect a learner&#8217;s syntactic competence (Briscoe et al., 2010).<\/li>\n<\/ul>\n<p>We also used 15 indices of phrasal complexity related to: (a) the average number of dependents per each of the seven noun phrase types and (b) the occurrence of particular independent types regardless of the nominal phrases they occur in. Table 5 details measures of clausal and phrasal complexity employed in this study.<\/p>\n<p><strong>Table 5. Clausal and phrasal complexity from TAASSC<\/strong><\/p>\n<table>\n<tbody>\n<tr>\n<td><strong>Index Name<\/strong><\/td>\n<td><strong>Description <\/strong><\/td>\n<td colspan=\"2\"><strong>Example (Kyle, 2016, p. 55)<\/strong><\/td>\n<\/tr>\n<tr>\n<td colspan=\"4\"><strong>Clausal complexity<\/strong><\/td>\n<\/tr>\n<tr>\n<td>acomp<\/td>\n<td colspan=\"2\">Adjective complement<\/td>\n<td><em>She looks [beautiful]<sub>acomp<\/sub><\/em><\/td>\n<\/tr>\n<tr>\n<td>advcl<\/td>\n<td colspan=\"2\">Adverbial clauses<\/td>\n<td><em>The accident happened [as night fell]<sub>advcl<\/sub><\/em><\/td>\n<\/tr>\n<tr>\n<td>advmod<\/td>\n<td colspan=\"2\">Adverbial modifier<\/td>\n<td><em>[Accordingly]<sub>advmod<\/sub>, I ate pizza<\/em><\/td>\n<\/tr>\n<tr>\n<td>ccomp<\/td>\n<td colspan=\"2\">Clausal complement<\/td>\n<td><em>I am certain [that he did it]<sub>ccomp<\/sub><\/em><\/td>\n<\/tr>\n<tr>\n<td>cc<\/td>\n<td colspan=\"2\">Clausal coordination<\/td>\n<td><em>Jill runs and [Jack jumps]<sub>cc<\/sub><\/em><\/td>\n<\/tr>\n<tr>\n<td>conj<\/td>\n<td colspan=\"2\">Conjunction<\/td>\n<td><em>He runs and [jumps]<sub>conj<\/sub><\/em><\/td>\n<\/tr>\n<tr>\n<td>mark<\/td>\n<td colspan=\"2\">Subordinating conjunction<\/td>\n<td><em>Forces engaged in fighting [after]<sub>mark<\/sub> insurgents attacked<\/em><\/td>\n<\/tr>\n<tr>\n<td>pcomp<\/td>\n<td colspan=\"2\">Prepositional complement<\/td>\n<td><em>They heard about [you missing classes]<sub>pcomp<\/sub><\/em><\/td>\n<\/tr>\n<tr>\n<td>csubj<\/td>\n<td colspan=\"2\">Clausal subject<\/td>\n<td><em>[What she said]<sub>csubj<\/sub> is not true<\/em><\/td>\n<\/tr>\n<tr>\n<td>xsubj<\/td>\n<td colspan=\"2\">Controlling subject<\/td>\n<td><em>[Tom]<sub>xsubj<\/sub> likes to eat fish<\/em><\/td>\n<\/tr>\n<tr>\n<td>nsubj<\/td>\n<td colspan=\"2\">Nominal subject<\/td>\n<td><em>The [baby]<sub>nsubj<\/sub> is cute<\/em><\/td>\n<\/tr>\n<tr>\n<td>dobj<\/td>\n<td colspan=\"2\">Direct object<\/td>\n<td><em>She gave me [a raise]<sub>dobj<\/sub><\/em><\/td>\n<\/tr>\n<tr>\n<td>iobj<\/td>\n<td colspan=\"2\">Indirect object<\/td>\n<td><em>She gave [me]<sub>iobj<\/sub> a raise<\/em><\/td>\n<\/tr>\n<tr>\n<td>ncomp<\/td>\n<td colspan=\"2\">Nominal compliment<\/td>\n<td><em>He is [a teacher]<sub>ncomp<\/sub><\/em><\/td>\n<\/tr>\n<tr>\n<td>xcomp<\/td>\n<td colspan=\"2\">Open clausal compliment<\/td>\n<td><em>I am ready [to leave]<sub>xcomp<\/sub><\/em><\/td>\n<\/tr>\n<tr>\n<td colspan=\"4\"><strong>Phrasal complexity<\/strong><\/td>\n<\/tr>\n<tr>\n<td>nsubj_deps<\/td>\n<td colspan=\"2\">Dependents per nominal subject<\/td>\n<td><em>[[The]<sub>deps<\/sub> man [in the red hat]<sub>deps<\/sub>]<sub>nsubj<\/sub> gave the tall man the money.<\/em><\/td>\n<\/tr>\n<tr>\n<td>ncomp_deps<\/td>\n<td colspan=\"2\">Dependents per nominal compliment<\/td>\n<td><em>He is [[a]<sub>deps<\/sub> [tall]<sub>deps<\/sub> man]<sub>ncomp<\/sub><\/em><\/td>\n<\/tr>\n<tr>\n<td>dobj_deps<\/td>\n<td colspan=\"2\">Dependents per direct object<\/td>\n<td><em>The man in the red hat gave the tall man [[the]<sub>deps<\/sub> money]<sub>dobj<\/sub><\/em><\/td>\n<\/tr>\n<tr>\n<td>iobj_deps<\/td>\n<td colspan=\"2\">Dependents per indirect object<\/td>\n<td><em>The man in the red hat gave [[the]<sub>deps<\/sub> [tall]<sub>deps<\/sub> man]<sub>iobj<\/sub> the money<\/em><\/td>\n<\/tr>\n<tr>\n<td>pobj_deps<\/td>\n<td colspan=\"2\">Dependents per prepositional object<\/td>\n<td><em>The man in [[the]<sub>deps<\/sub> [red]<sub>deps<\/sub> hat]<sub>pobj<\/sub> gave the tall man the money<\/em><\/td>\n<\/tr>\n<tr>\n<td>det_nominal<\/td>\n<td colspan=\"2\">Determiners per nominal phrases<\/td>\n<td><em>[The]<sub>det<\/sub> man in [the]<sub>det<\/sub> red hat gave [the]<sub>det<\/sub> tall man [the]<sub>det<\/sub> money<\/em><\/td>\n<\/tr>\n<tr>\n<td>amod_nominal<\/td>\n<td colspan=\"2\">Adjective modifiers per nominal phrases<\/td>\n<td><em>The man in the [red]<sub>amod<\/sub> hat gave the [tall]<sub>amod<\/sub> man the money<\/em><\/td>\n<\/tr>\n<tr>\n<td>prep_nominal<\/td>\n<td colspan=\"2\">Prepositional phrases per nominal phrases<\/td>\n<td><em>The man [in the red hat]<sub>prep<\/sub> gave the tall man the money<\/em><\/td>\n<\/tr>\n<tr>\n<td>poss_nominal<\/td>\n<td colspan=\"2\">Possessives per nominal phrases<\/td>\n<td><em>That is [her]<sub>poss<\/sub> red car<\/em><\/td>\n<\/tr>\n<tr>\n<td>vmod_nominal<\/td>\n<td colspan=\"2\">Verbal modifiers per nominal phrases<\/td>\n<td><em>I don\u2019t have anything [to say]<sub>vmod<\/sub> to you<\/em><\/td>\n<\/tr>\n<tr>\n<td>nn_nominal<\/td>\n<td colspan=\"2\">Nouns as modifiers per nominal phrases<\/td>\n<td><em>[Oil]<sub>nn<\/sub> prices are rising<\/em><\/td>\n<\/tr>\n<tr>\n<td>rcmod_nominal<\/td>\n<td colspan=\"2\">Relative clause modifiers per nominal phrases<\/td>\n<td><em>I saw the man [you love]<sub>rcmod<\/sub><\/em><\/td>\n<\/tr>\n<tr>\n<td>advmod_nominal<\/td>\n<td colspan=\"2\">Adverbial modifiers per nominal phrases<\/td>\n<td><em>We will drive the red car [tomorrow]<sub>advmod<\/sub><\/em><\/td>\n<\/tr>\n<tr>\n<td>conj_and_nominal<\/td>\n<td colspan=\"2\">Conjunctions \u201cand\u201d per nominal phrases<\/td>\n<td><em>Jack [and]<sub>conj_and<\/sub> Jill<\/em><\/td>\n<\/tr>\n<tr>\n<td>conj_or_nominal<\/td>\n<td colspan=\"2\">Conjunctions \u201cor\u201d per nominal phrases<\/td>\n<td><em>Jack [or]<sub>conj_or<\/sub> Jill<\/em><\/td>\n<\/tr>\n<tr>\n<td><\/td>\n<td><\/td>\n<td width=\"12\"><\/td>\n<td><\/td>\n<\/tr>\n<\/tbody>\n<\/table>\n<p>Discourse text complexity. Coh-Metrix is commonly used to analyze cohesion among lexical, syntactic, and semantic properties of texts (Graesser et al., 2004). However, Coh-Metrix has a limited number of indices and does not allow for batch processing. Crossley et al. (2016), therefore, introduced the Tool for the Automatic Analysis of Cohesion (TAACO) that enables batch processing and the examination of local (i.e., sentence-level), global (paragraph-level), and overall text (i.e., text-level) cohesion. Table 6 details measures of cohesion in this study.<\/p>\n<p><strong>Table 6. Cohesion measures from TAACO<\/strong><\/p>\n<table>\n<tbody>\n<tr>\n<td><strong>Index Name<\/strong><\/td>\n<td><strong>Description<\/strong><\/td>\n<\/tr>\n<tr>\n<td colspan=\"2\"><strong>Lexical overlap<\/strong><strong> (<\/strong><strong>Local and global<\/strong><strong>)<\/strong> <strong>&#8220;adja<\/strong><strong>cent<\/strong><strong>_overlap_\u201d<\/strong><\/td>\n<\/tr>\n<tr>\n<td>2_all_sent<\/td>\n<td id=\"back2\">Lemma [<a href=\"#note\">2<\/a>] types that occur at least once in the next two sentences<\/td>\n<\/tr>\n<tr>\n<td>2_argument_sent<\/td>\n<td>Noun and pronoun lemma types that occur at least once in the next two sentences<\/td>\n<\/tr>\n<tr>\n<td>binary_all_sent<\/td>\n<td>Sentences with any lemma types overlapping with the next sentences<\/td>\n<\/tr>\n<tr>\n<td>binary_argument_sent<\/td>\n<td>Sentences with any noun and pronoun lemma overlapping with the next sentences<\/td>\n<\/tr>\n<tr>\n<td>2_all_para<\/td>\n<td>Lemma types repeated between paragraphs<\/td>\n<\/tr>\n<tr>\n<td>2_argument_para<\/td>\n<td>Noun and pronoun lemma types repeated between paragraphs<\/td>\n<\/tr>\n<tr>\n<td>binary_all_para<\/td>\n<td>Paragraphs with any lemma types overlapping with the next paragraphs<\/td>\n<\/tr>\n<tr>\n<td>binary_argument_para<\/td>\n<td>Paragraphs with any noun and pronoun lemma types overlapping with the next paragraphs<\/td>\n<\/tr>\n<tr>\n<td colspan=\"2\"><strong>Semantic overlap<\/strong> <strong>(Local and global)<\/strong><\/td>\n<\/tr>\n<tr>\n<td>lsa_1_all_sent<\/td>\n<td>Average latent semantic analysis cosine similarity between all adjacent sentences (with a one-sentence interval)<\/td>\n<\/tr>\n<tr>\n<td>lsa_2_all_sent<\/td>\n<td>\u201c\u201d (with a two-sentence interval)<\/td>\n<\/tr>\n<tr>\n<td>lsa_1_all_para<\/td>\n<td>\u201c\u201d all adjacent paragraphs (with a one-paragraph interval)<\/td>\n<\/tr>\n<tr>\n<td>lsa_2_all_para<\/td>\n<td>\u201c\u201d (with a two-paragraph interval)<\/td>\n<\/tr>\n<tr>\n<td colspan=\"2\"><strong>Connectives (Text cohesion)<\/strong><\/td>\n<\/tr>\n<tr>\n<td>basic_connectives<\/td>\n<td>Basic connectives (e.g., for, and, or)<\/td>\n<\/tr>\n<tr>\n<td>all_demonstratives<\/td>\n<td>Demonstratives (e.g., this, that, these, those)<\/td>\n<\/tr>\n<tr>\n<td>all_additive<\/td>\n<td>Additive connectives (e.g., after all)<\/td>\n<\/tr>\n<tr>\n<td>all_logical<\/td>\n<td>Logical connectives (e.g., consequently)<\/td>\n<\/tr>\n<tr>\n<td colspan=\"2\" id=\"back3\"><strong>Givenness [<a href=\"#note\">3<\/a>] (Text cohesion)<\/strong><\/td>\n<\/tr>\n<tr>\n<td>repeated_content_lemma<\/td>\n<td>Content words repeated at least once divided by all words in the text<\/td>\n<\/tr>\n<tr>\n<td>repeated_content_<br \/>\nand_pronoun_lemma<\/td>\n<td>Content words and third person pronouns repeated at least once divided by all words in the text<\/td>\n<\/tr>\n<\/tbody>\n<\/table>\n<h4>Data Analysis<\/h4>\n<p>TAALES (version 2.2), TAALED (version 1.4.1), TAASSC (version 1.3.8, including all the indices in L2SCA), and TAACO (version 2.1.3) were employed to generate data on all measures of text complexity at the lexical, syntactic and discourse levels. The datasets were compiled into spreadsheets, enumerating results per text across IRT and IRPrT sources.<\/p>\n<p>Shapiro-Wilk tests were conducted on each measure to check for normality. Independent t-tests were used when the measure was found to be normal in both corpora, and a Mann-Whitney test was used when the measure was not found to be normal in either one or both corpora (Sainani, 2012). Effect size was measured as Cohen\u2019s <em>d<\/em> for parametric measures and Spearman\u2019s Rho <em>rs<\/em> for non-parametric measures. These were interpreted according to Plonsky and Oswald (2014).<\/p>\n<h3>Findings<\/h3>\n<h4>Lexical Text Complexity<\/h4>\n<p>Table 7 shows the results of the comparisons of lexical sophistication between the two corpora. IRT and IRPrT did not significantly differ in text coverage for the 3,000-word and 5,000-word levels. There was also no significant difference in word frequency and word range between IRT and IRPrT. In terms of academic language, IRT and IRPrt had no significant difference in the frequency of use for academic words and phrases. There was also no significant difference between IRT and IRPrT regarding word recognition, contextual distinctiveness and age of exposure.<\/p>\n<p>Taken together, IRT and IRPrT indicated no significant difference in lexical sophistication, density and diversity. It is therefore concluded that IRT and IRPrT do not differ in terms of text complexity at the lexical level.<\/p>\n<p><strong>Table 7. Comparisons of measures of lexical sophistication <\/strong><\/p>\n<table>\n<tbody>\n<tr>\n<td rowspan=\"2\" style=\"vertical-align: middle;\"><strong>Category<\/strong><\/td>\n<td rowspan=\"2\" style=\"vertical-align: middle;\"><strong>Indices<\/strong><\/td>\n<td colspan=\"2\" style=\"text-align: center;\"><strong><em>M<\/em><\/strong><strong> (<em>SD<\/em>)<\/strong><\/td>\n<td rowspan=\"2\" style=\"vertical-align: middle;\"><strong>Significance Testing<\/strong><\/td>\n<\/tr>\n<tr>\n<td style=\"text-align: center;\"><strong>IRT<\/strong><\/td>\n<td style=\"text-align: center;\"><strong>IRPrT<\/strong><\/td>\n<\/tr>\n<tr>\n<td rowspan=\"2\" style=\"vertical-align: middle;\">Text coverage<\/td>\n<td>3000_level<\/td>\n<td>110.62 (2.468)<\/td>\n<td>106.38 (2.507)<\/td>\n<td><em>z<\/em> = -.499, <em>p<\/em> = .62,<em> rs<\/em> = .03<\/td>\n<\/tr>\n<tr>\n<td>5000_level<\/td>\n<td>111.23 (1.546)<\/td>\n<td>105.77 (1.701)<\/td>\n<td><em>z<\/em> = -.642, <em>p<\/em> = .52,<em> rs<\/em> = .04<\/td>\n<\/tr>\n<tr>\n<td rowspan=\"3\" style=\"vertical-align: middle;\">Word frequency<\/td>\n<td>AW_Log<\/td>\n<td>-.23 (.085)<\/td>\n<td>-.22 (.091)<\/td>\n<td><em>t<\/em> = -.953, <em>p<\/em> = .34, <em>d<\/em> = -.70<\/td>\n<\/tr>\n<tr>\n<td>CW_Log<\/td>\n<td>-1.03 (.120)<\/td>\n<td>-1.01 (.124)<\/td>\n<td><em>t<\/em> = -1.524, <em>p<\/em> = .13, <em>d<\/em> = -.21<\/td>\n<\/tr>\n<tr>\n<td>FW_Log<\/td>\n<td>1.05 (.073)<\/td>\n<td>1.03 (.068)<\/td>\n<td><em>t<\/em> = -1.313, <em>p<\/em> = .19, <em>d<\/em> = .17<\/td>\n<\/tr>\n<tr>\n<td rowspan=\"3\" style=\"vertical-align: middle;\">Word range<\/td>\n<td>AW<\/td>\n<td>69.62 (2.714)<\/td>\n<td>70.10 (2.901)<\/td>\n<td><em>t<\/em> = -1.234, <em>p<\/em> = .22, <em>d<\/em> = .18<\/td>\n<\/tr>\n<tr>\n<td>CW<\/td>\n<td>51.72 (4.061)<\/td>\n<td>52.46 (4.154)<\/td>\n<td><em>t<\/em> = -1.321, <em>p<\/em> = .19, <em>d<\/em> = .17<\/td>\n<\/tr>\n<tr>\n<td>FW<\/td>\n<td>110.19 (.544)<\/td>\n<td>106.81 (.497)<\/td>\n<td><em>z<\/em> = -.396, <em>p<\/em> = .69,<em> rs<\/em> = .03<\/td>\n<\/tr>\n<tr>\n<td style=\"vertical-align: middle;\">Academic language<\/td>\n<td>All AWL<\/td>\n<td>110.47 (.028)<\/td>\n<td>106.53 (.025)<\/td>\n<td><em>z<\/em> = -.396, <em>p<\/em> = .69,<em> rs<\/em> = .03<\/td>\n<\/tr>\n<tr>\n<td rowspan=\"4\" style=\"vertical-align: middle;\">Word recognition norms<\/td>\n<td>LD_RT<\/td>\n<td>-.533 (.018)<\/td>\n<td>-.532 (.02)<\/td>\n<td><em>t<\/em> = -.087, <em>p<\/em> = .93, <em>d<\/em> = .05<\/td>\n<\/tr>\n<tr>\n<td>LD_Acc.<\/td>\n<td>114.84 (.004)<\/td>\n<td>102.16 (.004)<\/td>\n<td><em>z<\/em> = -1.497, <em>p<\/em> = .14,<em> rs<\/em> =.10<\/td>\n<\/tr>\n<tr>\n<td>WN_Zscore<\/td>\n<td>-.488 (.019)<\/td>\n<td>-.492 (.02)<\/td>\n<td><em>t<\/em> = 1.484, <em>p<\/em> = .14, <em>d<\/em> = .21<\/td>\n<\/tr>\n<tr>\n<td>WN_Acc.<\/td>\n<td>115.01 (.002)<\/td>\n<td>101.99 (.002)<\/td>\n<td><em>z<\/em> = -1.560, <em>p<\/em> = .12,<em> rs<\/em> = .11<\/td>\n<\/tr>\n<\/tbody>\n<\/table>\n<p>Table 8 shows the results of the comparisons of lexical density between the two corpora. IRT and IRPrT did not significantly differ in the rate of content word types and tokens used.<\/p>\n<p><strong>Table 8.<\/strong> <strong>Comparisons of measures of lexical density<\/strong><\/p>\n<table>\n<tbody>\n<tr>\n<td rowspan=\"2\" width=\"87\" style=\"vertical-align: middle;\"><strong>Category<\/strong><\/td>\n<td rowspan=\"2\" width=\"98\" style=\"vertical-align: middle;\"><strong>Indices<\/strong><\/td>\n<td colspan=\"2\" width=\"224\" style=\"text-align: center;\"><strong><em>M<\/em><\/strong><strong> (<em>SD<\/em>)<\/strong><\/td>\n<td rowspan=\"2\" width=\"192\" style=\"vertical-align: middle;\"><strong>Significance Testing<\/strong><\/td>\n<\/tr>\n<tr>\n<td width=\"114\" style=\"text-align: center;\"><strong>IRT<\/strong><\/td>\n<td width=\"110\" style=\"text-align: center;\"><strong>IRPrT<\/strong><\/td>\n<\/tr>\n<tr>\n<td rowspan=\"2\" width=\"87\" style=\"vertical-align: middle;\">Lexical density<\/td>\n<td width=\"98\">types<\/td>\n<td width=\"114\">.766 (.026)<\/td>\n<td width=\"110\">.762 (.025)<\/td>\n<td width=\"192\"><em>t<\/em> = 1.098, <em>p<\/em> = .27, <em>d<\/em> = .16<\/td>\n<\/tr>\n<tr>\n<td width=\"98\">tokens<\/td>\n<td width=\"114\">.502 (.032)<\/td>\n<td width=\"110\">.500 (.032)<\/td>\n<td width=\"192\"><em>t<\/em> = .537, <em>p<\/em> = .59, <em>d<\/em> = .06<\/td>\n<\/tr>\n<\/tbody>\n<\/table>\n<p>Table 9 shows the results of the comparisons of lexical diversity between the two corpora. No significant differences were observed in any of the measures of TTR, MATTR, MTLD, MTLD-MA-Wrap between IRT and IRPrT.<\/p>\n<p><strong>Table 9.<\/strong> <strong>Comparisons of measures of lexical diversity<\/strong><\/p>\n<table>\n<tbody>\n<tr>\n<td rowspan=\"2\" style=\"vertical-align: middle;\"><strong>Category<\/strong><\/td>\n<td rowspan=\"2\" style=\"vertical-align: middle;\"><strong>Indices<\/strong><\/td>\n<td colspan=\"2\" style=\"text-align: center;\"><strong><em>M<\/em><\/strong><strong> (<em>SD<\/em>)<\/strong><\/td>\n<td rowspan=\"2\" style=\"vertical-align: middle;\"><strong>Significance Testing<\/strong><\/td>\n<\/tr>\n<tr>\n<td style=\"text-align: center;\"><strong>IRT<\/strong><\/td>\n<td style=\"text-align: center;\"><strong>IRPrT<\/strong><\/td>\n<\/tr>\n<tr>\n<td rowspan=\"3\" style=\"vertical-align: middle;\">TTR<\/td>\n<td>aw<\/td>\n<td>.430 (.031)<\/td>\n<td>.434 (.034)<\/td>\n<td><em>t<\/em> = -1.509, <em>p<\/em> = .29, <em>d<\/em> = .12<\/td>\n<\/tr>\n<tr>\n<td>cw<\/td>\n<td>104.81 (.058)<\/td>\n<td>112.19 (.068)<\/td>\n<td><em>z<\/em> = -.868, <em>p<\/em> = .39,<em> rs<\/em> = .06<\/td>\n<\/tr>\n<tr>\n<td>fw<\/td>\n<td>.202 (.020)<\/td>\n<td>.206 (.019)<\/td>\n<td><em>t<\/em> = -1.669, <em>p<\/em> = .10, <em>d<\/em> = .20<\/td>\n<\/tr>\n<tr>\n<td rowspan=\"3\" style=\"vertical-align: middle;\">MATTR<\/td>\n<td>aw<\/td>\n<td>.793 (.026)<\/td>\n<td>.794 (.023)<\/td>\n<td><em>t<\/em> = -.280, <em>p<\/em> = .78, <em>d<\/em> = .04<\/td>\n<\/tr>\n<tr>\n<td>cw<\/td>\n<td>106.47 (.033)<\/td>\n<td>110.53 (.037)<\/td>\n<td><em>z<\/em> = -.477, <em>p<\/em> = .63,<em> rs<\/em> = .03<\/td>\n<\/tr>\n<tr>\n<td>fw<\/td>\n<td>.517 (.041)<\/td>\n<td>.519 (.037)<\/td>\n<td><em>t<\/em> = -.398, <em>p<\/em> = .69, <em>d<\/em> = .05<\/td>\n<\/tr>\n<tr>\n<td rowspan=\"3\" style=\"vertical-align: middle;\">MTLD<\/td>\n<td>aw<\/td>\n<td>109.75 (16.552)<\/td>\n<td>107.25 (15.862)<\/td>\n<td><em>z<\/em> = -.295, <em>p<\/em> = .77,<em> rs<\/em> = .02<\/td>\n<\/tr>\n<tr>\n<td>cw<\/td>\n<td>106.39 (98.404)<\/td>\n<td>110.61 (121.152)<\/td>\n<td><em>z<\/em> = -.496, <em>p<\/em> = .62,<em> rs<\/em> = .03<\/td>\n<\/tr>\n<tr>\n<td>fw<\/td>\n<td>104.39 (2.996)<\/td>\n<td>112.07 (2.613)<\/td>\n<td><em>z<\/em> = -.839, <em>p<\/em> = .40,<em> rs<\/em> = .06<\/td>\n<\/tr>\n<tr>\n<td rowspan=\"3\" style=\"vertical-align: middle;\">MTLD-MA-Wrap<\/td>\n<td>aw<\/td>\n<td>109.59 (16.771)<\/td>\n<td>107.41 (15.628)<\/td>\n<td><em>z<\/em> = -.257, <em>p<\/em> = .80,<em> rs<\/em> = .02<\/td>\n<\/tr>\n<tr>\n<td>cw<\/td>\n<td>106.62 (89.892)<\/td>\n<td>110.38 (104.612)<\/td>\n<td><em>z<\/em> = -.443, <em>p<\/em> = .66,<em> rs<\/em> = .03<\/td>\n<\/tr>\n<tr>\n<td>fw<\/td>\n<td>104.68 (2.885)<\/td>\n<td>112.32 (2.672)<\/td>\n<td><em>z<\/em> = -.898, <em>p<\/em> = .37,<em> rs<\/em> = .06<\/td>\n<\/tr>\n<\/tbody>\n<\/table>\n<h4>Syntactic Text Complexity<\/h4>\n<p>Table 10 shows the comparisons of syntactic complexity between the two corpora. The IRT exhibited greater length of t-units (MLT) with more verb phrases per t-unit than IRPrT, with both showing small effect sizes (<em>rs<\/em> = 0.15 and <em>rs<\/em> = 0.18, respectively), making these results somewhat inconclusive. Furthermore, the IRT exhibited more subordination than the IRPrT for all measuers with small effect sizes: C\/T, <em>rs<\/em> = 0.15; CT\/T, <em>d<\/em> = 0.33; DC\/C, <em>d<\/em> = 0.30; DC\/T, <em>rs<\/em> = 0.16.<\/p>\n<p><strong>Table 10<\/strong><strong>. <\/strong><strong>Comparisons of measures of<\/strong> <strong>syntactic complexity<\/strong><\/p>\n<table style=\"margin-bottom: 2px;\">\n<tbody>\n<tr>\n<td rowspan=\"2\" style=\"vertical-align: middle;\"><strong>Category<\/strong><\/td>\n<td rowspan=\"2\" style=\"vertical-align: middle;\"><strong>Indices<\/strong><\/td>\n<td colspan=\"2\" style=\"text-align: center;\"><strong><em>M<\/em><\/strong><strong> (<em>SD<\/em>)<\/strong><\/td>\n<td rowspan=\"2\" style=\"vertical-align: middle;\"><strong>Significance Testing<\/strong><\/td>\n<\/tr>\n<tr>\n<td style=\"text-align: center;\"><strong>IRT<\/strong><\/td>\n<td style=\"text-align: center;\"><strong>IRPrT<\/strong><\/td>\n<\/tr>\n<tr>\n<td rowspan=\"3\" style=\"vertical-align: middle;\">Length of production unit<\/td>\n<td>MLC<\/td>\n<td>109.03 (1.944)<\/td>\n<td>107.97 (1.844)<\/td>\n<td><em>z<\/em> = -.124, <em>p<\/em> = .90,<em> rs<\/em> = .01<\/td>\n<\/tr>\n<tr>\n<td>MLS<\/td>\n<td>23.459 (3.010)<\/td>\n<td>22.852 (3.284)<\/td>\n<td><em>t<\/em> = 1.416, <em>p<\/em> = .16, <em>d<\/em> = .20<\/td>\n<\/tr>\n<tr>\n<td>MLT<\/td>\n<td>118.15 (2.914)<\/td>\n<td>98.85 (3.052)<\/td>\n<td><em>z<\/em> = -2.270, <em>p<\/em> = .<strong>02*<\/strong>,<em> rs<\/em> = .15<\/td>\n<\/tr>\n<tr>\n<td style=\"vertical-align: middle;\">Sentence complexity<\/td>\n<td>C\/S<\/td>\n<td>115.52 (.295)<\/td>\n<td>101.48 (.283)<\/td>\n<td><em>z<\/em> = -1.652, <em>p<\/em> = .10,<em> rs<\/em> = .11<\/td>\n<\/tr>\n<tr>\n<td rowspan=\"4\" style=\"vertical-align: middle;\">Subordination<\/td>\n<td>C\/T<\/td>\n<td>117.60 (.268)<\/td>\n<td>99.40 (.261)<\/td>\n<td><em>z<\/em> = -2.139, <em>p<\/em> = <strong>.03*<\/strong>,<em> rs<\/em> = .15<\/td>\n<\/tr>\n<tr>\n<td>CT\/T<\/td>\n<td>.517 (.125)<\/td>\n<td>.478 (.114)<\/td>\n<td><em>t<\/em> = 2.377, <em>p<\/em> = <strong>.02*<\/strong>, <em>d<\/em> = .33<\/td>\n<\/tr>\n<tr>\n<td>DC\/C<\/td>\n<td>.404 (.077)<\/td>\n<td>.381 (.077)<\/td>\n<td><em>t<\/em> = 2.190, <em>p<\/em> = <strong>.03*<\/strong>, <em>d<\/em> = .30<\/td>\n<\/tr>\n<tr>\n<td>DC\/T<\/td>\n<td>118.25 (.236)<\/td>\n<td>98.75 (.228)<\/td>\n<td><em>z<\/em> = -2.292, <em>p<\/em> = <strong>.02*<\/strong>,<em> rs<\/em> = .16<\/td>\n<\/tr>\n<tr>\n<td rowspan=\"3\" style=\"vertical-align: middle;\">Coordination<\/td>\n<td>CP\/C<\/td>\n<td>102.29 (.143)<\/td>\n<td>114.71 (.119)<\/td>\n<td><em>z<\/em> = -1.460, <em>p<\/em> = .14,<em> rs<\/em> = .10<\/td>\n<\/tr>\n<tr>\n<td>CP\/T<\/td>\n<td>105.09 (.227)<\/td>\n<td>111.91 (.185)<\/td>\n<td><em>z<\/em> = -.802, <em>p<\/em> = .28,<em> rs<\/em> = .05<\/td>\n<\/tr>\n<tr>\n<td>T\/S<\/td>\n<td>103.90 (.077)<\/td>\n<td>113.10 (.085)<\/td>\n<td><em>z<\/em> = -1.082, <em>p<\/em> = .28,<em> rs<\/em> = .07<\/td>\n<\/tr>\n<tr>\n<td rowspan=\"3\" style=\"vertical-align: middle;\">Particular structures<\/td>\n<td>CN\/C<\/td>\n<td>109.11 (.319)<\/td>\n<td>107.89 (.352)<\/td>\n<td><em>z<\/em> = -.144, <em>p<\/em> = .89,<em> rs<\/em> = .01<\/td>\n<\/tr>\n<tr>\n<td>CN\/T<\/td>\n<td>116.01 (.540)<\/td>\n<td>100.99 (.598)<\/td>\n<td><em>z<\/em> = -1.766, <em>p<\/em> = .08,<em> rs<\/em> = .12<\/td>\n<\/tr>\n<tr>\n<td>VP\/T<\/td>\n<td>119.71 (.399)<\/td>\n<td>97.29 (.367)<\/td>\n<td><em>z<\/em> = -2.637, <em>p<\/em> = <strong>.01*<\/strong>,<em> rs<\/em> = .18<\/td>\n<\/tr>\n<\/tbody>\n<\/table>\n<p><em>* Significant difference at 0.05 level (2-tailed)<\/em><\/p>\n<p>Table 11 shows the comparisons of clausal and phrasal complexity between the two corpora. In terms of clausal complexity, the IRT exhibited greater number of direct objects per clause than the IRPrT with the small effect size (<em>d<\/em> = 0.30). In terms of phrasal complexity, the IRPrT exhibited more dependents per nominal subject than the IRT with the small effect size (<em>rs<\/em> = 0.19). However, the IRT exhibited more relative clause modifiers per nominal phrases than the IRPrT with the small effect size (<em>rs<\/em> = 0.13).<\/p>\n<p><strong>Table 11. Comparisons of measures of clausal and phrasal complexity<\/strong><\/p>\n<table style=\"margin-bottom: 2px;\">\n<tbody>\n<tr>\n<td rowspan=\"2\" style=\"vertical-align: middle;\"><strong>Indices<\/strong><\/td>\n<td colspan=\"2\" style=\"text-align: center;\"><strong><em>M<\/em><\/strong><strong> (<em>SD<\/em>)<\/strong><\/td>\n<td rowspan=\"2\" style=\"vertical-align: middle;\"><strong>Significance Testing<\/strong><\/td>\n<\/tr>\n<tr>\n<td style=\"text-align: center;\"><strong>IRT<\/strong><\/td>\n<td style=\"text-align: center;\"><strong>IRPrT<\/strong><\/td>\n<\/tr>\n<tr>\n<td colspan=\"4\" style=\"text-align: center;\"><strong>Clausal complexity<\/strong><\/td>\n<\/tr>\n<tr>\n<td>acomp<\/td>\n<td>104.60 (.030)<\/td>\n<td>112.40 (.037)<\/td>\n<td><em>z<\/em> = -.918, <em>p<\/em> = .36,<em> rs<\/em> = .06<\/td>\n<\/tr>\n<tr>\n<td>advcl<\/td>\n<td>106.35 (.024)<\/td>\n<td>110.65 (.028)<\/td>\n<td><em>z<\/em> = -.505, <em>p<\/em> = .61,<em> rs<\/em> = .03<\/td>\n<\/tr>\n<tr>\n<td>advmod<\/td>\n<td>.214 (.052)<\/td>\n<td>.210 (.058)<\/td>\n<td><em>t<\/em> = .477, <em>p<\/em> = .63, <em>d<\/em> = .08<\/td>\n<\/tr>\n<tr>\n<td>ccomp<\/td>\n<td>115.91 (.050)<\/td>\n<td>101.09 (.047)<\/td>\n<td><em>z<\/em> = -1.742, <em>p<\/em> = .08,<em> rs<\/em> = .12<\/td>\n<\/tr>\n<tr>\n<td>cc<\/td>\n<td>107.50 (.014)<\/td>\n<td>109.50 (.018)<\/td>\n<td><em>z<\/em> = -.236, <em>p<\/em> = .81,<em> rs<\/em> = .02<\/td>\n<\/tr>\n<tr>\n<td>conj<\/td>\n<td>106.20 (.040)<\/td>\n<td>110.80 (.037)<\/td>\n<td><em>z<\/em> = -.541, <em>p<\/em> = .59,<em> rs<\/em> = .04<\/td>\n<\/tr>\n<tr>\n<td>mark<\/td>\n<td>100.38 (.045)<\/td>\n<td>116.63 (.045)<\/td>\n<td><em>z<\/em> = -1.911, <em>p<\/em> = .06,<em> rs<\/em> = .13<\/td>\n<\/tr>\n<tr>\n<td>pcomp<\/td>\n<td>105.00 (.002)<\/td>\n<td>112.00 (.004)<\/td>\n<td><em>z<\/em> = -1.765, <em>p<\/em> = .08,<em> rs<\/em> = .12<\/td>\n<\/tr>\n<tr>\n<td>csubj<\/td>\n<td>107.82 (.011)<\/td>\n<td>109.18 (.009)<\/td>\n<td><em>z<\/em> = -.169, <em>p<\/em> = .87,<em> rs<\/em> = .01<\/td>\n<\/tr>\n<tr>\n<td>xsubj<\/td>\n<td>N\/A<\/td>\n<td>N\/A<\/td>\n<td>N\/A<\/td>\n<\/tr>\n<tr>\n<td>nsubj<\/td>\n<td>.609 (.089)<\/td>\n<td>.602 (.082)<\/td>\n<td><em>t<\/em> = .565, <em>p<\/em> = .57, <em>d<\/em> = .08<\/td>\n<\/tr>\n<tr>\n<td>dobj<\/td>\n<td>.400 (.072)<\/td>\n<td>.378 (.075)<\/td>\n<td><em>t<\/em> = 2.225, <em>p<\/em> = <strong>.03*<\/strong>, <em>d<\/em> = .30<\/td>\n<\/tr>\n<tr>\n<td>iobj<\/td>\n<td>105.92 (.005)<\/td>\n<td>111.08 (.006)<\/td>\n<td><em>z<\/em> = -.784, <em>p<\/em> = .43,<em> rs<\/em> = .05<\/td>\n<\/tr>\n<tr>\n<td>ncomp<\/td>\n<td>103.03 (.031)<\/td>\n<td>113.97 (.035)<\/td>\n<td><em>z<\/em> = -1.286, <em>p<\/em> = .20,<em> rs<\/em> = .09<\/td>\n<\/tr>\n<tr>\n<td>xcomp<\/td>\n<td>114.79 (.032)<\/td>\n<td>102.21 (.035)<\/td>\n<td><em>z<\/em> = -1.479, <em>p<\/em> = .14,<em> rs<\/em> = .10<\/td>\n<\/tr>\n<tr>\n<td colspan=\"4\" style=\"text-align: center;\"><strong>Phrasal complexity<\/strong><\/td>\n<\/tr>\n<tr>\n<td>nsubj_deps<\/td>\n<td>96.86 (.200)<\/td>\n<td>120.14 (.244)<\/td>\n<td><em>z<\/em> = -2.738, <em>p<\/em> = <strong>.01*<\/strong>,<em> rs<\/em> = .19<\/td>\n<\/tr>\n<tr>\n<td>ncomp_deps<\/td>\n<td>110.17 (.828)<\/td>\n<td>106.83 (.803)<\/td>\n<td><em>z<\/em> = -.393, <em>p<\/em> = .69,<em> rs<\/em> = .03<\/td>\n<\/tr>\n<tr>\n<td>dobj_deps<\/td>\n<td>107.25 (.235)<\/td>\n<td>109.75 (.240)<\/td>\n<td><em>z<\/em> = -.294, <em>p<\/em> = .77,<em> rs<\/em> = .02<\/td>\n<\/tr>\n<tr>\n<td>iobj_deps<\/td>\n<td>107.37 (.218)<\/td>\n<td>109.63 (.418)<\/td>\n<td><em>z<\/em> = -.585, <em>p<\/em> = .56,<em> rs<\/em> = .04<\/td>\n<\/tr>\n<tr>\n<td>pobj_deps<\/td>\n<td>1.343 (.126)<\/td>\n<td>1.313 (.125)<\/td>\n<td><em>t<\/em> = 1.735, <em>p<\/em> = .08, <em>d<\/em> = .24<\/td>\n<\/tr>\n<tr>\n<td>det_nominal<\/td>\n<td>.312 (.064)<\/td>\n<td>.315 (.053)<\/td>\n<td><em>t<\/em> = -.301, <em>p<\/em> = .76, <em>d<\/em> = .05<\/td>\n<\/tr>\n<tr>\n<td>amod_nominal<\/td>\n<td>103.14 (.050)<\/td>\n<td>113.86 (.058)<\/td>\n<td><em>z<\/em> = -1.261, <em>p<\/em> = .21,<em> rs<\/em> = .09<\/td>\n<\/tr>\n<tr>\n<td>prep_nominal<\/td>\n<td>.220 (.044)<\/td>\n<td>.218 (.049)<\/td>\n<td><em>t<\/em> = .335, <em>p<\/em> = .74, <em>d<\/em> = .05<\/td>\n<\/tr>\n<tr>\n<td>poss_nominal<\/td>\n<td>115.39 (.023)<\/td>\n<td>101.61 (.025)<\/td>\n<td><em>z<\/em> = -1.621, <em>p<\/em> = .11,<em> rs<\/em> = .11<\/td>\n<\/tr>\n<tr>\n<td>vmod_nominal<\/td>\n<td>106.17 (.014)<\/td>\n<td>110.83 (.014)<\/td>\n<td><em>z<\/em> = -.549, <em>p<\/em> = .58,<em> rs<\/em> = .04<\/td>\n<\/tr>\n<tr>\n<td>nn_nominal<\/td>\n<td>106.81 (.056)<\/td>\n<td>110.19 (.055)<\/td>\n<td><em>z<\/em> = -.398, <em>p<\/em> = .69,<em> rs<\/em> = .03<\/td>\n<\/tr>\n<tr>\n<td>rcmod_nominal<\/td>\n<td>116.75 (.014)<\/td>\n<td>100.25 (.014)<\/td>\n<td><em>z<\/em> = -1.941, <em>p<\/em> = <strong>.05*<\/strong>,<em> rs<\/em> = .13<\/td>\n<\/tr>\n<tr>\n<td>advmod_nominal<\/td>\n<td>103.84 (.013)<\/td>\n<td>113.16 (.121)<\/td>\n<td><em>z<\/em> = -1.096, <em>p<\/em> = .27,<em> rs<\/em> = .07<\/td>\n<\/tr>\n<tr>\n<td>conj_and_nominal<\/td>\n<td>101.63 (.025)<\/td>\n<td>115.37 (.026)<\/td>\n<td><em>z<\/em> = -1.615, <em>p<\/em> = .11,<em> rs<\/em> = .11<\/td>\n<\/tr>\n<tr>\n<td>conj_or_nominal<\/td>\n<td>103.37 (.008)<\/td>\n<td>113.63 (.010)<\/td>\n<td><em>z<\/em> = -1.221, <em>p<\/em> = .22,<em> rs<\/em> = .08<\/td>\n<\/tr>\n<\/tbody>\n<\/table>\n<p><em>* Significant difference at 0.05 level (2-tailed)<\/em><\/p>\n<h4>Discourse Text Complexity<\/h4>\n<p>Table 12 shows the comparisons of the discourse level indices of the IRT and IRPrT. The IRT had higher lexical overlap than IRPrT in all indices at both sentence and paragraph levels with the large effect sizes. This indicates that the IRT had more repetition of words, including nouns and pronouns in subsequent sentences than the IRPrT. The IRT also had much more semantic overlap than the IRPrT, as all of the indices exhibited significant differences with the large effect sizes. This means that the IRT had much more semantic similarity between all adjacent sentences, and paragraphs than the IRPrT.<\/p>\n<p>Finally, the results also show that the IRPrT employed more basic connectives, demonstratives, additives and logical connectors than IRT with the large effect sizes. Furthermore, the IRPrT had much higher givenness indices than the IRT as indicated by the large effect sizes. This suggests that content words and third person pronouns were repeated far more in the IRPrT than in the IRT.<\/p>\n<p><strong>Table 12. Comparisons of measures of discourse text complexity<\/strong><\/p>\n<table style=\"margin-bottom: 2px;\">\n<tbody>\n<tr>\n<td rowspan=\"2\" style=\"vertical-align: middle;\"><strong>Indices<\/strong><\/td>\n<td colspan=\"2\" style=\"text-align: center;\"><strong><em>M<\/em><\/strong><strong> (<em>SD<\/em>)<\/strong><\/td>\n<td rowspan=\"2\" style=\"vertical-align: middle;\"><strong>Significance Testing<\/strong><\/td>\n<\/tr>\n<tr>\n<td style=\"text-align: center;\"><strong>IRT<\/strong><\/td>\n<td style=\"text-align: center;\"><strong>IRPrT<\/strong><\/td>\n<\/tr>\n<tr>\n<td colspan=\"4\" style=\"text-align: center;\"><strong>Lexical Overlap<\/strong><\/td>\n<\/tr>\n<tr>\n<td>2_argument_sent<\/td>\n<td>161.98 (.041)<\/td>\n<td>55.02 (.028)<\/td>\n<td><em>z<\/em> = -12.557, <em>p<\/em> = <strong>.00<\/strong>*,<em> rs<\/em> = .85<\/td>\n<\/tr>\n<tr>\n<td>binary_all_sent<\/td>\n<td>162.50 (.053)<\/td>\n<td>54.50 (.058)<\/td>\n<td><em>z<\/em> = -12.705, <em>p<\/em> = <strong>.00*<\/strong>,<em> rs<\/em> = .86<\/td>\n<\/tr>\n<tr>\n<td>binary_argument_sent<\/td>\n<td>162.48 (.131)<\/td>\n<td>54.52 (.047)<\/td>\n<td><em>z<\/em> = -12.695, <em>p<\/em> = <strong>.00*<\/strong>,<em> rs<\/em> = .86<\/td>\n<\/tr>\n<tr>\n<td>2_all_para<\/td>\n<td>161.63 (.058)<\/td>\n<td>55.38 (.018)<\/td>\n<td><em>z<\/em> = -12.508, <em>p<\/em> = <strong>.00*<\/strong>,<em> rs<\/em> = .85<\/td>\n<\/tr>\n<tr>\n<td>2_argument_para<\/td>\n<td>159.81 (.063)<\/td>\n<td>57.19 (.037)<\/td>\n<td><em>z<\/em> = -12.065, <em>p<\/em> = <strong>.00*<\/strong>,<em> rs<\/em> = .82<\/td>\n<\/tr>\n<tr>\n<td>binary_all_para<\/td>\n<td>162.50 (.065)<\/td>\n<td>54.50 (.045)<\/td>\n<td><em>z<\/em> = -13.120, <em>p<\/em> = <strong>.00*<\/strong>,<em> rs<\/em> = .89<\/td>\n<\/tr>\n<tr>\n<td>binary_argument_para<\/td>\n<td>162.50 (.135)<\/td>\n<td>54.50 (.004)<\/td>\n<td><em>z<\/em> = -13.399, <em>p<\/em> = <strong>.00*<\/strong>,<em> rs<\/em> = .91<\/td>\n<\/tr>\n<tr>\n<td colspan=\"4\" style=\"text-align: center;\"><strong>Semantic overlap<\/strong><\/td>\n<\/tr>\n<tr>\n<td>lsa_1_all_sent<\/td>\n<td>162.50 (.071)<\/td>\n<td>54.50 (.009)<\/td>\n<td><em>z<\/em> = -12.780, <em>p<\/em> = <strong>.00*<\/strong>,<em> rs<\/em> = .87<\/td>\n<\/tr>\n<tr>\n<td>lsa_2_all_sent<\/td>\n<td>162.50 (.047)<\/td>\n<td>54.50 (.000)<\/td>\n<td><em>z<\/em> = -13.575, <em>p<\/em> = <strong>.00*<\/strong>,<em> rs<\/em> = .92<\/td>\n<\/tr>\n<tr>\n<td>lsa_1_all_para<\/td>\n<td>146.13 (.082)<\/td>\n<td>70.88 (.119)<\/td>\n<td><em>z<\/em> = -8.848, <em>p<\/em> = <strong>.00*<\/strong>,<em> rs<\/em> = .60<\/td>\n<\/tr>\n<tr>\n<td>lsa_2_all_para<\/td>\n<td>160.81 (.109)<\/td>\n<td>56.19 (.075)<\/td>\n<td><em>z<\/em> = -12.301, <em>p<\/em> = <strong>.00*<\/strong>,<em> rs<\/em> = .84<\/td>\n<\/tr>\n<tr>\n<td colspan=\"4\" style=\"text-align: center;\"><strong>Connectives<\/strong><\/td>\n<\/tr>\n<tr>\n<td>basic_connectives<\/td>\n<td>54.56 (.006)<\/td>\n<td>162.44 (.008)<\/td>\n<td><em>z<\/em> = -12.938, <em>p<\/em> = <strong>.00*<\/strong>,<em> rs<\/em> = .88<\/td>\n<\/tr>\n<tr>\n<td>all_demonstratives<\/td>\n<td>63.66 (.007)<\/td>\n<td>153.34 (.035)<\/td>\n<td><em>z<\/em> = -10.546, <em>p<\/em> = <strong>.00*<\/strong>,<em> rs<\/em> = .72<\/td>\n<\/tr>\n<tr>\n<td>all_additive<\/td>\n<td>64.95 (.010)<\/td>\n<td>152.05 (.035)<\/td>\n<td><em>z<\/em> = -10.242, <em>p<\/em> = <strong>.00*<\/strong>,<em> rs<\/em> = .70<\/td>\n<\/tr>\n<tr>\n<td>all_logical<\/td>\n<td>54.50 (.009)<\/td>\n<td>162.50 (.244)<\/td>\n<td><em>z<\/em> = -12.701, <em>p<\/em> = <strong>.00*<\/strong>,<em> rs<\/em> = .86<\/td>\n<\/tr>\n<tr>\n<td colspan=\"4\" style=\"text-align: center;\"><strong>Givenness<\/strong><\/td>\n<\/tr>\n<tr>\n<td>repeated_content_lemma<\/td>\n<td>56.50 (.039)<\/td>\n<td>160.50 (.803)<\/td>\n<td><em>z<\/em> = -12.230, <em>p<\/em> = <strong>.00*<\/strong>,<em> rs<\/em> = .83<\/td>\n<\/tr>\n<tr>\n<td>repeated_content_and_pronoun_lemma<\/td>\n<td>54.50 (.039)<\/td>\n<td>162.50 (.240)<\/td>\n<td><em>z<\/em> = -12.699, <em>p<\/em> = <strong>.00*<\/strong>,<em> rs<\/em> = .86<\/td>\n<\/tr>\n<\/tbody>\n<\/table>\n<p><em>* Significant difference at 0.05 level (2-tailed)<\/em><\/p>\n<h3>Discussion<\/h3>\n<h4>Lexical Text Complexity<\/h4>\n<p>The findings indicate that the IRT and IRPrT did not significantly differ in lexical sophistication. Specifically, the IRT and IRPrT had relatively equal text coverage of the most frequent 3,000 and 5,000 words; around 94% of texts in both being covered at the 3,000-word level, and approximately 97% being covered at the 5,000-word level. This suggests that the vocabulary loads in the IRT and IRPrT are quite similar. Moreover, no significant difference was shown in word frequency, word range, and contextual distinctiveness between IRT and IRPrT. This means that lexical items, either content or function words, are encountered a similar number of times across similar contexts in both corpora. This suggests that the IRPrT reflect the IRT well in terms of vocabulary size; therefore, if learners develop vocabulary knowledge through the IRPrT, they are likely to encounter words of the same frequency levels in the IRT.<\/p>\n<p>The findings also indicate that academic language was used at relatively equal frequency in the IRT and IRPrT. Therefore, the IRPrT are also likely to prepare learners for academic language that will appear on the IRT. However, it should be noted that the AWL coverage in both corpora was far lower than previous studies of academic written English texts (Chen &amp; Ge, 2007; Vongpumivitch et al., 2009). This raises the question as to whether the IRT is truly a valid measure of the academic vocabulary knowledge learners will need in real-life language domains later. However, future studies would be warranted to know this for sure.<\/p>\n<p>We also found no difference in the lexical sophistication of the IRT and IRPrT. Both word recognition norms (i.e., the time amount required to identify words correctly and to name them) and age of exposure (i.e., no words in any of the corpora required a more sophisticated link to other lexical items of relevant meaning for understanding the texts) measures were comparable across the corpora. This suggests a similar balance in cognitive processing of vocabulary when reading texts from either.<\/p>\n<p>Results for lexical density showed no difference between the IRT and IRPrT. Both IRT and IRPrT were comprised of almost 76% content words. This implies that the IRPrT expose learners to equal cognitive processing levels for semantically-rich words as the real test conditions, meaning that this sort of test preparation content is aligned quite well with the tests. Furthermore, the IRT and IRPrT had relatively equal density of content and function word tokens (approximately 50%). This further suggests that the IRPrT are providing a representative mix of content and function words, semantically tied together for comprehensibility.<\/p>\n<p>The IRT and IRPrT did not significantly differ in lexical diversity indices of TTR, MATTR, MTLD, and MTLD-MA. Both corpora contained around 43% different words overall, 66% different content words, and 20% different function words. This indicates less repetition of content words than function words in both. The high lexical diversity level of content words in the IRPrT may help test takers when taking the IRT, as there is a similar amount of variety in the real tests.<\/p>\n<p>Based on these results, the lexical alignment has positive implications for learners as follows:<\/p>\n<ul>\n<li>First, IRPrT can expose learners to texts with a similar informational load and mix of content and grammatical elements as found in IRT. This means that these materials provide truthful test practice at the word level and learners are likely developing relevant lexical knowledge through IRPrT.<\/li>\n<li>Second, as learners are aware that the words they learn in practice conditions can be of help in test conditions, vocabulary knowledge is more likely to be acquired and transferable to decoding meanings on IRT. This therefore implies positive washback &#8211; the practice tests motivate development of vocabulary knowledge construct-relevant to the actual test (Green, 2007). Seeing their word level improve on IRPrT may reinforce students\u2019 confidence to tackle the lexical complexity of IRT.<\/li>\n<\/ul>\n<p>Additionally, the lexical alignment can bear useful implications for instructors and IRPrT material developers. As for instructors, they can be further assured that assigning IRPrT materials can aid learners\u2019 expansion of lexicons in high-frequency and academic tiers, thereby being reasonably assured of parity with distributions in high-stakes IRT. Similarly, equipped with quantitative evidence that IRPrT exhibit comparable lexical sophistication, materials developers can utilize said evidence to validate the appropriateness of lexical profile in future IRPrT based on IRT benchmarks. If future IRPrT align with the lexical profile documented here, positive washback transpires as the future IRPrT may equip learners\u2019 lexicons to meet the demands of IRT, hence reinforcing score validity.<\/p>\n<h4>Syntactic Text Complexity<\/h4>\n<p>Clear differences were found in the syntactic complexity of the IRT and IRPrT corpora. Specifically, there might be a small difference in the length of syntactic units between the tests, as indicated by the significant difference found in MLT. However, since there were no differences in MLC and MLS, and there are differences in how older versions of TAASSC and Lu\u2019s (2010) original L2SCA calculate T-units, the difference found in MLT is somewhat inconclusive. These results might be affected by the fact that we used an older version of TAASSC which calculates T-units based on dependency tags, which is different from the tregex method used by Lu (2010). More importantly, subordination, a syntactic device that contributes to the depth and intricacy of sentences, was found to be used more frequently in the IRT than in the IRPrT across all measures of subordination (C\/T, CT\/T, DC\/C and DC\/T). For instance, the IRT exhibited a higher density of clauses per T-unit than the IRPrT. This pattern was echoed across the other subordination measures, i.e., the number of complex T-units divided by T-units as well as the numbers of dependent clauses per clause and T-unit. These findings indicate that the test uses far more of these complex structures than the practice materials, which is argued to be significant because potential test takers are likely to be under a more stressful cognitive load to decode the syntactic packaging as compared with practice conditions. For this reason, their reading comprehension might be impeded in the real-time processing of test conditions, as aligned with the previous studies (e.g., Kyle, 2016; Mesmer et al., 2012) where more syntactic complexity contributes to greater difficulty in reading comprehension. Moreover, because of no significant difference in lexical text complexity between the IRT and IRPrT as discussed earlier, greater syntactic complexity in the IRT may cause difficulty for those test takers having limited competence to negotiate with syntactic relations of the IRT in actual test conditions no matter how well-prepared they are for lexical knowledge in the IRPrT during test preparation. This interpretation corroborates Shiotsu and Weir\u2019s (2017) and Tong et al.\u2019s (2024) studies where syntactic knowledge and skills has more effect on lexical knowledge in text processing and comprehension.<\/p>\n<p>While our analyses included several indices of clausal and phrasal complexity, only a few indices mentioned above showed significant differences between the IRT and IRPrT. At the clausal level, a higher number of direct objects per clause in the IRT also means a more intricate description of the subject\u2019s actions for learners to interpret in the test. At the phrasal level, greater relative clause modifiers in the IRT can be mentally taxing to test takers in unpacking the meaning of sentences in the test condition. Although these isolated findings suggest caution in generalizing differences in clausal and phrasal complexity between the IRT and IRPrT, they may show that clausal and phrasal complexity is somewhat underrepresented in the IRPrT. Greater clausal embeddings to phrases and more direct objects per clauses can impose heavy cognitive loads on readers as well as impede processing time and comprehension. This is considered to be problematic, because of the limited capacity of one\u2019s working memory. \u201c[U]nder normal conditions, information can be remembered in working memory for about two seconds only. After that brief span, the representation is rapidly forgotten, unless it can be rehearsed subvocally\u201d (Ortega, 2009, p. 90). Based on this premise, if test takers encounter texts with greater clausal and phrasal complexity under the time constraint of the test condition than in the practice conditions, they may be under cognitive pressure to commit the information they read to the site of consciousness (Baars &amp; Franklin, 2003).<\/p>\n<p>However, the mean for dependents per nominal subject is higher in the IRPrT than the IRT. Given that a sentence with numerous dependents may be considered redundant (Crossley et al., 2017), we would argue that this difference could be problematic in terms of syntactic complexity. As redundant information could lack conciseness and hinder the effective communication of ideas, this may unnecessarily overload learners with a false impression about what to expect in the IRT and may arouse anxiety as a form of negative washback during test preparation (Nguyen, 2023).<\/p>\n<p>One pedagogical implication of these results is that classroom instructors can leverage these findings to better prepare learners for the sophisticated structures of IRT under strict time constraints. Making students aware of the syntactic gap could motivate them to seek additional practice to regulate learning and expand readiness for syntactic complexity in IRT. Furthermore, supplementary materials or instruction focused on comprehending longer T-units and clauses with multiple embedded elements may be warranted. Through targeted practice on the enhanced materials, students can develop strategies to unpack meaning from sophisticated syntax they are likely to encounter on IRT and dispel misconceptions about upcoming IRT. Such a coordination between instruction and further material use guided by empirical syntactic data serves to reduce negative washback during test preparation as the threat of construct-irrelevant variance to test scores. At the same time, future material developers of IRPrT could consider calibrating the syntax of IRPrT more closely to the sophistication of IRT.<\/p>\n<h4>Discourse Text Complexity<\/h4>\n<p>We found that the IRT had higher overlap than the IRPrT at both sentence and paragraph levels. This finding suggests that the IRPrT are not as lexically or semantically cohesive as the IRT, which might make them more difficult to comprehend (Gernsbacher, 2013). We would argue that having to read more challenging texts without lexical cohesion from test preparation materials at hand, learners may exhibit disengagement and reduced confidence when preparing for the IELTS.<\/p>\n<p>The semantic overlap results indicated that there was greater semantic similarity in the IRT than in the IRPrT between all adjacent sentences (both at one-sentence and two-sentence intervals) and between all adjacent paragraphs (both at one-paragraph and two-paragraph intervals). Therefore, it can be inferred that there was less similarity of ideas in the IRPrT than the IRT. Given that local (sentence) cohesion of ideas can facilitate moment-to-moment understanding of a text and that global (text) cohesion can be conducive to the overall integration and recall of ideas as the text progresses (Koda, 2005), the lack of local and global cohesion in the IRPrT can be linked to difficulty in reading comprehension during test preparation, as suggested by Ehrlich (1991).<\/p>\n<p>There are several potential reasons for the greater lexical and semantic overlap in the IRT than the IRPrT. One is that trade-off between cohesion and syntactic sophistication increased text cohesion but led to greater text length, density, and complexity (Beck et al., 1991; Ozuru et al., 2009). This could potentially also explain why MLT seemed to be slightly longer in IRT than in IRPrT. Another reason could be due to the fact that the IRPrT had more basic connectives, demonstratives, additives and logical connectors than IRT. Although having more givenness and connectives would support learners in linking ideas of a reading text easily to a certain extent during test preparation, we would argue that caution should be exercised in generalizing these findings to be a complete advantage for learners. Without a blend of semantic similarity at different discourse distances in IRPrT, learners may struggle to temporally commit mental representations of text information to their working memory in line with the text development. The disjointed flow of meaningful ideas during reading may lead them to lose focus on a text at hand and score low on reading comprehension items during test preparation, so IRPrT materials may consider reducing these connectors in favor of repetition in the future.<\/p>\n<p>Based on these findings, some implications can be drawn for classroom instructor, material developers and learners as follows:<\/p>\n<ul>\n<li>Classroom instructors should take greater care for selecting preparatory materials that cultivate the cohesive flow found in IRPrT, not just the usage of connectives. They should also be attentive that relying excessively on transitional phrases in lesson texts may not be conducive to the inherent lexical and semantic ties. Classroom exercises and readings should also provide exposure to and emphasize the importance of lexical overlap, phrasal repetition, and semantic connectivity threaded throughout texts, modeling the cohesion patterns on IRT.<\/li>\n<li>Materials developers should rigorously analyze the discourse-level cohesion of IRPrT by developing the passages exhibiting phrasal repetition and semantic similarity between sentences and paragraphs on par with IRT. Conscious efforts made to mirror the authentic cohesive flow of the IRT itself, not merely with focus on connective usage, will better prepare learners for the linguistic demands of IRT.<\/li>\n<li>Learners should not fall to the misconception that just having more connecting words means having better discourse-level cohesion. When engaging with IRPrT, learners should consciously observe how coherent flow arises in passages from lexical and phrasal repetition and tight semantic ties, not solely transitional words.<\/li>\n<\/ul>\n<p>This study also bears broader implications beyond IELTS for language testing researchers by offering valuable operationalization for investigating other high-stakes ELPT used for admission to academic contexts. A parallel multidimensional complexity audit comparing texts from such widely-used ELPT as TOEFL, IELTS, and PTE academic in relation to their corresponding practice tests would enable detailed comparability in linguistic features. Such empirical profiling illuminates if the text complexity features appropriately align across these standardized tests themselves as well as between the real tests and their preparatory materials. Given that English language testing increasingly plays a critical gatekeeping role for high-stakes decisions, an understanding of comparative text complexity features across the key ELPT can strengthen claims of score equivalence, and validity of score interpretations. Another implication pertains to profiling innovative informal assessments, such as the online Duolingo English Test, which have claimed increased acceptance for tertiary admissions. A systematic analysis of Duolingo\u2019s linguistic complexity in comparison with other traditional ELPT would determine areas of alignment and divergence across lexical, grammatical and discourse features. This provides a nuanced perspective on the relationships emerging between established standards and \u201cinformal\u201d online formats that promise greater access yet require empirical profiling of construct coverage, difficulty, cognitive processing and comparability to traditional benchmarks. In totality, by undertaking corpus-based multidimensional text complexity analysis, this study offers research-driven implications beyond merely IELTS preparation to examining the intricate relationships between traditional and technology-mediated ELPT, buttressing consequential decisions.<\/p>\n<p>While this study offers valuable insights, certain limitations should be noted. First, the analysis solely employed computational tools to quantify text complexity differences. The subjective experience of learner readers was not incorporated. Perception studies measuring learners\u2019 comprehension, engagement, and strategies in IRT versus IRPrT could complement the corpus-based computational comparisons. To further illuminate the comparability of text complexity between IRT and IRPrT, future studies could compare test-takers\u2019 reading performance on passages drawn from the two corpora. Simulated recall procedures made immediately after reading could further shed light on how the readers navigated and interpreted linguistic complexity variations between IRPrT and IRT. Second, the links between text complexity measures and actual reading performance remain unexplored. Relating the text complexity measures to test-takers\u2019 scores at different proficiency levels could better explain the implications. Despite such limitations, the study is one of the few corpus-based investigations to compare IRPrT with IRT in terms of text complexity and provide evidence on IRPrT usefulness in preparation for IRT.<\/p>\n<h3>Conclusions<\/h3>\n<p>The results of this study show that the IRT and IRPrT are genreally quite comparable in terms of objective measures of lexical text complexity, indicating that the IRPrT can provide practice that is likely to be lexically similar to IRT and prepare test takers for the real IELTS. This is important as it assures educators and learners that the IRPrT are providing authentic lexical practice. However, the results also indicate that the IRT might have slightly longer syntactic units and that it definitely uses much more subordination than IRPrT. Therefore, test takers using the IRPrT might want to be aware of this difference and supplement their learning with some other materials that help with reading sentences with lots of suboordination. Finally, we found that the IRT has a larger amount of cohesion, specifically semantic repetition, than the IRPrT texts, which instead favor connectors. Therefore, educators and learners using these materials might want to practice reading texts with more repetition, and IRPrT materials developers may want to consider increasing semantic repetition in future practice texts that they create.<\/p>\n<h3>About the Authors<\/h3>\n<p><strong>Huu Thanh Minh Nguyen <\/strong>is a lecturer in English as a Foreign Language at University of Foreign Language Studies, The University of Danang. He earned his Master\u2019s degree in TESOL Studies from the University of Leeds, UK. He is doing doctoral research at School of Languages and Linguistics, The University of Melbourne, Australia. His main research interests include Language Testing and Assessment, Second Language Writing, and Corpus Linguistics. ORCID ID: 0000-0002-9494-9641.<\/p>\n<p><strong>Nguyen Van Anh Le<\/strong> is a lecturer in English as a Foreign Language at University of Foreign Language Studies, The University of Danang. She earned her Master&#8217;s degree from the University of Huddersfield, UK, and her Ph.D. from Sophia University, Tokyo, Japan. Her main research interests include Phonetics, Phonology, Second Language Acquisition, and Sociolinguistics. ORCID ID: 0009-0000-7378-1432<\/p>\n<h3>Acknowledgements and Funding<\/h3>\n<p>We owe our immense thanks to the editorial board and four anonymous reviewers whose insightful and critical reviews have contributed to critical improvements in the earlier versions of the manuscript. We would also want to extend our deepest gratitude to Mr. Ngo Pham Anh Huy whose coding support was an indispensable part of our work.<\/p>\n<p>Research for this article was funded by University of Foreign Language Studies, The University of Danang (Project number: T2023-05-15) and Funds for Science and Technology Development, The University of Danang.<\/p>\n<h3 id=\"note\">Note<\/h3>\n<p>[1] T-unit is known as one main clause plus any subordinate clause or non-clausal structure to which it is attached or within which it is embedded (Vajjala &amp; Meurers, 2013). [<a href=\"#back1\">back<\/a>]<\/p>\n<p>[2] A lemma is a \u201cdictionary headword; an abstract representation, subsuming all the formal lexical variations which may apply: the verb <em>walk<\/em>, for example, subsumes <em>walking<\/em>, <em>walks<\/em> and <em>walked<\/em>\u201d (Knowles &amp; Mohd Don, 2004, p. 70, original emphasis). [<a href=\"#back2\">back<\/a>]<\/p>\n<p>[3] Givenness indices are reflected of the approximated proportion of given information to new information (Crossley et al., 2016). [<a href=\"#back3\">back<\/a>]<\/p>\n<h3>To Cite this Article<\/h3>\n<p>Nguyen, H. T. M., &#038; Le, N. V. A. (2024). Text Complexity of Cambridge-delivered IELTS Academic Reading Tests: Comparability with IELTS Academic Reading Practice Tests from Other Publishers. <em>Teaching English as a Second Language Electronic Journal (TESL-EJ), 28<\/em>(2). https:\/\/doi.org\/10.55593\/ej.28110a4<\/p>\n<h3>References<\/h3>\n<p>Alderson, J. C. (2000). Assessing reading. Cambridge University Press.<\/p>\n<p>Baars, B. J., &amp; Franklin, S. (2003) How conscious experience and working memory interact. <em>TRENDS in Cognitive Sciences<\/em>, <em>7<\/em>(4), 166\u2013172. <a href=\"https:\/\/doi.org\/10.1016\/S1364-6613(03)00056-1\" target=\"_blank\" rel=\"noopener\">https:\/\/doi.org\/10.1016\/S1364-6613(03)00056-1<\/a><\/p>\n<p>Bachman, L. F., Davidson, F., &amp; Milanovic, M. (1996). The use of test method characteristics in the content analysis and design of EFL proficiency tests. <em>Language Testing, 13<\/em>(2), 125\u2013150. <a href=\"https:\/\/doi.org\/10.1177\/02655322960130020\" target=\"_blank\" rel=\"noopener\">https:\/\/doi.org\/10.1177\/02655322960130020<\/a><\/p>\n<p>Bardovi-Harlig, K. (2000). <em>Tense and aspect in second language acquisition: Form, meaning and use<\/em>. Blackwell.<\/p>\n<p>Beck, I. L., McKeown, M. G., Sinatra, G. M., &amp; Loxterman, J. A. (1991). Revising social studies text from a text-processing perspective: evidence of improved comprehensibility. <em>Reading Research Quarterly<\/em>, <em>26<\/em>(3), 251\u2013276. <a href=\"https:\/\/psycnet.apa.org\/doi\/10.2307\/747763\" target=\"_blank\" rel=\"noopener\">https:\/\/doi.org\/10.2307\/747763<\/a><\/p>\n<p>Biber, D., Gray, B., &amp; Poonpon, K. (2013). Pay attention to the phrasal structures: Going beyond t\u2010units-a response to Weiwei Yang. <em>TESOL Quarterly<\/em>, <em>47<\/em>(1), 192\u2013201. <a href=\"https:\/\/doi.org\/10.1002\/tesq.84\" target=\"_blank\" rel=\"noopener\">https:\/\/doi.org\/10.1002\/tesq.84<\/a><\/p>\n<p>Biber, D., Gray, B., &amp; Staples, S. (2014). Predicting patterns of grammatical complexity across language exam task types and proficiency levels. <em>Applied Linguistics<\/em>, <em>37<\/em>(5), 639\u2013668. <a href=\"https:\/\/doi.org\/10.1093\/applin\/amu059\" target=\"_blank\" rel=\"noopener\">https:\/\/doi.org\/10.1093\/applin\/amu059<\/a><\/p>\n<p>Brown, J. D. (1998). An EFL readability index. <em>JALT Journal<\/em>, <em>20<\/em>, 7\u201336.<\/p>\n<p>Briscoe, T., Medlock, B., &amp; Andersen, \u00d8. E. (2010). <em>Automated assessment of ESOL free text examinations<\/em>. University of Cambridge: Computer Laboratory. <a href=\"https:\/\/www.cl.cam.ac.uk\/techreports\/UCAM-CL-TR-790.pdf\" target=\"_blank\" rel=\"noopener\">https:\/\/www.cl.cam.ac.uk\/techreports\/UCAM-CL-TR-790.pdf<\/a><\/p>\n<p>Bult\u00e9, B., &amp; Housen, A. (2012). Defining and operationalising L2 complexity. In A. House, F. Kuiken, &amp; I. Vedder (Eds.), <em>Dimensions of L2 performance and proficiency: Complexity, accuracy and fluency in SLA<\/em> (pp. 21\u201346). John Benjamins.<\/p>\n<p>Carrell, P. L. (1987). Readability in ESL. <em>Reading in a Foreign Language<\/em>, <em>4<\/em>(1), 21\u201340.<\/p>\n<p>Chen, Q. &amp; Ge, G. (2007). A corpus-based lexical study on frequency and distribution of Coxhead\u2019s AWL word families in medical research articles (RAs). <em>English for Specific Purposes<\/em>, <em>26<\/em>, 502\u2013504. <a href=\"https:\/\/doi.org\/10.1016\/j.esp.2007.04.003\" target=\"_blank\" rel=\"noopener\">https:\/\/doi.org\/10.1016\/j.esp.2007.04.003<\/a><\/p>\n<p>Cobb, T. (2009). <em>The Compleat lexical tutor<\/em>. <a href=\"http:\/\/www.lextutor.ca\/\" target=\"_blank\" rel=\"noopener\">http:\/\/www.lextutor.ca\/<\/a>.<\/p>\n<p>Crossley, S. A., Greenfield, J., &amp; McNamara, D. S. (2008). Assessing text readability using cognitively based indices. <em>TESOL Quarterly<\/em>, <em>42<\/em>(3), 475\u2013493. <a href=\"https:\/\/doi.org\/10.1002\/j.1545-7249.2008.tb00142.x\" target=\"_blank\" rel=\"noopener\">https:\/\/doi.org\/10.1002\/j.1545-7249.2008.tb00142.x<\/a><\/p>\n<p>Crossley, S. A., Kyle, K., &amp; McNamara, D. S. (2016). The tool for the automatic analysis of text cohesion (TAACO): Automatic assessment of local, global, and text cohesion. <em>Behavior Research Methods<\/em>, <em>48<\/em>(4), 1227\u20131237. <a href=\"https:\/\/doi.org\/10.3758\/s13428-015-0651-7\" target=\"_blank\" rel=\"noopener\">https:\/\/doi.org\/10.3758\/s13428-015-0651-7<\/a><\/p>\n<p>Crossley, S. A., Skalicky, S., Dascalu, M., McNamara, D. S., &amp; Kyle, K. (2017). Predicting Text Comprehension, Processing, and Familiarity in Adult Readers: New Approaches to Readability Formulas. <em>Discourse Processes<\/em>, <em>54<\/em>(5-6), 340\u2013359. <a href=\"https:\/\/doi.org\/10.1080\/0163853X.2017.1296264\" target=\"_blank\" rel=\"noopener\">https:\/\/doi.org\/10.1080\/0163853X.2017.1296264<\/a><\/p>\n<p>Ehrlich, M. F. (1991). The processing of cohesion devices in text comprehension. <em>Psychological Research, 53<\/em>(2), 169\u2013174. <a href=\"https:\/\/doi.org\/10.1007\/BF01371825\" target=\"_blank\" rel=\"noopener\">https:\/\/doi.org\/10.1007\/BF01371825<\/a><\/p>\n<p>Everett, R., Coleman, J. (2003, April 17). <em>A critical analysis of selected IELTS preparation materials. <\/em>IELTS. <a href=\"https:\/\/ielts.org\/researchers\/our-research\/research-reports\/a-critical-analysis-of-selected-ielts-preparation-materials\" target=\"_blank\" rel=\"noopener\">https:\/\/ielts.org\/researchers\/our-research\/research-reports\/a-critical-analysis-of-selected-ielts-preparation-materials<\/a><\/p>\n<p>Fang, Z., &amp; Pace, B. G. (2013). Teaching with challenging texts in the disciplines: Text complexity and close reading. <em>Journal of Adolescent &amp; Adult Literacy<\/em>, <em>57<\/em>(2), 104\u2013108. <a href=\"https:\/\/doi.org\/10.1002\/JAAL.229\" target=\"_blank\" rel=\"noopener\">https:\/\/doi.org\/10.1002\/JAAL.229<\/a><\/p>\n<p>Fulmer, S. M., D\u2019Mello, S. K., Strain, A., &amp; Graesser, A. C. (2015). Interest-based text preference moderates the effect of text difficulty on engagement and learning. <em>Contemporary Educational Psychology<\/em>, <em>41<\/em>, 98\u2013100. <a href=\"https:\/\/doi.org\/10.1016\/j.cedpsych.2014.12.005\" target=\"_blank\" rel=\"noopener\">https:\/\/doi.org\/10.1016\/j.cedpsych.2014.12.005<\/a>.<\/p>\n<p>Garnier, M. &amp; Schmitt, N. (2015). The PHaVE List: A pedagogical list of PVs and their most frequent meaning senses. <em>Language Teaching Research<\/em>, <em>19<\/em>(6), 645\u2013666. <a href=\"https:\/\/doi.org\/10.1177\/1362168814559798\" target=\"_blank\" rel=\"noopener\">https:\/\/doi.org\/10.1177\/1362168814559798<\/a><\/p>\n<p>Gernsbacher, M. A. (2013). <em>Language comprehension as structure building<\/em>. Psychology Press.<\/p>\n<p>Goldman, S. R., &amp; Rakestraw, J. A. (2000). Structural aspects of constructing meaning from text. In M. L. Kamil, P. B. Rosenthal, P. D. Pearson, &amp; R. Barr (Eds.), <em>Handbook of reading research<\/em>, (pp. 311\u2013335). Lawrence Erlbaum.<\/p>\n<p>Gough, C., &amp; Hutchison. S. (2015). <em>Exam Essential Practice Tests: IELTS 2<\/em>. Cengage Learning.<\/p>\n<p>Graesser, A. C., McNamara, D. S., Louwerse, M. M., &amp; Cai, Z. (2004). Coh-Metrix: Analysis of text on cohesion and language. <em>Behavior Research Methods, Instruments, &amp; Computers, 36<\/em>(2), 193\u2013202. <a href=\"http:\/\/doi.org\/10.3758\/BF03195564\" target=\"_blank\" rel=\"noopener\">http:\/\/doi.org\/10.3758\/BF03195564<\/a><\/p>\n<p>Green, A. (2006). Washback to the learner: Learner and teacher perspectives on IELTS preparation course expectations and outcomes, <em>Assessing Writing<\/em>, <em>11<\/em>(2), 113\u2013134. <a href=\"https:\/\/doi.org\/10.1016\/j.asw.2006.07.002\" target=\"_blank\" rel=\"noopener\">https:\/\/doi.org\/10.1016\/j.asw.2006.07.002<\/a><\/p>\n<p>Green, A. (2007). <em>IELTS washback in context: Preparation for academic writing in higher education<\/em>. Cambridge University Press.<\/p>\n<p>Greenfield, G. (1999). <em>Classic readability formulas in an EFL context: Are they valid for Japanese speakers?<\/em> Unpublished doctoral dissertation. Temple University.<\/p>\n<p>Guthrie, J. T., Klauda, S. L., &amp; Ho, A. N. (2013). Modeling the relationships among reading instruction, motivation, engagement, and achievement for adolescents. <em>Reading Research Quarterly<\/em>, <em>48<\/em>(1), 9\u201326. <a href=\"https:\/\/doi.org\/10.1002\/rrq.035\" target=\"_blank\" rel=\"noopener\">https:\/\/doi.org\/10.1002\/rrq.035<\/a><\/p>\n<p>Halliday, M. A. K. (2004). <em>An introduction to functional grammar<\/em>. Hodder Education.<\/p>\n<p>Halliday, M. A. K., &amp; Hasan, R. (2014). <em>Cohesion in English<\/em>. Routledge.<\/p>\n<p>Harrison, M., &amp; Whitehead, R. (2015). <em>Exam Essential Practice Tests: IELTS 1<\/em>. Cengage Learning.<\/p>\n<p>Hoeksema, J., &amp; Napoli, D. J. (1993). Paratactic and subordinative <em>so<\/em>. <em>Journal of Linguistics<\/em>, <em>29<\/em>(2), 291\u2013314. <a href=\"https:\/\/doi.org\/10.1017\/s0022226700000347\" target=\"_blank\" rel=\"noopener\">https:\/\/doi.org\/10.1017\/s0022226700000347<\/a><\/p>\n<p>Jakeman, V., &amp; McDowell, C. (2001). <em>IELTS Practice Tests Plus<\/em>. Pearson.<\/p>\n<p>Kirby, P. (2016). <em>Shadow schooling: Private tuition and social mobility in the UK<\/em>. The Sutton Trust.<\/p>\n<p>Kim, M., Crossley, S. A., &amp; Kyle, K. (2018). Lexical sophistication as a multidimensional phenomenon: Relations to second language lexical proficiency, development, and writing quality. <em>The Modern Language Journal<\/em>, <em>102<\/em>(1), 120\u2013141. <a href=\"https:\/\/doi.org\/10.1111\/modl.12447\" target=\"_blank\" rel=\"noopener\">https:\/\/doi.org\/10.1111\/modl.12447<\/a><\/p>\n<p>Knowles, G., &amp; Mohd Don, Z. (2004). The notion of a \u201clemma\u201d: Headwords, roots and lexical sets. <em>International Journal of Corpus Linguistics<\/em>, <em>9<\/em>(1), 69\u201381. <a href=\"https:\/\/doi.org\/10.1075\/ijcl.9.1.04kno\" target=\"_blank\" rel=\"noopener\">https:\/\/doi.org\/10.1075\/ijcl.9.1.04kno<\/a><\/p>\n<p>Koda, K. (2005). <em>Insights into second language reading: A cross-linguistic approach.<\/em> Cambridge University Press.<\/p>\n<p>Kunnan, A. J., &amp; Carr, N. T. (2017). A comparability study between the General English Proficiency Test-Advanced and the Internet-Based Test of English as a Foreign Language. <em>Language Testing in Asia, 7<\/em>(1), 7\u201317. <a href=\"https:\/\/doi.org\/10.1186\/s40468-017-0048-x\" target=\"_blank\" rel=\"noopener\">https:\/\/doi.org\/10.1186\/s40468-017-0048-x<\/a><\/p>\n<p>Kyle, K. (2016). <em>Measuring syntactic development in L2 writing: Fine grained indices of syntactic complexity and usage-based indices of syntactic sophistication. <\/em>Doctoral dissertation, Georgia State University. <a href=\"https:\/\/scholarworks.gsu.edu\/alesl_diss\/35\" target=\"_blank\" rel=\"noopener\">https:\/\/scholarworks.gsu.edu\/alesl_diss\/35<\/a><\/p>\n<p>Kyle, K., &amp; Crossley, S. A. (2015). Automatically assessing lexical sophistication: Indices, tools, findings, and application. <em>TESOL Quarterly<\/em>, <em>49<\/em>(4), 757\u2013786. <a href=\"https:\/\/doi.org\/10.1002\/tesq.194\" target=\"_blank\" rel=\"noopener\">https:\/\/doi.org\/10.1002\/tesq.194<\/a><\/p>\n<p>Kyle, K., Crossley, S. A., &amp; Jarvis, S. (2021). Assessing the validity of lexical diversity indices using direct judgements. <em>Language Assessment Quarterly<\/em>. Advance online publication. <a href=\"https:\/\/doi.org\/10.1080\/15434303.2020.1844205\" target=\"_blank\" rel=\"noopener\">https:\/\/doi.org\/10.1080\/15434303.2020.1844205<\/a><\/p>\n<p>Laufer, B., &amp; Nation, P. (1995). Vocabulary size and use: Lexical richness in L2 written production. <em>Applied Linguistics, 16<\/em>(3), 307\u2013322. <a href=\"https:\/\/doi.org\/10.1093\/applin\/16.3.307\" target=\"_blank\" rel=\"noopener\">https:\/\/doi.org\/10.1093\/applin\/16.3.307<\/a><\/p>\n<p>Laufer, B., &amp; Ravenhorst-Kalovski, G. C. (2010). Lexical threshold revisited: Lexical text coverage, learner\u2019s vocabulary size and reading comprehension. <em>Reading in a Foreign Language<\/em>, <em>22<\/em>(1), 15\u201330.<\/p>\n<p>Lu, X. (2010). Automatic analysis of syntactic complexity in second language writing. <em>International Journal of Corpus Linguistics,<\/em> <em>15<\/em>, 474\u2013496. <a href=\"https:\/\/doi.org\/10.1075\/ijcl.15.4.02lu\" target=\"_blank\" rel=\"noopener\">https:\/\/doi.org\/10.1075\/ijcl.15.4.02lu<\/a><\/p>\n<p>Matthews, M., &amp; Salisbury, K. (2011). <em>IELTS Practice Tests Plus 3<\/em>. Pearson.<\/p>\n<p>McCarter, S., &amp; Ash, J. (2008). <em>IELTS Test Builder 1<\/em>. Macmillan.<\/p>\n<p>McCarter, S. (2008). <em>IELTS Test Builder 2<\/em>. Macmillan.<\/p>\n<p>McEnery, T. (2006). A2. Representativeness, balance and sampling. In T. McEnery, R. Xiao, &amp; Y. Tono (Eds.), <em>Corpus-based language studies: <\/em><em>An<\/em><em> advanced resource book<\/em> (pp. 13\u201321). Routledge.<\/p>\n<p>Mesmer, H. A., Cunningham, J. W., &amp; Hiebert, E. H. (2012). Toward a theoretical model of text complexity for the early grades: Learning from the past, anticipating the future. Reading Research Quarterly, 47(3), 235\u2013258. <a href=\"https:\/\/doi.org\/10.1002\/rrq.019\" target=\"_blank\" rel=\"noopener\">https:\/\/doi.org\/10.1002\/rrq.019<\/a><\/p>\n<p>Messenger, K., Branigan, H. P., McLean, J. F., &amp; Sorace, A. (2012). Is young children\u2019s passive syntax semantically constrained? Evidence from syntactic priming. <em>Journal of Memory and Language<\/em>, <em>66<\/em>(4), 568-587. <a href=\"https:\/\/doi.org\/10.1016\/j.jml.2012.03.008\" target=\"_blank\" rel=\"noopener\">https:\/\/doi.org\/10.1016\/j.jml.2012.03.008<\/a><\/p>\n<p>Messick, S.\u00a0(1989).\u00a0Validity. In\u00a0R. L. Linn, (Ed.),\u00a0<em>Educational measurement<\/em>\u00a0(3rd ed.) (pp.\u00a013\u2013103). American Council on Education and Macmillan.<\/p>\n<p>Michel, M. (2017). Complexity, Accuracy and Fluency (CAF). In: L. Shawn, &amp; S. Masatoshi (Eds.), <em>The Routledge Handbook of Instructed Second Language Acquisition<\/em> (pp. 50\u201368). Routledge.<\/p>\n<p>Nation, I. P. (2006). How large a vocabulary is needed for reading and listening? <em>Canadian Modern Language Review<\/em>, <em>63<\/em>(1), 59\u201382. <a href=\"https:\/\/doi.org\/10.3138\/cmlr.63.1.59\" target=\"_blank\" rel=\"noopener\">https:\/\/doi.org\/10.3138\/cmlr.63.1.59<\/a><\/p>\n<p>Nguyen, H. T. M. (2023). The washback of the International English Language Testing System (IELTS) as an English language proficiency exit test on the learning of final-year English majors. <em>Teaching English as a Second Language Electronic Journal (TESL-EJ), 27<\/em>(2), 1\u201334. <a href=\"https:\/\/doi.org\/10.55593\/ej.27106a8\" target=\"_blank\" rel=\"noopener\">https:\/\/doi.org\/10.55593\/ej.27106a8<\/a><\/p>\n<p>Ortega, L. (2003). Syntactic complexity measures and their relationship to L2 proficiency: A research synthesis of college-level L2 writing. <em>Applied Linguistics<\/em>, <em>24<\/em>, 492\u2013518. <a href=\"http:\/\/doi.org\/10.1093\/applin\/24.4.492\" target=\"_blank\" rel=\"noopener\">http:\/\/doi.org\/10.1093\/applin\/24.4.492<\/a><\/p>\n<p>Ortega, L. (2009). <em>Understanding Second language acquisition<\/em>. Hodder Education.<\/p>\n<p>O\u2019Sullivan, B., Dunn, K., &amp; Berry, V. (2021). Test preparation: an international comparison of test takers\u2019 preferences. <em>Assessment in Education: Principles, Policy &amp; Practice<\/em>, <em>28<\/em>(1), 13\u201336. <a href=\"https:\/\/doi.org\/10.1080\/0969594X.2019.1637820\" target=\"_blank\" rel=\"noopener\">https:\/\/doi.org\/10.1080\/0969594X.2019.1637820<\/a><\/p>\n<p>Ozuru, Y., Dempsey, K., &amp; McNamara, D. S. (2009). Prior knowledge, reading skill, and text cohesion in the comprehension of science texts. <em>Learning and Instruction<\/em>, <em>19<\/em>(3), 228\u2013242. <a href=\"https:\/\/doi.org\/10.1016\/j.learninstruc.2008.04.003\" target=\"_blank\" rel=\"noopener\">https:\/\/doi.org\/10.1016\/j.learninstruc.2008.04.003<\/a><\/p>\n<p>Pearson, W. S. (2019). Critical perspectives on the IELTS test. <em>ELT Journal<\/em>, <em>73<\/em>(2), 197\u2013206. <a href=\"https:\/\/doi.org\/10.1093\/elt\/ccz006\" target=\"_blank\" rel=\"noopener\">https:\/\/doi.org\/10.1093\/elt\/ccz006<\/a><\/p>\n<p>Perfetti C., &amp; Stafura J. (2014). Word knowledge in a theory of reading comprehension.\u00a0<em>Scientific studies of Reading<\/em>, <em>18<\/em>(1), 22\u201337. <a href=\"https:\/\/doi.org\/10.1080\/10888438.2013.827687\" target=\"_blank\" rel=\"noopener\">https:\/\/doi.org\/10.1080\/10888438.2013.827687<\/a><\/p>\n<p>Plonsky, L., &amp; Oswald, F. L. (2014). How big is \u201cbig\u201d? Interpreting effect sizes in L2 research. <em>Language Learning<\/em>, <em>64<\/em>(4), 878\u2013912. <a href=\"https:\/\/doi.org\/10.1111\/lang.12079\" target=\"_blank\" rel=\"noopener\">https:\/\/doi.org\/10.1111\/lang.12079<\/a><\/p>\n<p>Read, J. (2000). <em>Assessing vocabulary<\/em>. Cambridge University Press.<\/p>\n<p>Sainani, K. L. (2012). Dealing with Non-normal Data. <em>PM&amp;R<\/em>, <em>4<\/em>(12), 1001\u20131005. <a href=\"https:\/\/doi.org\/10.1016\/j.pmrj.2012.10.013\" target=\"_blank\" rel=\"noopener\">https:\/\/doi.org\/10.1016\/j.pmrj.2012.10.013<\/a><\/p>\n<p>Shiotsu, T., &amp; Weir, C. J. (2007). The relative significance of syntactic knowledge and vocabulary breadth in the prediction of reading comprehension test performance. <em>Language Testing<\/em>, <em>24<\/em>(1), 99\u2013128. <a href=\"https:\/\/doi.org\/10.1177%2F0265532207071513\" target=\"_blank\" rel=\"noopener\">https:\/\/doi.org\/10.1177%2F0265532207071513<\/a><\/p>\n<p>Snow, C. (2002). <em>Reading for understanding: Toward an R&amp;D program in reading comprehension<\/em>. Rand Education.<\/p>\n<p>Spring, R. (2019). <em>From Linguistic Theory to the Classroom: A Practical Guide and Case Study<\/em>. Cambridge Scholars.<\/p>\n<p>Terry, M., &amp; Wilson, J. (2005). <em>IELTS Practice Tests Plus 2<\/em>. Pearson.<\/p>\n<p>Thompson, G. (2014). <em>Introducing functional grammar<\/em>. Routledge.<\/p>\n<p>Tong, X., Yu, L., &amp; Deacon, S. H. (2024). A Meta-Analysis of the Relation Between Syntactic Skills and Reading Comprehension: A Cross-Linguistic and Developmental Investigation.\u00a0<em>Review of Educational Research<\/em>.\u00a0<a href=\"https:\/\/doi.org\/10.3102\/00346543241228185\" target=\"_blank\" rel=\"noopener\">https:\/\/doi.org\/10.3102\/00346543241228185<\/a><\/p>\n<p>Vajjala, S., &amp; Meurers, D. (2013). On the applicability of readability models to web texts. In S. Williams, A. Siddharthan, &amp; A. Nenkova (Eds.), <em>Proceedings of the 2nd workshop on predicting and improving text readability for target reader populations<\/em> (pp. 59\u201368). Association for Computational Linguistics. <a href=\"https:\/\/aclanthology.org\/W13-2907.pdf\" target=\"_blank\" rel=\"noopener\">https:\/\/aclanthology.org\/W13-2907.pdf<\/a><\/p>\n<p>Vongpumivitch, V., Huang J., &amp; Chang Y. (2009). Frequency analysis of the words in the Academic Word List (AWL) and non-AWL content words in applied linguistics research papers. <em>English for Specific Purposes<\/em>, <em>28<\/em>, 33\u201341. <a href=\"https:\/\/doi.org\/10.1016\/j.esp.2007.04.003\" target=\"_blank\" rel=\"noopener\">https:\/\/doi.org\/10.1016\/j.esp.2007.04.003<\/a><\/p>\n<p>Yang, Y.-H., Chu, H.-C., &amp; Tseng, W.-T. (2021). Text Difficulty in Extensive Reading: Reading Comprehension and Reading Motivation. <em>Reading in a Foreign Language<\/em>, <em>33<\/em>(1), 78\u2013102.<\/p>\n<p>Yu, X. (2021). Text Complexity of Reading Comprehension Passages in the National Matriculation English Test in China: The Development from 1996 to 2020. <em>International Journal of Language Testing<\/em>, <em>11<\/em>(2), 142\u2013167.<\/p>\n<table border=\"0\" width=\"80%\" align=\"center\">\n<tbody>\n<tr>\n<td>Copyright of articles rests with the authors. Please cite TESL-EJ appropriately.<br \/>\n<strong>Editor\u2019s Note:<\/strong> The HTML version contains no page numbers. Please use the <a href=\"https:\/\/tesl-ej.org\/pdf\/ej110\/a4.pdf\">PDF version<\/a> of this article for citations.<\/td>\n<\/tr>\n<\/tbody>\n<\/table>\n","protected":false},"excerpt":{"rendered":"<p>August 2024 &#8211; Volume 28, Number 2 https:\/\/doi.org\/10.55593\/ej.28110a4 Huu Thanh Minh Nguyen University of Foreign Language Studies, The University of Danang &lt;nhtminhufl.udn.vn&gt; Nguyen Van Anh Le University of Foreign Language Studies, The University of Danang &lt;lnvanhufl.udn.vn&gt; Abstract Comparing language tests and test preparation materials holds important implications for the latter\u2019s validity and reliability. However, not [&hellip;]<\/p>\n","protected":false},"author":2,"featured_media":0,"parent":21535,"menu_order":4,"comment_status":"closed","ping_status":"closed","template":"","meta":{"_genesis_hide_title":false,"_genesis_hide_breadcrumbs":false,"_genesis_hide_singular_image":false,"_genesis_hide_footer_widgets":false,"_genesis_custom_body_class":"","_genesis_custom_post_class":"","_genesis_layout":"","footnotes":""},"class_list":["post-21603","page","type-page","status-publish","entry"],"featured_image_src":null,"featured_image_src_square":null,"_links":{"self":[{"href":"https:\/\/tesl-ej.org\/wordpress\/wp-json\/wp\/v2\/pages\/21603","targetHints":{"allow":["GET"]}}],"collection":[{"href":"https:\/\/tesl-ej.org\/wordpress\/wp-json\/wp\/v2\/pages"}],"about":[{"href":"https:\/\/tesl-ej.org\/wordpress\/wp-json\/wp\/v2\/types\/page"}],"author":[{"embeddable":true,"href":"https:\/\/tesl-ej.org\/wordpress\/wp-json\/wp\/v2\/users\/2"}],"replies":[{"embeddable":true,"href":"https:\/\/tesl-ej.org\/wordpress\/wp-json\/wp\/v2\/comments?post=21603"}],"version-history":[{"count":5,"href":"https:\/\/tesl-ej.org\/wordpress\/wp-json\/wp\/v2\/pages\/21603\/revisions"}],"predecessor-version":[{"id":21753,"href":"https:\/\/tesl-ej.org\/wordpress\/wp-json\/wp\/v2\/pages\/21603\/revisions\/21753"}],"up":[{"embeddable":true,"href":"https:\/\/tesl-ej.org\/wordpress\/wp-json\/wp\/v2\/pages\/21535"}],"wp:attachment":[{"href":"https:\/\/tesl-ej.org\/wordpress\/wp-json\/wp\/v2\/media?parent=21603"}],"curies":[{"name":"wp","href":"https:\/\/api.w.org\/{rel}","templated":true}]}}