[{"data":1,"prerenderedAt":-1},["ShallowReactive",2],{"doc-detail-113561-en":3,"doc-seo-113561-105":31,"detail-sidebar-cat-0-en-105":93},{"code":4,"msg":5,"data":6},0,"success",{"doc_id":7,"user_id":8,"nickname":9,"user_avatar":10,"doc_module":4,"category_id":11,"category_name":12,"doc_title":13,"doc_description":14,"doc_content":15,"file_id":16,"file_url":17,"file_type":18,"file_size":19,"view_count":20,"is_deleted":4,"is_public":21,"is_downloadable":21,"audit_status":21,"page_count":22,"language":23,"language_code":24,"site_id":25,"html_lang":24,"table_of_contents":26,"faqs":27,"seo_title":28,"seo_description":14,"update_tm":29,"read_time":30},113561,962075114101,"Seraphina","https://ap-avatar.wpscdn.com/avatar/e000253a75eb197efd?x-image-process=image/resize,m_fixed,w_180,h_180&k=1780044092746381165",8,"Research & Report","Learner Corpus Research in Applied Linguistics - Research overview","Learner corpora—collections of texts produced by second or additional language learners—are used in applied linguistics to support methodologically grounded analysis of learner language. The article reviews the brief history of learner corpus methodology, introduces key conceptual underpinnings, and presents a typology of learner corpora. It then summarizes two recent studies as models of best practice and extracts insights for understanding learner language development. The article concludes with research guidelines and future task directions for sustaining rigorous work and improving language pedagogy.","Learner Corpus Research in Applied Linguistics  \nUTE R€OMER-BARRON   \nDepartment of Applied Linguistics and ESL, Georgia State University  \n25 Park Place NE, Suite 1500, Atlanta, Georgia, 30303, USA  \nAbstract  \nThis article provides an overview of how learner corpora, deﬁned as collections of texts produced by second or additional language learners, have been used in applied linguistics. It starts by offering a brief history of the methodology of learner corpus analysis, introduces a typology of learner corpora, and outlines a few foundational concepts and analytic approaches in learner corpus research. To illustrate the impact that learner corpora have had on our ﬁeld, the article then summarizes two recent studies that model best practices in learner corpus research (Gries & Wulff, 2021; Wang, Wang & Wang, 2024) and the insights they have provided into learner language development. The article also offers guidelines for research based on learner corpora that can help us gain a better understanding of second language learning and positively impact language pedagogy. It closes with a discussion of future tasks for learner corpus researchers, focusing on required resources and research activities.  \ndoi: 10.1002/tesq.70140  \nBRIEF HISTORY, CONCEPTUAL UNDERPINNINGS, AND KEY CONTRIBUTIONS  \nlittle over 30 years ago, leading learner corpus researcher Sylviane  \nGrapusnogferLeparrovided anner Englishea(rIlyCLpEr)ogarendsscreallpeodrtthoenlethearneInternatr corpusioansala new source of data “a revolution in applied linguistics” (1994, p. 25) . In a TESOL Quarterly article published shortly after the release of ICLE’s ﬁrst edition (Granger, Dagneaux, & Meunier, 2002),  \nTESOL QUARTERLY Vol. 60, No. 2, June 2026 887  \n􀀁 2026 TESOL International Association.  \nGranger (2003) highlighted some of the beneﬁts of learner corpora, deﬁned as “electronic collection[s] of authentic texts produced by foreign or second language learners” (p. 538), and described their potential uses in second language learning and teaching as well as in second language acquisition (SLA) research. The present article aims to illustrate the relevance of learner corpus research for applied linguistics by providing a short history of the methodology and its foundational concepts, and by illustrating the impact that learner corpora have had in our ﬁeld. The article will also offer guidelines and a list of future tasks for rigorous learner corpus research.  \nThe ﬁrst learner corpora were compiled and made available in the late 1990s and early 2000s, ﬁrst to team members of the respective corpus project teams and then to researchers more broadly. Among them were ICLE, the Longman Learners’ Corpus, and several corpora that captured data from one L1-speciﬁc learner population, such as JEFLL, the Japanese English as a Foreign Language Learner corpus (Tono, 2000), and USE, the Uppsala Student English corpus of L1 Swedish learner language (Axelsson, 2000; see Pravec, 2002 for an overview of early learner corpora) . Early learner corpora were predominantly written rather than spoken and consisted of (mostly advanced) learners’ writing produced in high school or university contexts. Corpora of learner speech lagged behind their written counterparts by about a decade, mostly due to challenges in accessing and collecting learner oral productions in a systematic way, but are now more widely available and include collections of data produced by learners from various L1s such as LINDSEI, the Louvain International Database of Spoken English Interlanguage (Gilquin, De Cock, & Granger, 2010), and L1-speciﬁc learner data collections such as NICT JLE, the National Institute of Information and Communications Technology Japanese Learner English corpus (Izumi, Uchimoto, & Isahara, 2004) .  \nMode aside, learner corpora can be divided into three general types based on how the learner data has been collected: (1) synchronically from a single group of learners, (2) synchronically from multipl","cbCair8reSaQw42P","https://ap.wps.com/l/cbCair8reSaQw42P","pdf",187973,5,1,16,"English","en",105,"# Brief history of learner corpus analysis\n## Conceptual underpinning and key contributions\n## Typology of learner corpora\n## Guidelines and future tasks","[{\"question\":\"How does the article define learner corpora and why are they important for applied linguistics?\",\"answer\":\"Learner corpora are defined as collections of texts produced by second or additional language learners. The article emphasizes their value as a data source for analyzing learner language and supporting applied linguistics research and practice.\"},{\"question\":\"What learner corpus types does the article distinguish?\",\"answer\":\"The article groups learner corpora into three types based on collection timing and design: synchronically from a single learner group, synchronically from multiple groups (cross-sectional), and diachronically (longitudinal corpora).\"},{\"question\":\"What does the article use recent studies to illustrate?\",\"answer\":\"Two recent studies are summarized to model best practices in learner corpus research and to show how corpus-based analyses can yield insights into learner language development.\"}]","Learner Corpus Research in Applied Linguistics - Research overview | PDF",1784703358,40,{"code":4,"msg":32,"data":33},"ok",{"site_id":25,"language":24,"slug":34,"title":13,"keywords":35,"description":14,"schema_data":36,"social_meta":88,"head_meta":90,"extra_data":92,"updated_unix":29},"learner-corpus-research-in-applied-linguistics-research-overview","",{"@graph":37,"@context":87},[38,55,70],{"@type":39,"itemListElement":40},"BreadcrumbList",[41,45,49,52],{"item":42,"name":43,"@type":44,"position":21},"https://docshare.wps.com","Home","ListItem",{"item":46,"name":47,"@type":44,"position":48},"https://docshare.wps.com/document/","Document",2,{"item":50,"name":12,"@type":44,"position":51},"https://docshare.wps.com/document/research-report/",3,{"item":53,"name":13,"@type":44,"position":54},"https://docshare.wps.com/document/learner-corpus-research-in-applied-linguistics-research-overview/113561/",4,{"url":53,"name":13,"@type":56,"author":57,"headline":13,"publisher":59,"fileFormat":62,"inLanguage":24,"description":14,"dateModified":63,"datePublished":64,"encodingFormat":62,"isAccessibleForFree":65,"interactionStatistic":66},"DigitalDocument",{"name":9,"@type":58},"Person",{"url":42,"name":60,"@type":61},"DocShare","Organization","application/pdf","2026-07-29","2026-07-22",true,{"@type":67,"interactionType":68,"userInteractionCount":20},"InteractionCounter",{"@type":69},"ViewAction",{"@type":71,"mainEntity":72},"FAQPage",[73,79,83],{"name":74,"@type":75,"acceptedAnswer":76},"How does the article define learner corpora and why are they important for applied linguistics?","Question",{"text":77,"@type":78},"Learner corpora are defined as collections of texts produced by second or additional language learners. The article emphasizes their value as a data source for analyzing learner language and supporting applied linguistics research and practice.","Answer",{"name":80,"@type":75,"acceptedAnswer":81},"What learner corpus types does the article distinguish?",{"text":82,"@type":78},"The article groups learner corpora into three types based on collection timing and design: synchronically from a single learner group, synchronically from multiple groups (cross-sectional), and diachronically (longitudinal corpora).",{"name":84,"@type":75,"acceptedAnswer":85},"What does the article use recent studies to illustrate?",{"text":86,"@type":78},"Two recent studies are summarized to model best practices in learner corpus research and to show how corpus-based analyses can yield insights into learner language development.","https://schema.org",{"og:url":53,"og:type":89,"og:title":13,"og:site_name":60,"og:description":14},"article",{"robots":91,"canonical":53},"index,follow",{"doc_id":7,"site_id":25},{"code":4,"msg":5,"data":94},[95,99,103,107,111,116,120,123,128,131,135],{"id":21,"doc_module":4,"doc_module_name":47,"category_name":96,"show_sort_weight":97,"slug":98},"Story & Novel",90,"story-novel",{"id":48,"doc_module":4,"doc_module_name":47,"category_name":100,"show_sort_weight":101,"slug":102},"Literature",80,"literature",{"id":54,"doc_module":4,"doc_module_name":47,"category_name":104,"show_sort_weight":105,"slug":106},"Exam",70,"exam",{"id":20,"doc_module":4,"doc_module_name":47,"category_name":108,"show_sort_weight":109,"slug":110},"Comic",60,"comic",{"id":112,"doc_module":4,"doc_module_name":47,"category_name":113,"show_sort_weight":114,"slug":115},6,"Technology",50,"technology",{"id":117,"doc_module":4,"doc_module_name":47,"category_name":118,"show_sort_weight":30,"slug":119},7,"Healthcare","healthcare",{"id":11,"doc_module":4,"doc_module_name":47,"category_name":12,"show_sort_weight":121,"slug":122},30,"research-report",{"id":124,"doc_module":4,"doc_module_name":47,"category_name":125,"show_sort_weight":126,"slug":127},9,"Religion & Spirituality",20,"religion-spirituality",{"id":126,"doc_module":4,"doc_module_name":47,"category_name":129,"show_sort_weight":126,"slug":130},"World Cup","world-cup",{"id":132,"doc_module":4,"doc_module_name":47,"category_name":133,"show_sort_weight":132,"slug":134},10,"Lifestyle","lifestyle",{"id":136,"doc_module":4,"doc_module_name":47,"category_name":137,"show_sort_weight":20,"slug":138},19,"General","general"]