[{"data":1,"prerenderedAt":-1},["ShallowReactive",2],{"doc-detail-160528-en":3,"doc-seo-160528-105":31,"detail-sidebar-cat-0-en-105":92},{"code":4,"msg":5,"data":6},0,"success",{"doc_id":7,"user_id":8,"nickname":9,"user_avatar":10,"doc_module":4,"category_id":11,"category_name":12,"doc_title":13,"doc_description":14,"doc_content":15,"file_id":16,"file_url":17,"file_type":18,"file_size":19,"view_count":20,"is_deleted":4,"is_public":21,"is_downloadable":21,"audit_status":21,"page_count":22,"language":23,"language_code":24,"site_id":25,"html_lang":24,"table_of_contents":26,"faqs":27,"seo_title":28,"seo_description":14,"update_tm":29,"read_time":30},160528,962084931830,"Jacob","https://ap-avatar.wpscdn.com/davatar_a8503ba1806abce46bf441b54a3ca4cd",8,"Research & Report","Finding Time in Language Assessments - Maximizing Measurement Properties by Unit of Testing Time","Language assessment studies focus on how testing time shapes key usefulness criteria: validity, reliability, practicality, impact, authenticity, and interactivity. The work analyzes bandwidth and administration-length constraints across diverse testing contexts and explains how imposed time limits and task-aligned time constraints reduce random error while promoting real-world performance. It also examines computer-adaptive testing efficiency, simulation-based item-removal strategies, and timing effects on automaticity and scoring measurement. Practical guidance is provided for designing item types with appropriate maximum times per unit.","Finding time in language assessments: Maximizing  \nmeasurement properties by unit of testing time  \nSarah Goodwin, Geoffrey T. LaFlair, J.R. Lockwood, Steven W. Nydick, & Alina A. von Davier American Association for Applied Linguistics March 21, 2023  \n1  \nTest usefulness (Bachman & Palmer, 2010)  \n● Validity  \n● Reliability  \n● Practicality  \n● Impact  \n● Authenticity  \n● Interactivity  \nPracticality and impact affect access  \n Bandwidth considerations  \n Given in 213 countries and regions  \n Short administration time  \nTiming considerations  \nTo avoid fatiguing or overwhelming examinees, we need to know what effects timing has on test items, all while ensuring reliability and validity  \nHow critical is timing to a test construct?  \n● Imposed time limits, which are the norm on large-scale standardized tests, reduce sources of random error  \n● Time constraints associated with real-world tasks (speaking, writing) reflected as much as possible in assessment tasks  \n● Though response time does not affect scoring of performance, some degree of automaticity is expected for proficient test takers to complete test tasks  \n(Kane, 2020; Sireci & Botha, 2020; Wise & Kuhfeld, 2020)  \nComputer-adaptive test (CAT)s  \n● CATs are shorter than fixed-form tests but must have enough items to achieve a reliable final ability estimate  \n● Efficiency: the amount of measurement information per unit of time an item takes to administer (Cheng et al., 2017; Fan et al., 2012)  \n● With some exceptions (e.g., Crabtree, 2016; Hughes & Clesham, 2021; Kim et al., 2020, 2022), timing has sparingly been used to support test validity arguments  \n6  \nFinding Time: Steps we followed  \n● Test development/Literature review of factors affecting timing  \n○ Discussion of how to proceed  \n● Simulation of existing test taker results recalculating what would happen if items removed  \n○ Discussion  \n● Items removed to make enough time for a new item type  \n8  \nItem type Measurement  \nopportunities  \nMaximum time / item  \n\n| Conversation and Comprehension | 9 per item set | 1:30 |\n| --- | --- | --- |\n| Literacy and Comprehension | 18 per item set | 1:00 |\n| Conversation, Comprehension, Production | [depends on scoring features] | 0:20 |\n| Literacy, Comprehension, Production | 10 to 18 per item set | 3:00 |\n\nLiteracy, Conversation, Production  \n[depends on scoring  \n1:00  \nFINDING TIME  \nAffecting timing: Test delivery and format  \n● Whether a recording can be heard multiple times  \n● At what rate a voice plays in listening input  \n● Length, e.g., single image versus a video, or the number of characters or words of input text  \n● Types of interaction  \nbetween/among participants and items  \n● Degree of speededness, e.g., whether candidates are meant to process a 100-word passage in the same time as, or different timing allotted to, a 200-word passage  \n● User interface/navigation, e.g., whether candidates must scroll or all information is viewable on a single screen  \n● Local Internet bandwidth","cbCaiu6LGbAe08LG","https://ap.wps.com/l/cbCaiu6LGbAe08LG","pdf",1167162,4,1,20,"English","en",105,"# Test usefulness\n## Validity, Reliability, Practicality\n## Impact, Authenticity, Interactivity\n# Practicality and impact considerations\n## Bandwidth and administration time\n## Timing and test constructs\n# Computer-adaptive tests (CATs)\n## Efficiency per unit time\n## Timing in validity arguments\n# Finding time: steps we followed\n## Literature review and development\n## Simulation and item removal\n# Item-type measurement timing\n## Maximum time per item set\n# Affecting timing: test delivery and format\n## Recording access and playback rate\n## Interaction, speededness, UI/navigation\n## Local internet bandwidth","[{\"question\":\"How does timing influence validity and reliability in language assessments?\",\"answer\":\"Imposed time limits can reduce sources of random error, supporting test validity arguments. Timing constraints tied to real-world tasks help reflect the intended construct, while reliability depends on maintaining enough information within the time budget.\"},{\"question\":\"What role do computer-adaptive tests (CATs) play in managing testing time?\",\"answer\":\"CATs shorten test length compared with fixed-form formats, but they require enough items to produce a reliable final ability estimate. Efficiency is framed as measurement information per unit of administration time.\"},{\"question\":\"What steps were followed to determine appropriate timing and item selection?\",\"answer\":\"The process included test development and a literature review on factors affecting timing, followed by simulation of existing results with recalculated scenarios when items were removed. Items were then removed to create sufficient time for introducing new item types.\"}]","Finding Time in Language Assessments - Maximizing Measurement Properties by Unit of Testing Time | PDF",1788068509,50,{"code":4,"msg":32,"data":33},"ok",{"site_id":25,"language":24,"slug":34,"title":13,"keywords":35,"description":14,"schema_data":36,"social_meta":87,"head_meta":89,"extra_data":91,"updated_unix":29},"finding-time-in-language-assessments-maximizing-measurement-properties-by-unit-of-testing-time","",{"@graph":37,"@context":86},[38,54,69],{"@type":39,"itemListElement":40},"BreadcrumbList",[41,45,49,52],{"item":42,"name":43,"@type":44,"position":21},"https://docshare.wps.com","Home","ListItem",{"item":46,"name":47,"@type":44,"position":48},"https://docshare.wps.com/document/","Document",2,{"item":50,"name":12,"@type":44,"position":51},"https://docshare.wps.com/document/research-report/",3,{"item":53,"name":13,"@type":44,"position":20},"https://docshare.wps.com/document/finding-time-in-language-assessments-maximizing-measurement-properties-by-unit-of-testing-time/160528/",{"url":53,"name":13,"@type":55,"author":56,"headline":13,"publisher":58,"fileFormat":61,"inLanguage":24,"description":14,"dateModified":62,"datePublished":63,"encodingFormat":61,"isAccessibleForFree":64,"interactionStatistic":65},"DigitalDocument",{"name":9,"@type":57},"Person",{"url":42,"name":59,"@type":60},"DocShare","Organization","application/pdf","2026-09-04","2026-08-30",true,{"@type":66,"interactionType":67,"userInteractionCount":20},"InteractionCounter",{"@type":68},"ViewAction",{"@type":70,"mainEntity":71},"FAQPage",[72,78,82],{"name":73,"@type":74,"acceptedAnswer":75},"How does timing influence validity and reliability in language assessments?","Question",{"text":76,"@type":77},"Imposed time limits can reduce sources of random error, supporting test validity arguments. Timing constraints tied to real-world tasks help reflect the intended construct, while reliability depends on maintaining enough information within the time budget.","Answer",{"name":79,"@type":74,"acceptedAnswer":80},"What role do computer-adaptive tests (CATs) play in managing testing time?",{"text":81,"@type":77},"CATs shorten test length compared with fixed-form formats, but they require enough items to produce a reliable final ability estimate. Efficiency is framed as measurement information per unit of administration time.",{"name":83,"@type":74,"acceptedAnswer":84},"What steps were followed to determine appropriate timing and item selection?",{"text":85,"@type":77},"The process included test development and a literature review on factors affecting timing, followed by simulation of existing results with recalculated scenarios when items were removed. Items were then removed to create sufficient time for introducing new item types.","https://schema.org",{"og:url":53,"og:type":88,"og:title":13,"og:site_name":59,"og:description":14},"article",{"robots":90,"canonical":53},"index,follow",{"doc_id":7,"site_id":25},{"code":4,"msg":5,"data":93},[94,98,102,106,111,115,120,123,127,130,134],{"id":21,"doc_module":4,"doc_module_name":47,"category_name":95,"show_sort_weight":96,"slug":97},"Story & Novel",90,"story-novel",{"id":48,"doc_module":4,"doc_module_name":47,"category_name":99,"show_sort_weight":100,"slug":101},"Literature",80,"literature",{"id":20,"doc_module":4,"doc_module_name":47,"category_name":103,"show_sort_weight":104,"slug":105},"Exam",70,"exam",{"id":107,"doc_module":4,"doc_module_name":47,"category_name":108,"show_sort_weight":109,"slug":110},5,"Comic",60,"comic",{"id":112,"doc_module":4,"doc_module_name":47,"category_name":113,"show_sort_weight":30,"slug":114},6,"Technology","technology",{"id":116,"doc_module":4,"doc_module_name":47,"category_name":117,"show_sort_weight":118,"slug":119},7,"Healthcare",40,"healthcare",{"id":11,"doc_module":4,"doc_module_name":47,"category_name":12,"show_sort_weight":121,"slug":122},30,"research-report",{"id":124,"doc_module":4,"doc_module_name":47,"category_name":125,"show_sort_weight":22,"slug":126},9,"Religion & Spirituality","religion-spirituality",{"id":22,"doc_module":4,"doc_module_name":47,"category_name":128,"show_sort_weight":22,"slug":129},"World Cup","world-cup",{"id":131,"doc_module":4,"doc_module_name":47,"category_name":132,"show_sort_weight":131,"slug":133},10,"Lifestyle","lifestyle",{"id":135,"doc_module":4,"doc_module_name":47,"category_name":136,"show_sort_weight":107,"slug":137},19,"General","general"]