[{"data":1,"prerenderedAt":-1},["ShallowReactive",2],{"doc-detail-209287-en":3,"doc-seo-209287-105":30,"detail-sidebar-cat-0-en-105":92},{"code":4,"msg":5,"data":6},0,"success",{"doc_id":7,"user_id":8,"nickname":9,"user_avatar":10,"doc_module":4,"category_id":11,"category_name":12,"doc_title":13,"doc_description":14,"doc_content":15,"file_id":16,"file_url":17,"file_type":18,"file_size":19,"view_count":20,"is_deleted":4,"is_public":20,"is_downloadable":20,"audit_status":20,"page_count":21,"language":22,"language_code":23,"site_id":24,"html_lang":23,"table_of_contents":25,"faqs":26,"seo_title":27,"seo_description":14,"update_tm":28,"read_time":29},209287,2336475401981,"Chumphorn","https://ap-avatar.wpscdn.com/avatar/22000c94efd8d5204d?x-image-process=image/resize,m_fixed,w_180,h_180&k=1786935347598174694",8,"Research & Report","Analyzing the Equity of the Brazilian National High School Exam - Validating the Item Response Theory’s Invariance","Several studies examine how economic, racial, and gender circumstances affect student performance in large-scale entrance exams such as Brazil’s ENEM. This work uses Item Response Theory to test the invariance premise, checking whether item characteristic curves are similar across subpopulations defined by gender, race, and income irrespective of students’ true abilities. It analyzes each group’s observed curve properties and applies a nonparametric ranking test to compare equity per item. Results show Languages and Codes questions favor male, white, and high-income participants, while Mathematics, Natural Sciences, and Human Sciences are more egalitarian.","Analyzing the Equity of the Brazilian National High School Exam by Validating the Item Response Theory’s Invariance  \nVitoria Guardieiro1 , Marcos M. Raimundo1 ,2 , Jorge Poco1  \n1 Getulio Vargas Foundation; 2 Universidade Federal do Rio de Janeiro [vitoriaguardieiro@gmail.com](vitoriaguardieiro@gmail.com) ; [marcosmrai@gmail.com](marcosmrai@gmail.com) ; [jorge.poco@fgv.br](jorge.poco@fgv.br)  \nABSTRACT  \nSeveral studies adopt different approaches to examining how economic, racial, and gender circumstances influence student performance in large-scale entrance exams, such as the National High School Exam (ENEM) . Using a methodology based on Item Response Theory, ENEM’s exam attempts to assess, for each item (question), the curve (function) that relates the participants’ abilities to their probabilities of correctly answering the item, which is assumed to hold whichever subgroup, a fundamental premise of IRT called invariance. This work analyzes whether the ENEM 2019 test presented similar curves for subpopulations defined by gender, race, and income, regardless of the participant’s actual abilities. Our approach is to analyze the properties of the observed curve for each group and then perform a nonparametric ranking test to compare the equity of each item (question) for each analyzed characteristic. We found that the ”Languages and Codes” questions consistently favored male, white, and high-income participants. At the sametime, the other three sets of questions (Mathematics, Natural Sciences, and Human Sciences) were considerably more egalitarian.  \nKeywords  \nHigher education entrance exams, Grading equity, Item Response Theory, Educational Data Mining  \n1. INTRODUCTION  \nThe Brazilian National High School Exam (ENEM, for its initials in Portuguese) is one of the most extensive entrance exams globally, having over 5 million participants registered in 2019 [11] . The exam has several functions; on an individual scale, it serves as an admission test to access the federal universities (through the Unified Selection System or SISU) and access to the federal scholarship programs (University for All Program or ProUni) . On a collective scale, this exam allows a comparison between schools and municipalities, and it also serves as an indicator for national public  \nV. Guardieiro, M. M. Raimundo, and J. Poco. Analyzing the equity of the Brazilian national high school exam by validating the item response theory’s invariance. In A. Mitrovic and N. Bosch, editors, Proceedings of the 15th International Conference on Educational Data Mining, pages 583–587, Durham, United Kingdom, July 2022 . International Educational Data Mining Society.  \n© 2022 Copyright is held by the author(s) . This work is distributed under the Creative Commons Attribution NonCommercial NoDerivatives 4.0 International (CC BY-NC-ND 4.0) license. [https://doi.org/10.5281/zenodo.6852946](https://doi.org/10.5281/zenodo.6852946)  \neducational policies at the national level [10] . Several studies investigated how sensible characteristics such as income, race, gender, and locality, affect the participants’ score [13, 20, 19] . However, most studies use the grade obtained asa direct indicator of the participants’ ability without investigating whether exams’ grading methodologies are unfairly favoring or disfavoring specific subpopulations.  \nSince 2009, ENEM’s participants’ grades have been assigned using Item Response Theory (IRT) methods, which consider the difficulty of each participant’s correct questions [7] to assign the grades, in contrast to the Classical Test Theory, where only the number of correct answers matter [4] . IRT creates a probability function that gives, for each question, the probability of a correct answer given the participant’s ability. Moreover, the primary assumption of IRT theory is that such function does not vary independently of subgroups of students [7, 16] . This assumption means that, given two groups based on a specific characteristic (e.g","cbCainr3Iz3iOPJb","https://ap.wps.com/l/cbCainr3Iz3iOPJb","pdf",1047645,1,5,"English","en",105,"# Abstract\n# Keywords\n# 1. Introduction\n# 2. Data","[{\"question\":\"What invariance assumption does the paper validate in ENEM grading?\",\"answer\":\"The study validates the Item Response Theory invariance premise, meaning item curve functions should not vary independently across subgroups. If subgroup-specific curves differ for many items, grading may systematically overestimate or underestimate groups’ abilities.\"},{\"question\":\"How does the methodology compare equity across gender, race, and income?\",\"answer\":\"The method approximates item characteristic curves from ENEM’s IRT-evaluated grades and answers, then compares curve-related properties per group. It uses a nonparametric ranking test to check whether observed item behavior is statistically consistent across the whole test.\"},{\"question\":\"Which subjects showed the most consistent bias according to the findings?\",\"answer\":\"The “Languages and Codes” questions consistently favored male, white, and high-income participants. The other three sets—Mathematics, Natural Sciences, and Human Sciences—were considerably more egalitarian.\"}]","Analyzing the Equity of the Brazilian National High School Exam - Validating the Item Response Theory’s Invariance | PDF",1788612168,13,{"code":4,"msg":31,"data":32},"ok",{"site_id":24,"language":23,"slug":33,"title":13,"keywords":34,"description":14,"schema_data":35,"social_meta":87,"head_meta":89,"extra_data":91,"updated_unix":28},"analyzing-the-equity-of-the-brazilian-national-high-school-exam-validating-the-item-response-theorys-invariance","",{"@graph":36,"@context":86},[37,54,69],{"@type":38,"itemListElement":39},"BreadcrumbList",[40,44,48,51],{"item":41,"name":42,"@type":43,"position":20},"https://docshare.wps.com","Home","ListItem",{"item":45,"name":46,"@type":43,"position":47},"https://docshare.wps.com/document/","Document",2,{"item":49,"name":12,"@type":43,"position":50},"https://docshare.wps.com/document/research-report/",3,{"item":52,"name":13,"@type":43,"position":53},"https://docshare.wps.com/document/analyzing-the-equity-of-the-brazilian-national-high-school-exam-validating-the-item-response-theorys-invariance/209287/",4,{"url":52,"name":13,"@type":55,"author":56,"headline":13,"publisher":58,"fileFormat":61,"inLanguage":23,"description":14,"dateModified":62,"datePublished":63,"encodingFormat":61,"isAccessibleForFree":64,"interactionStatistic":65},"DigitalDocument",{"name":9,"@type":57},"Person",{"url":41,"name":59,"@type":60},"DocShare","Organization","application/pdf","2026-09-11","2026-09-05",true,{"@type":66,"interactionType":67,"userInteractionCount":20},"InteractionCounter",{"@type":68},"ViewAction",{"@type":70,"mainEntity":71},"FAQPage",[72,78,82],{"name":73,"@type":74,"acceptedAnswer":75},"What invariance assumption does the paper validate in ENEM grading?","Question",{"text":76,"@type":77},"The study validates the Item Response Theory invariance premise, meaning item curve functions should not vary independently across subgroups. If subgroup-specific curves differ for many items, grading may systematically overestimate or underestimate groups’ abilities.","Answer",{"name":79,"@type":74,"acceptedAnswer":80},"How does the methodology compare equity across gender, race, and income?",{"text":81,"@type":77},"The method approximates item characteristic curves from ENEM’s IRT-evaluated grades and answers, then compares curve-related properties per group. It uses a nonparametric ranking test to check whether observed item behavior is statistically consistent across the whole test.",{"name":83,"@type":74,"acceptedAnswer":84},"Which subjects showed the most consistent bias according to the findings?",{"text":85,"@type":77},"The “Languages and Codes” questions consistently favored male, white, and high-income participants. The other three sets—Mathematics, Natural Sciences, and Human Sciences—were considerably more egalitarian.","https://schema.org",{"og:url":52,"og:type":88,"og:title":13,"og:site_name":59,"og:description":14},"article",{"robots":90,"canonical":52},"index,follow",{"doc_id":7,"site_id":24},{"code":4,"msg":5,"data":93},[94,98,102,106,110,115,120,123,128,131,135],{"id":20,"doc_module":4,"doc_module_name":46,"category_name":95,"show_sort_weight":96,"slug":97},"Story & Novel",90,"story-novel",{"id":47,"doc_module":4,"doc_module_name":46,"category_name":99,"show_sort_weight":100,"slug":101},"Literature",80,"literature",{"id":53,"doc_module":4,"doc_module_name":46,"category_name":103,"show_sort_weight":104,"slug":105},"Exam",70,"exam",{"id":21,"doc_module":4,"doc_module_name":46,"category_name":107,"show_sort_weight":108,"slug":109},"Comic",60,"comic",{"id":111,"doc_module":4,"doc_module_name":46,"category_name":112,"show_sort_weight":113,"slug":114},6,"Technology",50,"technology",{"id":116,"doc_module":4,"doc_module_name":46,"category_name":117,"show_sort_weight":118,"slug":119},7,"Healthcare",40,"healthcare",{"id":11,"doc_module":4,"doc_module_name":46,"category_name":12,"show_sort_weight":121,"slug":122},30,"research-report",{"id":124,"doc_module":4,"doc_module_name":46,"category_name":125,"show_sort_weight":126,"slug":127},9,"Religion & Spirituality",20,"religion-spirituality",{"id":126,"doc_module":4,"doc_module_name":46,"category_name":129,"show_sort_weight":126,"slug":130},"World Cup","world-cup",{"id":132,"doc_module":4,"doc_module_name":46,"category_name":133,"show_sort_weight":132,"slug":134},10,"Lifestyle","lifestyle",{"id":136,"doc_module":4,"doc_module_name":46,"category_name":137,"show_sort_weight":21,"slug":138},19,"General","general"]