{"industry":{"id":"e1a1da3c-52a8-4e99-88ad-d08ef17d7c3f","slug":"hr-technology","label":"HR Technology","description":"ATS, HRIS, payroll, and workforce management"},"topic":{"slug":"candidate-screening-assessment","label":"Candidate Screening & Assessment","description":"Background Checks, Pre-Employment Testing, Reference Checks, Skills Assessment","schemaKind":null},"answer":{"id":"72492b35-0af1-4a64-9156-e4f70d3a5618","slug":"ai-resume-screening-software-actually-work","question":"Does AI resume screening software actually work, and is it worth it (bias, accuracy)?","answerMarkdown":"AI resume screening reliably reproduces the ranking it was trained to reproduce, and independent audits show it does so with measurable demographic disparities: a University of Washington audit of embedding models ranking 500 real resumes against 500 job descriptions found white-associated names preferred in 85.1 percent of cases, female-associated names in 11.1 percent, and Black male-associated names disadvantaged in up to 100 percent of comparisons [1]. Whether it is worth buying turns on what the model was trained to predict, because vendors generally let the client pick the outcome, so a tool can be accurate at copying prior screening while adding nothing to hiring quality [4]. Corrected meta-analytic estimates also cut the validity of most selection procedures by .10 to .20 and left the structured interview highest ranked [7]. In New York City the tool needs a bias audit from the past year, published impact ratios, and candidate notice 10 business days before use [8]. This is general information, not legal advice.","answerText":"AI resume screening reliably reproduces the ranking it was trained to reproduce, and independent audits show it does so with measurable demographic disparities: a University of Washington audit of embedding models ranking 500 real resumes against 500 job descriptions found white-associated names preferred in 85.1 percent of cases, female-associated names in 11.1 percent, and Black male-associated names disadvantaged in up to 100 percent of comparisons [1]. Whether it is worth buying turns on what the model was trained to predict, because vendors generally let the client pick the outcome, so a tool can be accurate at copying prior screening while adding nothing to hiring quality [4]. Corrected meta-analytic estimates also cut the validity of most selection procedures by .10 to .20 and left the structured interview highest ranked [7]. In New York City the tool needs a bias audit from the past year, published impact ratios, and candidate notice 10 business days before use [8]. This is general information, not legal advice.","answerHtml":"<p>AI resume screening reliably reproduces the ranking it was trained to reproduce, and independent audits show it does so with measurable demographic disparities: a University of Washington audit of embedding models ranking 500 real resumes against 500 job descriptions found white-associated names preferred in 85.1 percent of cases, female-associated names in 11.1 percent, and Black male-associated names disadvantaged in up to 100 percent of comparisons <a href=\"https://arxiv.org/abs/2407.20371\" class=\"citation-ref\" data-citation-index=\"1\" target=\"_blank\" rel=\"noreferrer\">[1]</a>. Whether it is worth buying turns on what the model was trained to predict, because vendors generally let the client pick the outcome, so a tool can be accurate at copying prior screening while adding nothing to hiring quality <a href=\"https://arxiv.org/pdf/1906.09208\" class=\"citation-ref\" data-citation-index=\"4\" target=\"_blank\" rel=\"noreferrer\">[4]</a>. Corrected meta-analytic estimates also cut the validity of most selection procedures by .10 to .20 and left the structured interview highest ranked <a href=\"https://pubmed.ncbi.nlm.nih.gov/34968080/\" class=\"citation-ref\" data-citation-index=\"7\" target=\"_blank\" rel=\"noreferrer\">[7]</a>. In New York City the tool needs a bias audit from the past year, published impact ratios, and candidate notice 10 business days before use <a href=\"https://rules.cityofnewyork.us/wp-content/uploads/2023/04/DCWP-NOA-for-Use-of-Automated-Employment-Decisionmaking-Tools-2.pdf\" class=\"citation-ref\" data-citation-index=\"8\" target=\"_blank\" rel=\"noreferrer\">[8]</a>. This is general information, not legal advice.</p>\n","summary":"Independent audits find language-model resume rankers favor some demographic groups on otherwise identical resumes. The deeper issue is what these tools are trained to predict, usually a past hiring decision rather than job performance, so a high accuracy score can mean only that the model agrees with last year's recruiters. AI screening holds up on high-volume roles with concrete requirements, used to rank rather than reject, with human review and monitored adverse impact.","publishedAt":"2026-07-21T21:20:44.609","verifiedAt":"2026-07-21T00:00:00","editorialStatus":"APPROVED","lastReviewedAt":"2026-07-21T00:00:00","nextReviewDueAt":"2026-10-21T00:00:00","templateVersion":"v2","aliases":["Does AI resume screening actually work?","Is AI resume screening biased?","How accurate is AI resume screening software?","Is AI resume screening worth the money?","Do AI resume screeners discriminate against candidates?","Do large language models rank resumes fairly?","Is automated resume screening legal?","AI resume screening bias and accuracy evidence","Should we buy AI resume screening software?","What are the risks of AI resume screening?","Does AI resume ranking predict job performance?","AI candidate screening bias audit requirements"],"confidenceScore":82,"confidenceLabel":"High","canonicalUrl":null},"contributor":{"id":"ec39deab-44fe-48d8-9029-fefe993ab85a","slug":"answer-stack","displayName":"AnswerStack","websiteUrl":null},"contributorOrganizationProfile":{"entityId":"ec39deab-44fe-48d8-9029-fefe993ab85a","legalName":null,"description":null,"websiteUrl":null,"imageUrl":null,"slogan":null,"subtitle":null,"facts":[],"coiNote":null,"foundingDate":null,"numberOfEmployeesText":null,"contactPoint":null,"address":null,"headquartersText":null,"organizationType":null},"contributorPerson":{"slug":"answerstack-editorial-team","displayName":"AnswerStack Editorial Team"},"sections":[{"id":"52acf5af-7af5-43e1-b751-e677240238d9","sectionKey":"what_research_shows","sectionType":"markdown_section","heading":"What does the independent research show about AI resume screening?","introMarkdown":"The strongest evidence comes from resume audit studies, and it points where human-hiring audits have long pointed: identical resumes rank differently depending on the name at the top. Kyra Wilson and Aylin Caliskan of the University of Washington ran more than 500 public resumes against 500 job descriptions across nine occupations using Massive Text Embedding models, the class behind semantic resume matching in commercial tools [1]. White-associated names were preferred in 85.1 percent of cases against 11.1 percent for female-associated names, and Black male-associated names were disadvantaged in up to 100 percent of comparisons [1].\n\nA follow-up experiment tested what happens with a person in the loop. Across 528 participants and 1,526 screening decisions, people shown no recommendations picked across racial groups at roughly equal rates, but when the model favored a group they picked from it up to 90 percent of the time, even when they called the recommendations poor quality [2].\n\nNewer models complicate the picture rather than closing it. A June 2026 paired-resume audit of fourteen language models, at 24,024 paired postings each, found the one 2023-vintage model reproduced a pro-white callback gap of 2.12 percentage points, while every model from 2024 onward showed no significant gap or a reversal of up to 3.01 points [3].\n\nThe case people cite most is older and different in kind. Amazon started building a system in 2014 to score applicants from one to five stars, trained on ten years of resumes it had received; because most came from men, it learned to penalize the word \"women's\" and the names of certain all-women colleges, and Amazon scrapped it in 2017 [6]. That came from supervised learning on hiring records, not from a language model.","introHtml":"<p>The strongest evidence comes from resume audit studies, and it points where human-hiring audits have long pointed: identical resumes rank differently depending on the name at the top. Kyra Wilson and Aylin Caliskan of the University of Washington ran more than 500 public resumes against 500 job descriptions across nine occupations using Massive Text Embedding models, the class behind semantic resume matching in commercial tools <a href=\"https://arxiv.org/abs/2407.20371\" class=\"citation-ref\" data-citation-index=\"1\" target=\"_blank\" rel=\"noreferrer\">[1]</a>. White-associated names were preferred in 85.1 percent of cases against 11.1 percent for female-associated names, and Black male-associated names were disadvantaged in up to 100 percent of comparisons <a href=\"https://arxiv.org/abs/2407.20371\" class=\"citation-ref\" data-citation-index=\"1\" target=\"_blank\" rel=\"noreferrer\">[1]</a>.</p>\n<p>A follow-up experiment tested what happens with a person in the loop. Across 528 participants and 1,526 screening decisions, people shown no recommendations picked across racial groups at roughly equal rates, but when the model favored a group they picked from it up to 90 percent of the time, even when they called the recommendations poor quality <a href=\"https://arxiv.org/html/2509.04404v2\" class=\"citation-ref\" data-citation-index=\"2\" target=\"_blank\" rel=\"noreferrer\">[2]</a>.</p>\n<p>Newer models complicate the picture rather than closing it. A June 2026 paired-resume audit of fourteen language models, at 24,024 paired postings each, found the one 2023-vintage model reproduced a pro-white callback gap of 2.12 percentage points, while every model from 2024 onward showed no significant gap or a reversal of up to 3.01 points <a href=\"https://arxiv.org/abs/2606.28978\" class=\"citation-ref\" data-citation-index=\"3\" target=\"_blank\" rel=\"noreferrer\">[3]</a>.</p>\n<p>The case people cite most is older and different in kind. Amazon started building a system in 2014 to score applicants from one to five stars, trained on ten years of resumes it had received; because most came from men, it learned to penalize the word &quot;women&#39;s&quot; and the names of certain all-women colleges, and Amazon scrapped it in 2017 <a href=\"https://www.technologyreview.com/2018/10/10/139858/amazon-ditched-ai-recruitment-software-because-it-was-biased-against-women/\" class=\"citation-ref\" data-citation-index=\"6\" target=\"_blank\" rel=\"noreferrer\">[6]</a>. That came from supervised learning on hiring records, not from a language model.</p>\n","outroMarkdown":null,"outroHtml":null,"contentJson":{},"configJson":{},"noteMarkdown":null,"noteHtml":null,"sortOrder":10},{"id":"97e19685-31c5-46c1-b285-2e181deb27e4","sectionKey":"screening_types_table","sectionType":"table_section","heading":"Which kind of resume screening are you actually buying?","introMarkdown":"Four technologies get sold under one label, and the bias and accuracy story differs for each. Most coverage treats them as one thing, which is why it rarely helps.","introHtml":"<p>Four technologies get sold under one label, and the bias and accuracy story differs for each. Most coverage treats them as one thing, which is why it rarely helps.</p>\n","outroMarkdown":null,"outroHtml":null,"contentJson":{"rows":[{"cells":["Keyword and boolean matching","Literal term presence, plus knockout rules on fields like years of experience","Requirements written into the query, and applicant vocabulary differences","Whether the query matched the text, not anything about the candidate"]},{"cells":["Statistical ranking trained on past outcomes","Scores applicants against a target the employer chose, learned from past records","The past decisions, plus proxies that survive removal of protected attributes [5]","Agreement with the target, usually a prior decision, not performance [4]"]},{"cells":["Language model or embedding evaluation of free text","Semantic similarity to a job description, or a model's written judgment","Associations carried in the pretrained model, including name-based ones [1][3]","Agreement with a rater or a similarity score, not a validated prediction"]},{"cells":["Structured parsing and field extraction","Reads a resume into database fields with no score attached","Extraction failures on unusual formats, not a scoring rule","Whether the fields came out right, which you can check by hand [8]"]}],"columns":["Technology","How it ranks or filters","Where bias enters","What accuracy means here"]},"configJson":{},"noteMarkdown":null,"noteHtml":null,"sortOrder":20},{"id":"595db958-f2ab-4efa-af37-0a0347b91fd6","sectionKey":"how_each_behaves","sectionType":"markdown_section","heading":"How does each screening method actually behave?","introMarkdown":"### Keyword and boolean matching\n\nKeyword filters do what the query says and nothing more, so their errors are a recruiter's errors made faster: a string requiring \"CPA\" drops the candidate who wrote \"Certified Public Accountant.\" Bias here is written into the requirements rather than learned from data, so it is the easiest kind to fix. Pull the live query on an open requisition and count how many applicants each rule removes.\n\n### Statistical ranking trained on past hiring outcomes\n\nThis category produced the Amazon failure, and it is where bias is structural rather than incidental. The model learns from records of who was previously advanced or rated well, so patterns in those decisions become patterns in the score, and stripping name, gender, and race from the inputs does not remove them because proxies survive in schools, employers, and employment gaps [5]. Upturn documented a screener that had learned the name \"Jared\" and high school lacrosse predicted success, a real statistical association with no causal link to the work [5].\n\n### Language model and embedding evaluation of free text\n\nThese tools carry associations from pretraining rather than from your hiring history, which is why disparities appear with no employer data involved: the Washington audit ran on off-the-shelf embedding models with no fine-tuning and the name gaps appeared anyway [1]. Those associations move between releases, so treat the model version and prompt as a fixed configuration and re-audit whenever either changes [3].\n\n### Structured parsing and field extraction\n\nParsing carries the lowest risk of the four because it produces no score and expresses no preference. New York City's rule draws the same line, excluding from its definition of a simplified output tools that only translate or transcribe existing text, such as converting a resume from a PDF [8]. The real risk is quiet data loss, since a mangled skills section becomes a thin record every downstream filter inherits.","introHtml":"<h3>Keyword and boolean matching</h3>\n<p>Keyword filters do what the query says and nothing more, so their errors are a recruiter&#39;s errors made faster: a string requiring &quot;CPA&quot; drops the candidate who wrote &quot;Certified Public Accountant.&quot; Bias here is written into the requirements rather than learned from data, so it is the easiest kind to fix. Pull the live query on an open requisition and count how many applicants each rule removes.</p>\n<h3>Statistical ranking trained on past hiring outcomes</h3>\n<p>This category produced the Amazon failure, and it is where bias is structural rather than incidental. The model learns from records of who was previously advanced or rated well, so patterns in those decisions become patterns in the score, and stripping name, gender, and race from the inputs does not remove them because proxies survive in schools, employers, and employment gaps <a href=\"https://www.upturn.org/work/help-wanted/\" class=\"citation-ref\" data-citation-index=\"5\" target=\"_blank\" rel=\"noreferrer\">[5]</a>. Upturn documented a screener that had learned the name &quot;Jared&quot; and high school lacrosse predicted success, a real statistical association with no causal link to the work <a href=\"https://www.upturn.org/work/help-wanted/\" class=\"citation-ref\" data-citation-index=\"5\" target=\"_blank\" rel=\"noreferrer\">[5]</a>.</p>\n<h3>Language model and embedding evaluation of free text</h3>\n<p>These tools carry associations from pretraining rather than from your hiring history, which is why disparities appear with no employer data involved: the Washington audit ran on off-the-shelf embedding models with no fine-tuning and the name gaps appeared anyway <a href=\"https://arxiv.org/abs/2407.20371\" class=\"citation-ref\" data-citation-index=\"1\" target=\"_blank\" rel=\"noreferrer\">[1]</a>. Those associations move between releases, so treat the model version and prompt as a fixed configuration and re-audit whenever either changes <a href=\"https://arxiv.org/abs/2606.28978\" class=\"citation-ref\" data-citation-index=\"3\" target=\"_blank\" rel=\"noreferrer\">[3]</a>.</p>\n<h3>Structured parsing and field extraction</h3>\n<p>Parsing carries the lowest risk of the four because it produces no score and expresses no preference. New York City&#39;s rule draws the same line, excluding from its definition of a simplified output tools that only translate or transcribe existing text, such as converting a resume from a PDF <a href=\"https://rules.cityofnewyork.us/wp-content/uploads/2023/04/DCWP-NOA-for-Use-of-Automated-Employment-Decisionmaking-Tools-2.pdf\" class=\"citation-ref\" data-citation-index=\"8\" target=\"_blank\" rel=\"noreferrer\">[8]</a>. The real risk is quiet data loss, since a mangled skills section becomes a thin record every downstream filter inherits.</p>\n","outroMarkdown":null,"outroHtml":null,"contentJson":{},"configJson":{},"noteMarkdown":null,"noteHtml":null,"sortOrder":30},{"id":"ea225c94-6b17-449c-87c0-6d55c1917e24","sectionKey":"does_it_work","sectionType":"markdown_section","heading":"Does AI resume screening actually work?","introMarkdown":"\"Works\" has no agreed meaning in this market, and that is the first thing to settle before reading any accuracy claim. A screening model is trained against a target variable someone chose, and in commercial tools that choice usually belongs to the buyer. Raghavan, Barocas, Kleinberg, and Levy reviewed 18 vendors of algorithmic pre-employment assessments and found vendors leaving it to clients to decide what outcomes to predict, naming performance reviews, sales numbers, and retention time as examples [4]. Vendor websites generally do not make clear whether models are validated or how validation data was selected [4].\n\nSo a tool can post an excellent accuracy figure and still be worth nothing to you. If the label it learned was \"this applicant was advanced to a phone screen last year,\" high accuracy means it agrees with last year's recruiters. It has automated the existing screen, including whatever that screen was getting wrong, with no evidence about who does the job well.\n\nA tool trained on real performance data still meets a ceiling that predates this software. Sackett and colleagues re-examined the meta-analytic validity literature and concluded that range restriction corrections had substantially overstated the validity of most selection procedures, cutting mean estimates by .10 to .20 points, with the structured interview highest ranked [7]. Nothing in personnel selection predicts job performance with anything close to precision, so a tool claiming to find the best candidate is claiming something the field has never shown.","introHtml":"<p>&quot;Works&quot; has no agreed meaning in this market, and that is the first thing to settle before reading any accuracy claim. A screening model is trained against a target variable someone chose, and in commercial tools that choice usually belongs to the buyer. Raghavan, Barocas, Kleinberg, and Levy reviewed 18 vendors of algorithmic pre-employment assessments and found vendors leaving it to clients to decide what outcomes to predict, naming performance reviews, sales numbers, and retention time as examples <a href=\"https://arxiv.org/pdf/1906.09208\" class=\"citation-ref\" data-citation-index=\"4\" target=\"_blank\" rel=\"noreferrer\">[4]</a>. Vendor websites generally do not make clear whether models are validated or how validation data was selected <a href=\"https://arxiv.org/pdf/1906.09208\" class=\"citation-ref\" data-citation-index=\"4\" target=\"_blank\" rel=\"noreferrer\">[4]</a>.</p>\n<p>So a tool can post an excellent accuracy figure and still be worth nothing to you. If the label it learned was &quot;this applicant was advanced to a phone screen last year,&quot; high accuracy means it agrees with last year&#39;s recruiters. It has automated the existing screen, including whatever that screen was getting wrong, with no evidence about who does the job well.</p>\n<p>A tool trained on real performance data still meets a ceiling that predates this software. Sackett and colleagues re-examined the meta-analytic validity literature and concluded that range restriction corrections had substantially overstated the validity of most selection procedures, cutting mean estimates by .10 to .20 points, with the structured interview highest ranked <a href=\"https://pubmed.ncbi.nlm.nih.gov/34968080/\" class=\"citation-ref\" data-citation-index=\"7\" target=\"_blank\" rel=\"noreferrer\">[7]</a>. Nothing in personnel selection predicts job performance with anything close to precision, so a tool claiming to find the best candidate is claiming something the field has never shown.</p>\n","outroMarkdown":null,"outroHtml":null,"contentJson":{},"configJson":{},"noteMarkdown":null,"noteHtml":null,"sortOrder":40},{"id":"88a13314-b1e0-4d4a-a3b2-90e44c7e8f57","sectionKey":"measuring_bias","sectionType":"markdown_section","heading":"How would you know whether your tool is biased?","introMarkdown":"The measurement standard already exists in federal enforcement practice: compare selection rates across groups and flag any group below four-fifths of the highest rate. Under 29 CFR 1607.4(D), such a rate \"will generally be regarded by the Federal enforcement agencies as evidence of adverse impact\" [12], and the Uniform Guidelines require validity evidence only where a procedure adversely affects a group's opportunities [13].\n\nNew York City turned that arithmetic into a filing requirement. A bias audit under the DCWP rule must calculate the selection rate and impact ratio for every EEO-1 category, separately for sex, for race and ethnicity, and for intersectional categories, and must report how many assessed people fell into an unknown category [8]. Where the tool scores rather than selects, scoring rate above the sample median replaces selection rate [8].\n\nHow often any of this happens is a separate question. A 2024 study at the ACM Conference on Fairness, Accountability, and Transparency sent 155 investigators to 391 employer websites and found audit reports for 18 employers and transparency notices for 13 [9]. The authors call that \"null compliance\" rather than non-compliance, since employers decide for themselves whether a tool is in scope [9].","introHtml":"<p>The measurement standard already exists in federal enforcement practice: compare selection rates across groups and flag any group below four-fifths of the highest rate. Under 29 CFR 1607.4(D), such a rate &quot;will generally be regarded by the Federal enforcement agencies as evidence of adverse impact&quot; <a href=\"https://www.law.cornell.edu/cfr/text/29/1607.4\" class=\"citation-ref\" data-citation-index=\"12\" target=\"_blank\" rel=\"noreferrer\">[12]</a>, and the Uniform Guidelines require validity evidence only where a procedure adversely affects a group&#39;s opportunities <a href=\"https://www.eeoc.gov/laws/guidance/questions-and-answers-clarify-and-provide-common-interpretation-uniform-guidelines\" class=\"citation-ref\" data-citation-index=\"13\" target=\"_blank\" rel=\"noreferrer\">[13]</a>.</p>\n<p>New York City turned that arithmetic into a filing requirement. A bias audit under the DCWP rule must calculate the selection rate and impact ratio for every EEO-1 category, separately for sex, for race and ethnicity, and for intersectional categories, and must report how many assessed people fell into an unknown category <a href=\"https://rules.cityofnewyork.us/wp-content/uploads/2023/04/DCWP-NOA-for-Use-of-Automated-Employment-Decisionmaking-Tools-2.pdf\" class=\"citation-ref\" data-citation-index=\"8\" target=\"_blank\" rel=\"noreferrer\">[8]</a>. Where the tool scores rather than selects, scoring rate above the sample median replaces selection rate <a href=\"https://rules.cityofnewyork.us/wp-content/uploads/2023/04/DCWP-NOA-for-Use-of-Automated-Employment-Decisionmaking-Tools-2.pdf\" class=\"citation-ref\" data-citation-index=\"8\" target=\"_blank\" rel=\"noreferrer\">[8]</a>.</p>\n<p>How often any of this happens is a separate question. A 2024 study at the ACM Conference on Fairness, Accountability, and Transparency sent 155 investigators to 391 employer websites and found audit reports for 18 employers and transparency notices for 13 <a href=\"https://arxiv.org/html/2406.01399v1\" class=\"citation-ref\" data-citation-index=\"9\" target=\"_blank\" rel=\"noreferrer\">[9]</a>. The authors call that &quot;null compliance&quot; rather than non-compliance, since employers decide for themselves whether a tool is in scope <a href=\"https://arxiv.org/html/2406.01399v1\" class=\"citation-ref\" data-citation-index=\"9\" target=\"_blank\" rel=\"noreferrer\">[9]</a>.</p>\n","outroMarkdown":null,"outroHtml":null,"contentJson":{},"configJson":{},"noteMarkdown":null,"noteHtml":null,"sortOrder":50},{"id":"ca1d01d7-ef85-4990-9e54-3691c57a76b7","sectionKey":"regulation","sectionType":"markdown_section","heading":"What do the rules require as of July 2026?","introMarkdown":"Three jurisdictions impose specific obligations on employers using automated tools to screen candidates, while the federal baseline stays ordinary discrimination law rather than an AI statute. This is general information, not legal advice; confirm current requirements with counsel before deploying anything.\n\n### New York City\n\nLocal Law 144, implemented by DCWP rules effective July 5, 2023, bars use of an automated employment decision tool unless it has had a bias audit within the past year, a summary of results is published on your site before use, and notice reached candidates at least 10 business days beforehand with instructions for requesting an alternative process or accommodation [8]. The definition reaches any tool whose simplified output is relied on alone, weighted more heavily than any other criterion, or used to overrule human conclusions, so a ranking tool sitting beside recruiter judgment can still be covered [8].\n\n### Illinois\n\nHB 3773 became Public Act 103-0804 and took effect January 1, 2026, amending the Illinois Human Rights Act to bar artificial intelligence that discriminates on the basis of a protected class or that uses zip codes as a proxy for one, and requiring notice to workers when AI is used in hiring, promotion, discipline, and other employment decisions [10].\n\n### Colorado\n\nColorado's position moved twice, so older summaries are unreliable. SB26-189, signed May 14, 2026, repeals the 2024 Colorado AI Act and reenacts it as a disclosure and rights framework effective January 1, 2027 [11]. Deployers will have to notify workers before covered technology influences an employment decision, give a plain-language explanation of an adverse outcome within 30 days, and allow meaningful human review [11].\n\n### Federal\n\nThe Uniform Guidelines on Employee Selection Procedures still govern, and they reach any procedure used to make an employment decision, including application screening [13]. The EEOC's 2023 technical assistance document on adverse impact in algorithmic selection no longer resolves on eeoc.gov, so anchor your analysis to the Uniform Guidelines and 29 CFR 1607 instead [12][13].","introHtml":"<p>Three jurisdictions impose specific obligations on employers using automated tools to screen candidates, while the federal baseline stays ordinary discrimination law rather than an AI statute. This is general information, not legal advice; confirm current requirements with counsel before deploying anything.</p>\n<h3>New York City</h3>\n<p>Local Law 144, implemented by DCWP rules effective July 5, 2023, bars use of an automated employment decision tool unless it has had a bias audit within the past year, a summary of results is published on your site before use, and notice reached candidates at least 10 business days beforehand with instructions for requesting an alternative process or accommodation <a href=\"https://rules.cityofnewyork.us/wp-content/uploads/2023/04/DCWP-NOA-for-Use-of-Automated-Employment-Decisionmaking-Tools-2.pdf\" class=\"citation-ref\" data-citation-index=\"8\" target=\"_blank\" rel=\"noreferrer\">[8]</a>. The definition reaches any tool whose simplified output is relied on alone, weighted more heavily than any other criterion, or used to overrule human conclusions, so a ranking tool sitting beside recruiter judgment can still be covered <a href=\"https://rules.cityofnewyork.us/wp-content/uploads/2023/04/DCWP-NOA-for-Use-of-Automated-Employment-Decisionmaking-Tools-2.pdf\" class=\"citation-ref\" data-citation-index=\"8\" target=\"_blank\" rel=\"noreferrer\">[8]</a>.</p>\n<h3>Illinois</h3>\n<p>HB 3773 became Public Act 103-0804 and took effect January 1, 2026, amending the Illinois Human Rights Act to bar artificial intelligence that discriminates on the basis of a protected class or that uses zip codes as a proxy for one, and requiring notice to workers when AI is used in hiring, promotion, discipline, and other employment decisions <a href=\"https://www.ilga.gov/ftp/legislation/103/BillStatus/HTML/10300HB3773.html\" class=\"citation-ref\" data-citation-index=\"10\" target=\"_blank\" rel=\"noreferrer\">[10]</a>.</p>\n<h3>Colorado</h3>\n<p>Colorado&#39;s position moved twice, so older summaries are unreliable. SB26-189, signed May 14, 2026, repeals the 2024 Colorado AI Act and reenacts it as a disclosure and rights framework effective January 1, 2027 <a href=\"https://leg.colorado.gov/bills/sb26-189\" class=\"citation-ref\" data-citation-index=\"11\" target=\"_blank\" rel=\"noreferrer\">[11]</a>. Deployers will have to notify workers before covered technology influences an employment decision, give a plain-language explanation of an adverse outcome within 30 days, and allow meaningful human review <a href=\"https://leg.colorado.gov/bills/sb26-189\" class=\"citation-ref\" data-citation-index=\"11\" target=\"_blank\" rel=\"noreferrer\">[11]</a>.</p>\n<h3>Federal</h3>\n<p>The Uniform Guidelines on Employee Selection Procedures still govern, and they reach any procedure used to make an employment decision, including application screening <a href=\"https://www.eeoc.gov/laws/guidance/questions-and-answers-clarify-and-provide-common-interpretation-uniform-guidelines\" class=\"citation-ref\" data-citation-index=\"13\" target=\"_blank\" rel=\"noreferrer\">[13]</a>. The EEOC&#39;s 2023 technical assistance document on adverse impact in algorithmic selection no longer resolves on eeoc.gov, so anchor your analysis to the Uniform Guidelines and 29 CFR 1607 instead <a href=\"https://www.law.cornell.edu/cfr/text/29/1607.4\" class=\"citation-ref\" data-citation-index=\"12\" target=\"_blank\" rel=\"noreferrer\">[12]</a><a href=\"https://www.eeoc.gov/laws/guidance/questions-and-answers-clarify-and-provide-common-interpretation-uniform-guidelines\" class=\"citation-ref\" data-citation-index=\"13\" target=\"_blank\" rel=\"noreferrer\">[13]</a>.</p>\n","outroMarkdown":null,"outroHtml":null,"contentJson":{},"configJson":{},"noteMarkdown":null,"noteHtml":null,"sortOrder":60},{"id":"27242296-5ea6-48c3-9f62-1d94e0125c3b","sectionKey":"when_defensible","sectionType":"markdown_section","heading":"When is AI resume screening defensible, and when is it not?","introMarkdown":"AI screening holds up where applicant volume exceeds human capacity, requirements are concrete, the tool ranks instead of rejecting, and someone recalculates impact ratios on a set cadence.\n\n### Conditions where it holds up\n\nHigh-volume hourly and entry-level requisitions are the clearest case, because the realistic alternative is a recruiter skimming a thousand resumes at a few seconds each. A ranking that surfaces candidates for a human to read, with the full pool still reachable, changes the order of review instead of the outcome. A forklift certification is a fact you can verify, and a tool sorting on facts is auditable in a way a tool scoring \"culture fit\" is not [4]. The configuration that survives scrutiny is ranking, plus human review of the list, plus a standing adverse impact calculation [12].\n\n### Conditions where it is hard to defend\n\nSmall requisitions rarely justify the exposure, since forty applications are readable by a person and the audit overhead outweighs the time saved. Subjective and senior roles fit poorly, because the qualities that matter are the ones a model reaches through proxies, where the documented disparities live [1][5]. Automatic rejection without review turns a ranking error into a final decision, and it sits inside the NYC definition of a tool relied on alone [8]. Without category-level rates on your applicant flow, you cannot see whether the four-fifths threshold is being crossed [8][12].","introHtml":"<p>AI screening holds up where applicant volume exceeds human capacity, requirements are concrete, the tool ranks instead of rejecting, and someone recalculates impact ratios on a set cadence.</p>\n<h3>Conditions where it holds up</h3>\n<p>High-volume hourly and entry-level requisitions are the clearest case, because the realistic alternative is a recruiter skimming a thousand resumes at a few seconds each. A ranking that surfaces candidates for a human to read, with the full pool still reachable, changes the order of review instead of the outcome. A forklift certification is a fact you can verify, and a tool sorting on facts is auditable in a way a tool scoring &quot;culture fit&quot; is not <a href=\"https://arxiv.org/pdf/1906.09208\" class=\"citation-ref\" data-citation-index=\"4\" target=\"_blank\" rel=\"noreferrer\">[4]</a>. The configuration that survives scrutiny is ranking, plus human review of the list, plus a standing adverse impact calculation <a href=\"https://www.law.cornell.edu/cfr/text/29/1607.4\" class=\"citation-ref\" data-citation-index=\"12\" target=\"_blank\" rel=\"noreferrer\">[12]</a>.</p>\n<h3>Conditions where it is hard to defend</h3>\n<p>Small requisitions rarely justify the exposure, since forty applications are readable by a person and the audit overhead outweighs the time saved. Subjective and senior roles fit poorly, because the qualities that matter are the ones a model reaches through proxies, where the documented disparities live <a href=\"https://arxiv.org/abs/2407.20371\" class=\"citation-ref\" data-citation-index=\"1\" target=\"_blank\" rel=\"noreferrer\">[1]</a><a href=\"https://www.upturn.org/work/help-wanted/\" class=\"citation-ref\" data-citation-index=\"5\" target=\"_blank\" rel=\"noreferrer\">[5]</a>. Automatic rejection without review turns a ranking error into a final decision, and it sits inside the NYC definition of a tool relied on alone <a href=\"https://rules.cityofnewyork.us/wp-content/uploads/2023/04/DCWP-NOA-for-Use-of-Automated-Employment-Decisionmaking-Tools-2.pdf\" class=\"citation-ref\" data-citation-index=\"8\" target=\"_blank\" rel=\"noreferrer\">[8]</a>. Without category-level rates on your applicant flow, you cannot see whether the four-fifths threshold is being crossed <a href=\"https://rules.cityofnewyork.us/wp-content/uploads/2023/04/DCWP-NOA-for-Use-of-Automated-Employment-Decisionmaking-Tools-2.pdf\" class=\"citation-ref\" data-citation-index=\"8\" target=\"_blank\" rel=\"noreferrer\">[8]</a><a href=\"https://www.law.cornell.edu/cfr/text/29/1607.4\" class=\"citation-ref\" data-citation-index=\"12\" target=\"_blank\" rel=\"noreferrer\">[12]</a>.</p>\n","outroMarkdown":null,"outroHtml":null,"contentJson":{},"configJson":{},"noteMarkdown":null,"noteHtml":null,"sortOrder":70},{"id":"1af2ddec-69a5-4358-9792-cc40807d33de","sectionKey":"vendor_questions","sectionType":"markdown_section","heading":"What should you require from a vendor before you buy?","introMarkdown":"Ask for five things in writing, and treat a refusal as its own answer.\n\n### The most recent bias audit, with impact ratios by category\n\nRequest the audit output itself rather than a compliance badge: selection or scoring rates and impact ratios for each sex, race and ethnicity, and intersectional category, plus the audit date, data source, and count excluded as unknown [8]. An audit on the vendor's pooled data across many employers is not the same as one on your applicant flow, and the NYC rule expects the auditor to get your own historical data [8].\n\n### The target variable and the training population\n\nAsk what the model was trained to predict and over which people, in one sentence each. Vendors hand the outcome choice to the client and rarely publish validation methodology, so the answer will not be on the website [4]. If the target is a past decision or rating, the tool measures agreement with prior screening, and should be judged on that.\n\n### An explanation you could give a rejected candidate\n\nAsk the vendor to produce, for one real applicant, the reason that applicant ranked where they did. Colorado will require a plain-language explanation of an adverse decision within 30 days plus meaningful human review [11], so a tool that cannot generate one leaves that obligation to you.\n\n### The model version and the change policy\n\nLanguage-model behavior varies by release, and the fourteen-model comparison found demographic gaps flipping direction across generations [3]. Get the version identifier, the change-notice policy, and a written commitment to re-audit after an upgrade.\n\n### The evidence behind any performance number\n\nPerformance figures are marketing until the underlying study is on the table. Workday markets HiredScore AI for Recruiting with a claimed \"54% increase in recruiter capacity within 10 months of launch\" and \"unbiased, AI-driven candidate grading\" [14]; both are vendor claims, not published validation of screening accuracy. Ask which customers, over what period, against what baseline.","introHtml":"<p>Ask for five things in writing, and treat a refusal as its own answer.</p>\n<h3>The most recent bias audit, with impact ratios by category</h3>\n<p>Request the audit output itself rather than a compliance badge: selection or scoring rates and impact ratios for each sex, race and ethnicity, and intersectional category, plus the audit date, data source, and count excluded as unknown <a href=\"https://rules.cityofnewyork.us/wp-content/uploads/2023/04/DCWP-NOA-for-Use-of-Automated-Employment-Decisionmaking-Tools-2.pdf\" class=\"citation-ref\" data-citation-index=\"8\" target=\"_blank\" rel=\"noreferrer\">[8]</a>. An audit on the vendor&#39;s pooled data across many employers is not the same as one on your applicant flow, and the NYC rule expects the auditor to get your own historical data <a href=\"https://rules.cityofnewyork.us/wp-content/uploads/2023/04/DCWP-NOA-for-Use-of-Automated-Employment-Decisionmaking-Tools-2.pdf\" class=\"citation-ref\" data-citation-index=\"8\" target=\"_blank\" rel=\"noreferrer\">[8]</a>.</p>\n<h3>The target variable and the training population</h3>\n<p>Ask what the model was trained to predict and over which people, in one sentence each. Vendors hand the outcome choice to the client and rarely publish validation methodology, so the answer will not be on the website <a href=\"https://arxiv.org/pdf/1906.09208\" class=\"citation-ref\" data-citation-index=\"4\" target=\"_blank\" rel=\"noreferrer\">[4]</a>. If the target is a past decision or rating, the tool measures agreement with prior screening, and should be judged on that.</p>\n<h3>An explanation you could give a rejected candidate</h3>\n<p>Ask the vendor to produce, for one real applicant, the reason that applicant ranked where they did. Colorado will require a plain-language explanation of an adverse decision within 30 days plus meaningful human review <a href=\"https://leg.colorado.gov/bills/sb26-189\" class=\"citation-ref\" data-citation-index=\"11\" target=\"_blank\" rel=\"noreferrer\">[11]</a>, so a tool that cannot generate one leaves that obligation to you.</p>\n<h3>The model version and the change policy</h3>\n<p>Language-model behavior varies by release, and the fourteen-model comparison found demographic gaps flipping direction across generations <a href=\"https://arxiv.org/abs/2606.28978\" class=\"citation-ref\" data-citation-index=\"3\" target=\"_blank\" rel=\"noreferrer\">[3]</a>. Get the version identifier, the change-notice policy, and a written commitment to re-audit after an upgrade.</p>\n<h3>The evidence behind any performance number</h3>\n<p>Performance figures are marketing until the underlying study is on the table. Workday markets HiredScore AI for Recruiting with a claimed &quot;54% increase in recruiter capacity within 10 months of launch&quot; and &quot;unbiased, AI-driven candidate grading&quot; <a href=\"https://www.workday.com/en-us/products/talent-management/ai-recruiting.html\" class=\"citation-ref\" data-citation-index=\"14\" target=\"_blank\" rel=\"noreferrer\">[14]</a>; both are vendor claims, not published validation of screening accuracy. Ask which customers, over what period, against what baseline.</p>\n","outroMarkdown":null,"outroHtml":null,"contentJson":{},"configJson":{},"noteMarkdown":null,"noteHtml":null,"sortOrder":80},{"id":"8f303757-e907-4849-b132-128e32ee83b2","sectionKey":"contributor_perspective","sectionType":"markdown_section","heading":"How this answer was researched","introMarkdown":"This record was assembled from primary research papers and primary legal text rather than vendor marketing, because the vendor literature on this question is close to uniformly promotional. The audit studies were opened and read directly, and the statutes and rules were read in the original rather than through law firm summaries, which is why the Colorado section reflects the May 2026 replacement law instead of the 2024 act much secondary coverage still describes. One document others cite often, the EEOC's 2023 technical assistance on algorithmic selection, no longer resolves and is deliberately absent below. HR practitioners running these tools, and vendors who believe a claim here misstates their product, are invited to send corrections with evidence. Bias audit results from real applicant flows would be especially useful, since almost none are public.","introHtml":"<p>This record was assembled from primary research papers and primary legal text rather than vendor marketing, because the vendor literature on this question is close to uniformly promotional. The audit studies were opened and read directly, and the statutes and rules were read in the original rather than through law firm summaries, which is why the Colorado section reflects the May 2026 replacement law instead of the 2024 act much secondary coverage still describes. One document others cite often, the EEOC&#39;s 2023 technical assistance on algorithmic selection, no longer resolves and is deliberately absent below. HR practitioners running these tools, and vendors who believe a claim here misstates their product, are invited to send corrections with evidence. Bias audit results from real applicant flows would be especially useful, since almost none are public.</p>\n","outroMarkdown":null,"outroHtml":null,"contentJson":{},"configJson":{},"noteMarkdown":"This answer was written and reviewed by the AnswerStack Editorial Team, which has no commercial stake in the products, companies, or methods discussed. Every claim is cited inline and verified on the dates shown.","noteHtml":"<p>This answer was written and reviewed by the AnswerStack Editorial Team, which has no commercial stake in the products, companies, or methods discussed. Every claim is cited inline and verified on the dates shown.</p>\n","sortOrder":90}],"citations":[{"title":"Gender, Race, and Intersectional Bias in Resume Screening via Language Model Retrieval","url":"https://arxiv.org/abs/2407.20371","excerpt":"Using that framework, we then perform a resume audit study to determine whether a selection of Massive Text Embedding (MTE) models are biased in resume screening scenarios.","quoteText":null,"sourceRole":"PRIMARY","verifiedAt":"2026-07-21T00:00:00","supportsText":"85.1 percent preference for white-associated names, 11.1 percent for female-associated names, up to 100 percent disadvantage for Black male-associated names; 500+ resumes, 500 job descriptions, nine occupations, Massive Text Embedding models with no fine-tuning","domain":"arxiv.org","publisherName":"Kyra Wilson and Aylin Caliskan, University of Washington (AAAI/ACM AIES 2024)"},{"title":"No Thoughts Just AI: Biased LLM Hiring Recommendations Alter Human Decision Making and Limit Human Autonomy","url":"https://arxiv.org/html/2509.04404v2","excerpt":"When interacting with AI favoring a particular group, people select those candidates up to 90% of the time.","quoteText":null,"sourceRole":"INDEPENDENT","verifiedAt":"2026-07-21T00:00:00","supportsText":"528 participants and 1,526 resume screening decisions; equal selection without AI recommendations; selection of the favored group up to 90 percent of the time with biased recommendations; effect persisting among participants who rated the recommendations poorly","domain":"arxiv.org","publisherName":"Kyra Wilson, Mattea Sim, Anna-Maria Gueorguieva, Aylin Caliskan"},{"title":"Can LLMs Hire Fairly? Racial Bias in Resume Screening","url":"https://arxiv.org/abs/2606.28978","excerpt":"The sole 2023-vintage model reproduces the pro-White callback gap.","quoteText":null,"sourceRole":"INDEPENDENT","verifiedAt":"2026-07-21T00:00:00","supportsText":"Fourteen models, 24,024 paired postings per model; 2023-vintage model shows a +2.12 percentage point pro-white gap; 2024 and later models show null gaps or reversals up to 3.01 points; same generational pattern on gender","domain":"arxiv.org","publisherName":"Zhenyu Gao, Wenxi Jiang, Yutong Yan (arXiv preprint, June 2026)"},{"title":"Mitigating Bias in Algorithmic Hiring: Evaluating Claims and Practices","url":"https://arxiv.org/pdf/1906.09208","excerpt":"Vendors in general leave it up to clients to determine what outcomes they want to predict, including, for example, performance reviews, sales numbers, and retention time.","quoteText":null,"sourceRole":"INDEPENDENT","verifiedAt":"2026-07-21T00:00:00","supportsText":"18 vendors of algorithmic pre-employment assessments reviewed; eight build assessments from client employee data; vendors leave target variable choice to clients; websites generally do not make validation methodology clear","domain":"arxiv.org","publisherName":"Raghavan, Barocas, Kleinberg and Levy (ACM FAccT 2020)"},{"title":"Help Wanted: An Examination of Hiring Algorithms, Equity, and Bias","url":"https://www.upturn.org/work/help-wanted/","excerpt":"Removing or obscuring sensitive factors like gender and race will not prevent predictive models from reflecting patterns of bias.","quoteText":null,"sourceRole":"INDEPENDENT","verifiedAt":"2026-07-21T00:00:00","supportsText":"Models trained on historical hiring records reproduce those patterns; removing protected attributes does not prevent proxy effects; the Jared and high school lacrosse example of a meaningless learned correlation","domain":"upturn.org","publisherName":"Upturn"},{"title":"Amazon ditched AI recruitment software because it was biased against women","url":"https://www.technologyreview.com/2018/10/10/139858/amazon-ditched-ai-recruitment-software-because-it-was-biased-against-women/","excerpt":"The company lost confidence that the program was indeed gender neutral in all other areas.","quoteText":null,"sourceRole":"INDEPENDENT","verifiedAt":"2026-07-21T00:00:00","supportsText":"Amazon system started 2014, scrapped 2017, trained on ten years of submitted resumes, penalized resumes containing the word women's and the names of certain all-women colleges","domain":"technologyreview.com","publisherName":"MIT Technology Review"},{"title":"Revisiting meta-analytic estimates of validity in personnel selection: Addressing systematic overcorrection for restriction of range","url":"https://pubmed.ncbi.nlm.nih.gov/34968080/","excerpt":"Key findings are that most of the same selection procedures that ranked high in prior summaries remain high in rank, but with mean validity estimates reduced by .10-.20 points. Structured interviews emerged as the top-ranked selection procedure.","quoteText":null,"sourceRole":"INDEPENDENT","verifiedAt":"2026-07-21T00:00:00","supportsText":"Prior validity estimates for selection procedures substantially overstated; revised estimates reduced by .10 to .20; structured interviews ranked highest","domain":"pubmed.ncbi.nlm.nih.gov","publisherName":"Journal of Applied Psychology (Sackett, Zhang, Berry & Lievens, 2022) via PubMed"},{"title":"Notice of Adoption of Final Rule: Automated Employment Decision Tools (6 RCNY Subchapter T)","url":"https://rules.cityofnewyork.us/wp-content/uploads/2023/04/DCWP-NOA-for-Use-of-Automated-Employment-Decisionmaking-Tools-2.pdf","excerpt":"Provide notice on the employment section of its website in a clear and conspicuous manner at least 10 business days before use of an AEDT.","quoteText":null,"sourceRole":"PRIMARY","verifiedAt":"2026-07-21T00:00:00","supportsText":"AEDT definition and the substantially assist test; bias audit contents including selection rate, impact ratio, sex, race/ethnicity and intersectional categories, unknown-category count, scoring rate above median, the 2 percent exclusion; published results and six-month posting; 10 business days noti","domain":"rules.cityofnewyork.us","publisherName":"New York City Department of Consumer and Worker Protection"},{"title":"Null Compliance: NYC Local Law 144 and the Challenges of Algorithm Accountability","url":"https://arxiv.org/html/2406.01399v1","excerpt":"Null compliance describes a state in which the absence of evidence of compliance cannot be ascertained as non-compliance.","quoteText":null,"sourceRole":"INDEPENDENT","verifiedAt":"2026-07-21T00:00:00","supportsText":"155 investigators searched 391 employer websites; 18 posted audit reports and 13 posted transparency notices; definition and argument for null compliance","domain":"arxiv.org","publisherName":"Wright, Muenster, Vecchione, Qu, Cai, Smith, Metcalf and Matias (ACM FAccT 2024)"},{"title":"Illinois HB 3773 bill status, Public Act 103-0804","url":"https://www.ilga.gov/ftp/legislation/103/BillStatus/HTML/10300HB3773.html","excerpt":"Bill status page confirming Public Act 103-0804 and its January 1, 2026 effective date; the act bars AI use that discriminates on a protected class or uses zip codes as a proxy for one.","quoteText":null,"sourceRole":"PRIMARY","verifiedAt":"2026-07-21T00:00:00","supportsText":"Public Act 103-0804, effective January 1, 2026; amends the Illinois Human Rights Act on AI discrimination, zip codes as a proxy, and employee notice","domain":"ilga.gov","publisherName":"Illinois General Assembly"},{"title":"SB26-189 Automated Decision-Making Technology","url":"https://leg.colorado.gov/bills/sb26-189","excerpt":"Repeals and reenacts those provisions with new requirements.","quoteText":null,"sourceRole":"PRIMARY","verifiedAt":"2026-07-21T00:00:00","supportsText":"Signed May 14, 2026; repeals and reenacts the 2024 Colorado AI Act; effective January 1, 2027; deployer notice, 30-day plain-language explanation of adverse decisions, right to meaningful human review, attorney general enforcement with a 60-day cure period","domain":"leg.colorado.gov","publisherName":"Colorado General Assembly"},{"title":"29 CFR 1607.4: Information on impact (Uniform Guidelines on Employee Selection Procedures)","url":"https://www.law.cornell.edu/cfr/text/29/1607.4","excerpt":"A selection rate for any race, sex, or ethnic group which is less than four-fifths (4/5) (or eighty percent) of the rate for the group with the highest rate will generally be regarded by the Federal enforcement agencies as evidence of adverse impact.","quoteText":null,"sourceRole":"PRIMARY","verifiedAt":"2026-07-21T00:00:00","supportsText":"Four-fifths rule as evidence of adverse impact; requirement to keep records of impact by race, sex, and ethnic group","domain":"law.cornell.edu","publisherName":"Legal Information Institute, Cornell Law School"},{"title":"Questions and Answers to Clarify and Provide a Common Interpretation of the Uniform Guidelines on Employee Selection Procedures","url":"https://www.eeoc.gov/laws/guidance/questions-and-answers-clarify-and-provide-common-interpretation-uniform-guidelines","excerpt":"The Uniform Guidelines require users to produce evidence of validity only when the selection procedure adversely affects the opportunities of a race, sex, or ethnic group for hire, transfer, promotion, retention or other employment decision.","quoteText":null,"sourceRole":"PRIMARY","verifiedAt":"2026-07-21T00:00:00","supportsText":"Scope of employee selection procedures including application form screening; validity evidence required only where a procedure adversely affects a group","domain":"eeoc.gov","publisherName":"U.S. Equal Employment Opportunity Commission"},{"title":"HiredScore AI for Recruiting","url":"https://www.workday.com/en-us/products/talent-management/ai-recruiting.html","excerpt":"54% increase in recruiter capacity within 10 months of launch.","quoteText":null,"sourceRole":"SUPPORTING","verifiedAt":"2026-07-21T00:00:00","supportsText":"Example of a live vendor claim: 54 percent increase in recruiter capacity within 10 months of launch, and unbiased AI-driven candidate grading. Vendor claim, not independent validation.","domain":"workday.com","publisherName":"Workday"}],"revisions":[],"relatedAnswers":[{"id":"8b7d74e4-7ae0-4d1e-8cff-6b3dcda9e3d4","slug":"candidates-cheat-on-online-skills-assessments-and-does-proctoring","question":"Do candidates cheat on online skills assessments, and does proctoring / anti-cheat actually stop it (including AI/ChatGPT)?","publishedAt":"2026-07-21T21:20:42.255","confidenceScore":78,"confidenceLabel":"Medium","industry":{"id":"e1a1da3c-52a8-4e99-88ad-d08ef17d7c3f","slug":"hr-technology","label":"HR Technology","description":"ATS, HRIS, payroll, and workforce management"},"topic":{"slug":"candidate-screening-assessment","label":"Candidate Screening & Assessment","description":"Background Checks, Pre-Employment Testing, Reference Checks, Skills Assessment","schemaKind":null},"contributor":{"id":"ec39deab-44fe-48d8-9029-fefe993ab85a","slug":"answer-stack","displayName":"AnswerStack","websiteUrl":null},"snippet":"Cheating on remote skills assessments is real, and generative AI widened it for coding challenges and knowledge questions. Proctoring catches tab switching, pasted code, proxy test takers, and leaked items, and it cannot see a phone outside the webcam frame. Every published detection rate is a vendor figure. Biometric identity checks and recorded video also carry Illinois BIPA and AI Video Interview Act exposure.","url":"/q/candidates-cheat-on-online-skills-assessments-and-does-proctoring"},{"id":"7c1df232-4afc-4f90-acc6-373bb41cf94d","slug":"screening-background-check-software-integrate-with-my-ats","question":"Does screening / background check software integrate with my ATS or HRIS?","publishedAt":"2026-07-21T21:20:40.075","confidenceScore":88,"confidenceLabel":"High","industry":{"id":"e1a1da3c-52a8-4e99-88ad-d08ef17d7c3f","slug":"hr-technology","label":"HR Technology","description":"ATS, HRIS, payroll, and workforce management"},"topic":{"slug":"candidate-screening-assessment","label":"Candidate Screening & Assessment","description":"Background Checks, Pre-Employment Testing, Reference Checks, Skills Assessment","schemaKind":null},"contributor":{"id":"ec39deab-44fe-48d8-9029-fefe993ab85a","slug":"answer-stack","displayName":"AnswerStack","websiteUrl":null},"snippet":"Screening and assessment vendors ship certified ATS and HRIS connectors, so ordering, status, and results move without rekeying. The mechanics that decide whether the integration helps are narrower: where the FCRA disclosure and authorization are hosted, what writes back, who can read results, and whether adverse action fires from the screening platform or your ATS.","url":"/q/screening-background-check-software-integrate-with-my-ats"},{"id":"ef2314e9-45b8-4867-adcb-bb8bb029b6e7","slug":"pre-employment-assessments-and-background-checks-legal-eeoc-ada-fcra","question":"Are pre-employment assessments and background checks legal (EEOC, ADA, FCRA, disparate impact)?","publishedAt":"2026-07-21T21:20:37.515","confidenceScore":86,"confidenceLabel":"High","industry":{"id":"e1a1da3c-52a8-4e99-88ad-d08ef17d7c3f","slug":"hr-technology","label":"HR Technology","description":"ATS, HRIS, payroll, and workforce management"},"topic":{"slug":"candidate-screening-assessment","label":"Candidate Screening & Assessment","description":"Background Checks, Pre-Employment Testing, Reference Checks, Skills Assessment","schemaKind":null},"contributor":{"id":"ec39deab-44fe-48d8-9029-fefe993ab85a","slug":"answer-stack","displayName":"AnswerStack","websiteUrl":null},"snippet":"Screening is lawful in principle and regulated in the details. Third-party background reports trigger FCRA disclosure, authorization, and a two-step adverse action sequence. Tests and criminal-record screens face Title VII adverse impact analysis under the Uniform Guidelines. The ADA restricts what you may ask or examine before a conditional offer. Fair-chance, salary history, credit, and automated hiring tool rules vary by state and city. A June 2026 federal opinion narrowed enforcement posture on disparate impact without repealing the statute.","url":"/q/pre-employment-assessments-and-background-checks-legal-eeoc-ada-fcra"},{"id":"c0112a94-3c83-4520-9777-fff00c999af6","slug":"pre-employment-tests-skills-assessments-actually-worth-it-do-they","question":"Are pre-employment tests and skills assessments actually worth it, and do they predict job performance?","publishedAt":"2026-07-21T21:20:35.197","confidenceScore":88,"confidenceLabel":"High","industry":{"id":"e1a1da3c-52a8-4e99-88ad-d08ef17d7c3f","slug":"hr-technology","label":"HR Technology","description":"ATS, HRIS, payroll, and workforce management"},"topic":{"slug":"candidate-screening-assessment","label":"Candidate Screening & Assessment","description":"Background Checks, Pre-Employment Testing, Reference Checks, Skills Assessment","schemaKind":null},"contributor":{"id":"ec39deab-44fe-48d8-9029-fefe993ab85a","slug":"answer-stack","displayName":"AnswerStack","websiteUrl":null},"snippet":"Selection tests do predict performance, but the evidence base was revised downward in 2022 after researchers showed that decades of range restriction corrections had inflated published validity. Structured interviews now top the ranking at .42, while general cognitive ability sits at .31 and carries the largest Black-White subgroup difference of any common method. Value depends on volume, measurable job requirements, what the assessment replaces, and whether you can monitor selection rates by group.","url":"/q/pre-employment-tests-skills-assessments-actually-worth-it-do-they"}],"contributorStats":{"verifiedAnswers":224,"openDisputes":0},"schemaJson":{"@context":"https://schema.org","@type":"Question","name":"Does AI resume screening software actually work, and is it worth it (bias, accuracy)?","text":"Does AI resume screening software actually work, and is it worth it (bias, accuracy)?","url":"https://www.answerstack.io/q/ai-resume-screening-software-actually-work","answerCount":1,"datePublished":"2026-07-21T21:20:44.609","author":{"@type":"Person","name":"AnswerStack Editorial Team","worksFor":{"@type":"Organization","name":"AnswerStack"},"url":"https://www.answerstack.io/contributors/answer-stack"},"about":[{"@type":"Thing","name":"Candidate Screening & Assessment"},{"@type":"Thing","name":"HR Technology"}],"acceptedAnswer":{"@type":"Answer","text":"AI resume screening reliably reproduces the ranking it was trained to reproduce, and independent audits show it does so with measurable demographic disparities: a University of Washington audit of embedding models ranking 500 real resumes against 500 job descriptions found white-associated names preferred in 85.1 percent of cases, female-associated names in 11.1 percent, and Black male-associated names disadvantaged in up to 100 percent of comparisons [1]. Whether it is worth buying turns on what the model was trained to predict, because vendors generally let the client pick the outcome, so a tool can be accurate at copying prior screening while adding nothing to hiring quality [4]. Corrected meta-analytic estimates also cut the validity of most selection procedures by .10 to .20 and left the structured interview highest ranked [7]. In New York City the tool needs a bias audit from the past year, published impact ratios, and candidate notice 10 business days before use [8]. This is general information, not legal advice.","url":"https://www.answerstack.io/q/ai-resume-screening-software-actually-work","upvoteCount":0,"datePublished":"2026-07-21T21:20:44.609","dateModified":"2026-07-21T00:00:00","author":{"@type":"Person","name":"AnswerStack Editorial Team","worksFor":{"@type":"Organization","name":"AnswerStack"},"url":"https://www.answerstack.io/contributors/answer-stack"},"citation":[{"@type":"CreativeWork","name":"Gender, Race, and Intersectional Bias in Resume Screening via Language Model Retrieval","url":"https://arxiv.org/abs/2407.20371"},{"@type":"CreativeWork","name":"No Thoughts Just AI: Biased LLM Hiring Recommendations Alter Human Decision Making and Limit Human Autonomy","url":"https://arxiv.org/html/2509.04404v2"},{"@type":"CreativeWork","name":"Can LLMs Hire Fairly? Racial Bias in Resume Screening","url":"https://arxiv.org/abs/2606.28978"},{"@type":"CreativeWork","name":"Mitigating Bias in Algorithmic Hiring: Evaluating Claims and Practices","url":"https://arxiv.org/pdf/1906.09208"},{"@type":"CreativeWork","name":"Help Wanted: An Examination of Hiring Algorithms, Equity, and Bias","url":"https://www.upturn.org/work/help-wanted/"},{"@type":"CreativeWork","name":"Amazon ditched AI recruitment software because it was biased against women","url":"https://www.technologyreview.com/2018/10/10/139858/amazon-ditched-ai-recruitment-software-because-it-was-biased-against-women/"},{"@type":"CreativeWork","name":"Revisiting meta-analytic estimates of validity in personnel selection: Addressing systematic overcorrection for restriction of range","url":"https://pubmed.ncbi.nlm.nih.gov/34968080/"},{"@type":"CreativeWork","name":"Notice of Adoption of Final Rule: Automated Employment Decision Tools (6 RCNY Subchapter T)","url":"https://rules.cityofnewyork.us/wp-content/uploads/2023/04/DCWP-NOA-for-Use-of-Automated-Employment-Decisionmaking-Tools-2.pdf"},{"@type":"CreativeWork","name":"Null Compliance: NYC Local Law 144 and the Challenges of Algorithm Accountability","url":"https://arxiv.org/html/2406.01399v1"},{"@type":"CreativeWork","name":"Illinois HB 3773 bill status, Public Act 103-0804","url":"https://www.ilga.gov/ftp/legislation/103/BillStatus/HTML/10300HB3773.html"},{"@type":"CreativeWork","name":"SB26-189 Automated Decision-Making Technology","url":"https://leg.colorado.gov/bills/sb26-189"},{"@type":"CreativeWork","name":"29 CFR 1607.4: Information on impact (Uniform Guidelines on Employee Selection Procedures)","url":"https://www.law.cornell.edu/cfr/text/29/1607.4"},{"@type":"CreativeWork","name":"Questions and Answers to Clarify and Provide a Common Interpretation of the Uniform Guidelines on Employee Selection Procedures","url":"https://www.eeoc.gov/laws/guidance/questions-and-answers-clarify-and-provide-common-interpretation-uniform-guidelines"},{"@type":"CreativeWork","name":"HiredScore AI for Recruiting","url":"https://www.workday.com/en-us/products/talent-management/ai-recruiting.html"}]}}}