{"id":1996,"date":"2026-08-21T07:51:42","date_gmt":"2026-08-21T02:21:42","guid":{"rendered":"https:\/\/www.editage.com\/blog\/?p=1996"},"modified":"2026-08-19T09:22:30","modified_gmt":"2026-08-19T03:52:30","slug":"what-are-ai-hallucinations-in-research-causes-examples-and-risks","status":"publish","type":"post","link":"https:\/\/www.editage.com\/blog\/what-are-ai-hallucinations-in-research-causes-examples-and-risks\/","title":{"rendered":"What Are AI Hallucinations in Research: Causes, Examples, and Risks"},"content":{"rendered":"<p><strong>Key Takeaways<\/strong><\/p>\n<ul>\n<li>Hallucination is inherent in generative AI. Large language models generate statistically plausible text; they do not retrieve verified facts. A fabricated citation and a real one are produced by exactly the same process.<\/li>\n<li>Citations are the highest-risk output. Peer-reviewed studies report fabrication rates from 18% to 69%, varying by model, discipline, and prompt design.<\/li>\n<li>Plausibility is the danger. The typical fabricated reference names real authors, a real journal, and a working DOI that points to an unrelated paper.<\/li>\n<li>Only humans can verify AI output. Asking a chatbot whether its own output is real is not a check; it is a second chance to hallucinate.<\/li>\n<\/ul>\n<div id=\"ez-toc-container\" class=\"ez-toc-v2_0_85 counter-hierarchy ez-toc-counter ez-toc-grey ez-toc-container-direction\">\n<div class=\"ez-toc-title-container\">\n<p class=\"ez-toc-title\" style=\"cursor:inherit\">Table of Contents<\/p>\n<span class=\"ez-toc-title-toggle\"><a href=\"#\" class=\"ez-toc-pull-right ez-toc-btn ez-toc-btn-xs ez-toc-btn-default ez-toc-toggle\" aria-label=\"Toggle Table of Content\"><span class=\"ez-toc-js-icon-con\"><span class=\"\"><span class=\"eztoc-hide\" style=\"display:none;\">Toggle<\/span><span class=\"ez-toc-icon-toggle-span\"><svg style=\"fill: #999;color:#999\" xmlns=\"http:\/\/www.w3.org\/2000\/svg\" class=\"list-377408\" width=\"20px\" height=\"20px\" viewBox=\"0 0 24 24\" fill=\"none\"><path d=\"M6 6H4v2h2V6zm14 0H8v2h12V6zM4 11h2v2H4v-2zm16 0H8v2h12v-2zM4 16h2v2H4v-2zm16 0H8v2h12v-2z\" fill=\"currentColor\"><\/path><\/svg><svg style=\"fill: #999;color:#999\" class=\"arrow-unsorted-368013\" xmlns=\"http:\/\/www.w3.org\/2000\/svg\" width=\"10px\" height=\"10px\" viewBox=\"0 0 24 24\" version=\"1.2\" baseProfile=\"tiny\"><path d=\"M18.2 9.3l-6.2-6.3-6.2 6.3c-.2.2-.3.4-.3.7s.1.5.3.7c.2.2.4.3.7.3h11c.3 0 .5-.1.7-.3.2-.2.3-.5.3-.7s-.1-.5-.3-.7zM5.8 14.7l6.2 6.3 6.2-6.3c.2-.2.3-.5.3-.7s-.1-.5-.3-.7c-.2-.2-.4-.3-.7-.3h-11c-.3 0-.5.1-.7.3-.2.2-.3.5-.3.7s.1.5.3.7z\"\/><\/svg><\/span><\/span><\/span><\/a><\/span><\/div>\n<nav><ul class='ez-toc-list ez-toc-list-level-1 ' ><li class='ez-toc-page-1 ez-toc-heading-level-2'><a class=\"ez-toc-link ez-toc-heading-1\" href=\"https:\/\/www.editage.com\/blog\/what-are-ai-hallucinations-in-research-causes-examples-and-risks\/#What_Is_an_AI_Hallucination_in_Research\" >What Is an AI Hallucination in Research?<\/a><ul class='ez-toc-list-level-3' ><li class='ez-toc-heading-level-3'><a class=\"ez-toc-link ez-toc-heading-2\" href=\"https:\/\/www.editage.com\/blog\/what-are-ai-hallucinations-in-research-causes-examples-and-risks\/#AI_hallucination_vs_factual_error_vs_bias\" >AI hallucination vs. factual error vs. bias<\/a><\/li><li class='ez-toc-page-1 ez-toc-heading-level-3'><a class=\"ez-toc-link ez-toc-heading-3\" href=\"https:\/\/www.editage.com\/blog\/what-are-ai-hallucinations-in-research-causes-examples-and-risks\/#Why_the_term_%E2%80%9Challucination%E2%80%9D_is_contested\" >Why the term &#8220;hallucination&#8221; is contested<\/a><\/li><li class='ez-toc-page-1 ez-toc-heading-level-3'><a class=\"ez-toc-link ez-toc-heading-4\" href=\"https:\/\/www.editage.com\/blog\/what-are-ai-hallucinations-in-research-causes-examples-and-risks\/#Where_hallucinations_enter_the_research_workflow\" >Where hallucinations enter the research workflow<\/a><\/li><li class='ez-toc-page-1 ez-toc-heading-level-3'><a class=\"ez-toc-link ez-toc-heading-5\" href=\"https:\/\/www.editage.com\/blog\/what-are-ai-hallucinations-in-research-causes-examples-and-risks\/#Sample_prompt_scoping_which_is_low_risk\" >Sample prompt: scoping, which is low risk<\/a><\/li><\/ul><\/li><li class='ez-toc-page-1 ez-toc-heading-level-2'><a class=\"ez-toc-link ez-toc-heading-6\" href=\"https:\/\/www.editage.com\/blog\/what-are-ai-hallucinations-in-research-causes-examples-and-risks\/#Why_Do_AI_Hallucinate_The_Core_Causes\" >Why Do AI Hallucinate? The Core Causes<\/a><ul class='ez-toc-list-level-3' ><li class='ez-toc-heading-level-3'><a class=\"ez-toc-link ez-toc-heading-7\" href=\"https:\/\/www.editage.com\/blog\/what-are-ai-hallucinations-in-research-causes-examples-and-risks\/#Why_do_AI_hallucinate_when_asked_for_citations_specifically\" >Why do AI hallucinate when asked for citations specifically?<\/a><\/li><li class='ez-toc-page-1 ez-toc-heading-level-3'><a class=\"ez-toc-link ez-toc-heading-8\" href=\"https:\/\/www.editage.com\/blog\/what-are-ai-hallucinations-in-research-causes-examples-and-risks\/#Prediction_optimizes_for_plausibility_not_truth\" >Prediction optimizes for plausibility, not truth<\/a><\/li><li class='ez-toc-page-1 ez-toc-heading-level-3'><a class=\"ez-toc-link ez-toc-heading-9\" href=\"https:\/\/www.editage.com\/blog\/what-are-ai-hallucinations-in-research-causes-examples-and-risks\/#Training_data_gaps_staleness_and_long-tail_topics\" >Training data gaps, staleness, and long-tail topics<\/a><\/li><li class='ez-toc-page-1 ez-toc-heading-level-3'><a class=\"ez-toc-link ez-toc-heading-10\" href=\"https:\/\/www.editage.com\/blog\/what-are-ai-hallucinations-in-research-causes-examples-and-risks\/#Evaluation_and_training_that_reward_guessing_over_abstention\" >Evaluation and training that reward guessing over abstention<\/a><\/li><li class='ez-toc-page-1 ez-toc-heading-level-3'><a class=\"ez-toc-link ez-toc-heading-11\" href=\"https:\/\/www.editage.com\/blog\/what-are-ai-hallucinations-in-research-causes-examples-and-risks\/#Ambiguous_prompts_and_leading_questions\" >Ambiguous prompts and leading questions<\/a><\/li><\/ul><\/li><li class='ez-toc-page-1 ez-toc-heading-level-2'><a class=\"ez-toc-link ez-toc-heading-12\" href=\"https:\/\/www.editage.com\/blog\/what-are-ai-hallucinations-in-research-causes-examples-and-risks\/#How_Hallucination_in_AI_Models_Works_in_Practice\" >How Hallucination in AI Models Works in Practice<\/a><ul class='ez-toc-list-level-3' ><li class='ez-toc-heading-level-3'><a class=\"ez-toc-link ez-toc-heading-13\" href=\"https:\/\/www.editage.com\/blog\/what-are-ai-hallucinations-in-research-causes-examples-and-risks\/#Hallucination_in_AI_models_intrinsic_vs_extrinsic\" >Hallucination in AI models: intrinsic vs. extrinsic<\/a><\/li><li class='ez-toc-page-1 ez-toc-heading-level-3'><a class=\"ez-toc-link ez-toc-heading-14\" href=\"https:\/\/www.editage.com\/blog\/what-are-ai-hallucinations-in-research-causes-examples-and-risks\/#Does_retrieval-augmented_generation_RAG_reduce_AI_hallucinations\" >Does retrieval-augmented generation (RAG) reduce AI hallucinations?<\/a><\/li><li class='ez-toc-page-1 ez-toc-heading-level-3'><a class=\"ez-toc-link ez-toc-heading-15\" href=\"https:\/\/www.editage.com\/blog\/what-are-ai-hallucinations-in-research-causes-examples-and-risks\/#Settings_and_conditions_that_change_the_risk_profile\" >Settings and conditions that change the risk profile<\/a><\/li><\/ul><\/li><li class='ez-toc-page-1 ez-toc-heading-level-2'><a class=\"ez-toc-link ez-toc-heading-16\" href=\"https:\/\/www.editage.com\/blog\/what-are-ai-hallucinations-in-research-causes-examples-and-risks\/#Examples_of_AI_Hallucination_Across_Subject_Areas\" >Examples of AI Hallucination Across Subject Areas<\/a><ul class='ez-toc-list-level-3' ><li class='ez-toc-heading-level-3'><a class=\"ez-toc-link ez-toc-heading-17\" href=\"https:\/\/www.editage.com\/blog\/what-are-ai-hallucinations-in-research-causes-examples-and-risks\/#Examples_of_AI_hallucination_in_law_the_Mata_v_Avianca_sanction\" >Examples of AI hallucination in law: the Mata v. Avianca sanction<\/a><\/li><li class='ez-toc-page-1 ez-toc-heading-level-3'><a class=\"ez-toc-link ez-toc-heading-18\" href=\"https:\/\/www.editage.com\/blog\/what-are-ai-hallucinations-in-research-causes-examples-and-risks\/#Medicine_and_biomedical_research\" >Medicine and biomedical research<\/a><\/li><li class='ez-toc-page-1 ez-toc-heading-level-3'><a class=\"ez-toc-link ez-toc-heading-19\" href=\"https:\/\/www.editage.com\/blog\/what-are-ai-hallucinations-in-research-causes-examples-and-risks\/#Psychology_and_mental_health_research\" >Psychology and mental health research<\/a><\/li><li class='ez-toc-page-1 ez-toc-heading-level-3'><a class=\"ez-toc-link ez-toc-heading-20\" href=\"https:\/\/www.editage.com\/blog\/what-are-ai-hallucinations-in-research-causes-examples-and-risks\/#Economics_statistics_and_quantitative_social_science\" >Economics, statistics, and quantitative social science<\/a><\/li><li class='ez-toc-page-1 ez-toc-heading-level-3'><a class=\"ez-toc-link ez-toc-heading-21\" href=\"https:\/\/www.editage.com\/blog\/what-are-ai-hallucinations-in-research-causes-examples-and-risks\/#Sample_prompt_extracting_numbers_with_a_hallucination_trap_built_in\" >Sample prompt: extracting numbers with a hallucination trap built in<\/a><\/li><li class='ez-toc-page-1 ez-toc-heading-level-3'><a class=\"ez-toc-link ez-toc-heading-22\" href=\"https:\/\/www.editage.com\/blog\/what-are-ai-hallucinations-in-research-causes-examples-and-risks\/#Computer_science_and_software_engineering\" >Computer science and software engineering<\/a><\/li><li class='ez-toc-page-1 ez-toc-heading-level-3'><a class=\"ez-toc-link ez-toc-heading-23\" href=\"https:\/\/www.editage.com\/blog\/what-are-ai-hallucinations-in-research-causes-examples-and-risks\/#History_and_the_humanities\" >History and the humanities<\/a><\/li><\/ul><\/li><li class='ez-toc-page-1 ez-toc-heading-level-2'><a class=\"ez-toc-link ez-toc-heading-24\" href=\"https:\/\/www.editage.com\/blog\/what-are-ai-hallucinations-in-research-causes-examples-and-risks\/#ChatGPT_Hallucination_Rates_What_the_Studies_Actually_Measure\" >ChatGPT Hallucination Rates: What the Studies Actually Measure<\/a><ul class='ez-toc-list-level-3' ><li class='ez-toc-heading-level-3'><a class=\"ez-toc-link ez-toc-heading-25\" href=\"https:\/\/www.editage.com\/blog\/what-are-ai-hallucinations-in-research-causes-examples-and-risks\/#ChatGPT_hallucination_rates_reported_in_citation_studies\" >ChatGPT hallucination rates reported in citation studies<\/a><\/li><li class='ez-toc-page-1 ez-toc-heading-level-3'><a class=\"ez-toc-link ez-toc-heading-26\" href=\"https:\/\/www.editage.com\/blog\/what-are-ai-hallucinations-in-research-causes-examples-and-risks\/#Why_published_rates_will_not_predict_your_rate\" >Why published rates will not predict your rate<\/a><\/li><li class='ez-toc-page-1 ez-toc-heading-level-3'><a class=\"ez-toc-link ez-toc-heading-27\" href=\"https:\/\/www.editage.com\/blog\/what-are-ai-hallucinations-in-research-causes-examples-and-risks\/#Sample_prompt_measure_your_own_rate_before_trusting_a_workflow\" >Sample prompt: measure your own rate before trusting a workflow<\/a><\/li><\/ul><\/li><li class='ez-toc-page-1 ez-toc-heading-level-2'><a class=\"ez-toc-link ez-toc-heading-28\" href=\"https:\/\/www.editage.com\/blog\/what-are-ai-hallucinations-in-research-causes-examples-and-risks\/#The_Risks_of_AI_Hallucination_for_Researchers_and_Institutions\" >The Risks of AI Hallucination for Researchers and Institutions<\/a><ul class='ez-toc-list-level-3' ><li class='ez-toc-heading-level-3'><a class=\"ez-toc-link ez-toc-heading-29\" href=\"https:\/\/www.editage.com\/blog\/what-are-ai-hallucinations-in-research-causes-examples-and-risks\/#Personal_and_professional_risk\" >Personal and professional risk<\/a><\/li><li class='ez-toc-page-1 ez-toc-heading-level-3'><a class=\"ez-toc-link ez-toc-heading-30\" href=\"https:\/\/www.editage.com\/blog\/what-are-ai-hallucinations-in-research-causes-examples-and-risks\/#Contamination_of_the_scholarly_record\" >Contamination of the scholarly record<\/a><\/li><li class='ez-toc-page-1 ez-toc-heading-level-3'><a class=\"ez-toc-link ez-toc-heading-31\" href=\"https:\/\/www.editage.com\/blog\/what-are-ai-hallucinations-in-research-causes-examples-and-risks\/#Legal_regulatory_and_research-integrity_exposure\" >Legal, regulatory, and research-integrity exposure<\/a><\/li><li class='ez-toc-page-1 ez-toc-heading-level-3'><a class=\"ez-toc-link ez-toc-heading-32\" href=\"https:\/\/www.editage.com\/blog\/what-are-ai-hallucinations-in-research-causes-examples-and-risks\/#Automation_bias_the_reason_plausible_errors_survive\" >Automation bias: the reason plausible errors survive<\/a><\/li><\/ul><\/li><li class='ez-toc-page-1 ez-toc-heading-level-2'><a class=\"ez-toc-link ez-toc-heading-33\" href=\"https:\/\/www.editage.com\/blog\/what-are-ai-hallucinations-in-research-causes-examples-and-risks\/#How_to_Detect_and_Reduce_Hallucinations_in_Your_Workflow\" >How to Detect and Reduce Hallucinations in Your Workflow<\/a><ul class='ez-toc-list-level-3' ><li class='ez-toc-heading-level-3'><a class=\"ez-toc-link ez-toc-heading-34\" href=\"https:\/\/www.editage.com\/blog\/what-are-ai-hallucinations-in-research-causes-examples-and-risks\/#A_verification_checklist_for_every_AI-assisted_claim\" >A verification checklist for every AI-assisted claim<\/a><\/li><li class='ez-toc-page-1 ez-toc-heading-level-3'><a class=\"ez-toc-link ez-toc-heading-35\" href=\"https:\/\/www.editage.com\/blog\/what-are-ai-hallucinations-in-research-causes-examples-and-risks\/#Prompting_practices_that_measurably_lower_risk\" >Prompting practices that measurably lower risk<\/a><\/li><li class='ez-toc-page-1 ez-toc-heading-level-3'><a class=\"ez-toc-link ez-toc-heading-36\" href=\"https:\/\/www.editage.com\/blog\/what-are-ai-hallucinations-in-research-causes-examples-and-risks\/#Tools_that_help_and_where_each_stops_helping\" >Tools that help, and where each stops helping<\/a><\/li><li class='ez-toc-page-1 ez-toc-heading-level-3'><a class=\"ez-toc-link ez-toc-heading-37\" href=\"https:\/\/www.editage.com\/blog\/what-are-ai-hallucinations-in-research-causes-examples-and-risks\/#Disclosure_obligations\" >Disclosure obligations<\/a><\/li><\/ul><\/li><li class='ez-toc-page-1 ez-toc-heading-level-2'><a class=\"ez-toc-link ez-toc-heading-38\" href=\"https:\/\/www.editage.com\/blog\/what-are-ai-hallucinations-in-research-causes-examples-and-risks\/#Frequently_Asked_Questions\" >Frequently Asked Questions<\/a><ul class='ez-toc-list-level-3' ><li class='ez-toc-heading-level-3'><a class=\"ez-toc-link ez-toc-heading-39\" href=\"https:\/\/www.editage.com\/blog\/what-are-ai-hallucinations-in-research-causes-examples-and-risks\/#How_do_I_check_if_an_AI-generated_citation_is_real\" >How do I check if an AI-generated citation is real?<\/a><\/li><li class='ez-toc-page-1 ez-toc-heading-level-3'><a class=\"ez-toc-link ez-toc-heading-40\" href=\"https:\/\/www.editage.com\/blog\/what-are-ai-hallucinations-in-research-causes-examples-and-risks\/#Why_does_ChatGPT_make_up_DOIs_that_link_to_real_papers\" >Why does ChatGPT make up DOIs that link to real papers?<\/a><\/li><li class='ez-toc-page-1 ez-toc-heading-level-3'><a class=\"ez-toc-link ez-toc-heading-41\" href=\"https:\/\/www.editage.com\/blog\/what-are-ai-hallucinations-in-research-causes-examples-and-risks\/#Can_I_use_ChatGPT_for_a_literature_review\" >Can I use ChatGPT for a literature review?<\/a><\/li><li class='ez-toc-page-1 ez-toc-heading-level-3'><a class=\"ez-toc-link ez-toc-heading-42\" href=\"https:\/\/www.editage.com\/blog\/what-are-ai-hallucinations-in-research-causes-examples-and-risks\/#Do_newer_AI_models_still_hallucinate\" >Do newer AI models still hallucinate?<\/a><\/li><li class='ez-toc-page-1 ez-toc-heading-level-3'><a class=\"ez-toc-link ez-toc-heading-43\" href=\"https:\/\/www.editage.com\/blog\/what-are-ai-hallucinations-in-research-causes-examples-and-risks\/#What_is_the_difference_between_an_AI_hallucination_and_a_lie\" >What is the difference between an AI hallucination and a lie?<\/a><\/li><li class='ez-toc-page-1 ez-toc-heading-level-3'><a class=\"ez-toc-link ez-toc-heading-44\" href=\"https:\/\/www.editage.com\/blog\/what-are-ai-hallucinations-in-research-causes-examples-and-risks\/#Do_I_need_to_disclose_AI_use_when_submitting_to_a_journal\" >Do I need to disclose AI use when submitting to a journal?<\/a><\/li><li class='ez-toc-page-1 ez-toc-heading-level-3'><a class=\"ez-toc-link ez-toc-heading-45\" href=\"https:\/\/www.editage.com\/blog\/what-are-ai-hallucinations-in-research-causes-examples-and-risks\/#Can_AI_hallucinations_be_eliminated_completely\" >Can AI hallucinations be eliminated completely?<\/a><\/li><li class='ez-toc-page-1 ez-toc-heading-level-3'><a class=\"ez-toc-link ez-toc-heading-46\" href=\"https:\/\/www.editage.com\/blog\/what-are-ai-hallucinations-in-research-causes-examples-and-risks\/#Which_AI_tool_hallucinates_the_least_for_academic_research\" >Which AI tool hallucinates the least for academic research?<\/a><\/li><\/ul><\/li><\/ul><\/nav><\/div>\n<h2><span class=\"ez-toc-section\" id=\"What_Is_an_AI_Hallucination_in_Research\"><\/span>What Is an AI Hallucination in Research?<span class=\"ez-toc-section-end\"><\/span><\/h2>\n<p>An AI hallucination is output that reads as authoritative but has no basis in the model&#8217;s training data, in the sources you supplied, or in reality. In a research context, this rarely looks like nonsense. It looks like a citation, a statistic, a quotation, or a summary that is formatted correctly and fits the argument perfectly.<\/p>\n<p>The critical point is mechanical: a language model does not distinguish between generating a true sentence and generating a plausible one. Both come out of the same process. Fluency is not evidence of accuracy.<\/p>\n<h3><span class=\"ez-toc-section\" id=\"AI_hallucination_vs_factual_error_vs_bias\"><\/span>AI hallucination vs. factual error vs. bias<span class=\"ez-toc-section-end\"><\/span><\/h3>\n<p>The following 4 issues get conflated, but they need different fixes. Mislabeling a hallucination as &#8220;bias&#8221; or &#8220;just an error&#8221; leads researchers to the wrong remedy.<\/p>\n<table>\n<thead>\n<tr>\n<td><strong>Failure mode<\/strong><\/td>\n<td><strong>What it looks like<\/strong><\/td>\n<td><strong>What actually fixes it<\/strong><\/td>\n<\/tr>\n<\/thead>\n<tbody>\n<tr>\n<td>Hallucination<\/td>\n<td>A confident claim, citation, or quote with no real referent anywhere.<\/td>\n<td>External verification against primary sources.<\/td>\n<\/tr>\n<tr>\n<td>Factual error<\/td>\n<td>A real entity described incorrectly: wrong date, wrong figure, wrong attribution.<\/td>\n<td>Fact-checking; sometimes better retrieval.<\/td>\n<\/tr>\n<tr>\n<td>Bias<\/td>\n<td>Systematic skew in framing, coverage, or which perspectives get represented.<\/td>\n<td>Diverse sourcing; explicit counter-prompting.<\/td>\n<\/tr>\n<tr>\n<td>Sycophancy<\/td>\n<td>The model agreeing with your stated <a href=\"https:\/\/www.editage.com\/blog\/how-to-write-a-research-hypothesis-examples-formulation-thesis-research-papers\/\">hypothesis<\/a> regardless of evidence.<\/td>\n<td>Neutral prompt phrasing; asking for the opposing case.<\/td>\n<\/tr>\n<\/tbody>\n<\/table>\n<h3><span class=\"ez-toc-section\" id=\"Why_the_term_%E2%80%9Challucination%E2%80%9D_is_contested\"><\/span>Why the term &#8220;hallucination&#8221; is contested<span class=\"ez-toc-section-end\"><\/span><\/h3>\n<p>Many researchers object to the word, and the objection is worth understanding before you use it in a paper.<\/p>\n<ul>\n<li>It implies perception. The model is not misperceiving a real world; it has no perceptual access to one at all.<\/li>\n<li>It implies deviation from normal function. Generating plausible text is normal function; accuracy is the incidental outcome when the plausible text happens to be true.<\/li>\n<li>It borrows a clinical term for a computational phenomenon, which some argue trivializes psychiatric symptoms.<\/li>\n<li>Proposed alternatives include &#8220;confabulation,&#8221; &#8220;fabrication,&#8221; and simply &#8220;bullshit&#8221; in the technical, Frankfurtian sense of indifference to truth.<\/li>\n<\/ul>\n<p>&#8220;Hallucination&#8221; remains the dominant term in the literature, so this article uses it. Just be aware that it names a symptom, not a malfunction.<\/p>\n<h3><span class=\"ez-toc-section\" id=\"Where_hallucinations_enter_the_research_workflow\"><\/span>Where hallucinations enter the research workflow<span class=\"ez-toc-section-end\"><\/span><\/h3>\n<p>Risk is not evenly distributed across tasks. It concentrates wherever the model must produce specific, verifiable particulars that it cannot look up.<\/p>\n<table>\n<thead>\n<tr>\n<td><strong>Workflow stage<\/strong><\/td>\n<td><strong>Typical task<\/strong><\/td>\n<td><strong>Hallucination risk<\/strong><\/td>\n<\/tr>\n<\/thead>\n<tbody>\n<tr>\n<td>Topic scoping<\/td>\n<td>Explaining a concept, mapping a debate<\/td>\n<td>Low to moderate: broad claims are usually well-represented in training data.<\/td>\n<\/tr>\n<tr>\n<td><a href=\"https:\/\/www.editage.com\/blog\/ai-literature-search-how-to-use-ai-for-searching-evaluating-and-synthesizing-research\/\">Literature search<\/a><\/td>\n<td>Asking for relevant papers on a topic<\/td>\n<td>Very high: this is the single most dangerous use.<\/td>\n<\/tr>\n<tr>\n<td>Summarizing a supplied paper<\/td>\n<td>Condensing a PDF you uploaded<\/td>\n<td>Moderate: the source is real, but conclusions get overstated or reversed.<\/td>\n<\/tr>\n<tr>\n<td><a href=\"https:\/\/www.editage.com\/blog\/ai-literature-search-how-to-use-ai-for-searching-evaluating-and-synthesizing-research\/#Prompting_for_Structured_Extraction\">Data extraction<\/a><\/td>\n<td>Pulling figures from tables or transcripts<\/td>\n<td>High: numbers get transposed, averaged, or invented to fill gaps.<\/td>\n<\/tr>\n<tr>\n<td>Statistical interpretation<\/td>\n<td>Explaining what a result means<\/td>\n<td>High: models invent plausible p values and effect sizes.<\/td>\n<\/tr>\n<tr>\n<td><a href=\"https:\/\/www.editage.com\/blog\/how-to-use-ai-to-write-a-research-paper-tools-prompts-and-workflow-to-prevent-hallucination-and-sound-human\/\">Drafting and editing<\/a><\/td>\n<td>Improving prose you wrote<\/td>\n<td>Low if you have written the bulk of the text, high if you ask the AI tool to create sentences, paragraphs, or sections<\/td>\n<\/tr>\n<tr>\n<td>Citation formatting<\/td>\n<td>Converting an existing reference to APA, MLA, etc.<\/td>\n<td>Low if the reference is real; the model may still &#8220;correct&#8221; real details.<\/td>\n<\/tr>\n<\/tbody>\n<\/table>\n<h3><span class=\"ez-toc-section\" id=\"Sample_prompt_scoping_which_is_low_risk\"><\/span>Sample prompt: scoping, which is low risk<span class=\"ez-toc-section-end\"><\/span><\/h3>\n<p>Explain the main theoretical disagreements between the capability approach and standard welfare economics. Do not cite specific papers. I will find sources myself.<\/p>\n<p>This works because it asks for conceptual structure, which models handle well, and it explicitly blocks the output type that fails most often.<\/p>\n<p>&nbsp;<\/p>\n<h2><span class=\"ez-toc-section\" id=\"Why_Do_AI_Hallucinate_The_Core_Causes\"><\/span>Why Do AI Hallucinate? The Core Causes<span class=\"ez-toc-section-end\"><\/span><\/h2>\n<p>There is no single cause. Hallucination emerges from at least 5 separate pressures, which is why no single fix eliminates it. Understanding which pressure applies to your task tells you which mitigation will help.<\/p>\n<h3><span class=\"ez-toc-section\" id=\"Why_do_AI_hallucinate_when_asked_for_citations_specifically\"><\/span>Why do AI hallucinate when asked for citations specifically?<span class=\"ez-toc-section-end\"><\/span><\/h3>\n<p>Citations are the worst case for a text predictor, and the reason is structural rather than accidental.<\/p>\n<ul>\n<li><strong>The format is extremely regular.<\/strong> Author, year, title, journal, volume, pages. A model learns this shape perfectly from millions of examples, so it can generate flawless-looking references indefinitely.<\/li>\n<li><strong>The content is arbitrary.<\/strong> There is no pattern connecting a topic to the exact page numbers of a specific 2019 article. That mapping must be memorized, and most of it never was.<\/li>\n<li><strong>Individual references are long-tail.<\/strong> Any given paper appears rarely in training data. The format appears constantly. The model has learned the container far better than the contents.<\/li>\n<li><strong>Numeric fields fail first.<\/strong> In one study of 636 citations, volume, issue, page, and year errors were the most common defects even in otherwise-real references.<\/li>\n<li><strong>Book chapters are the worst category.<\/strong> In the same study, 70% of GPT-4&#8217;s cited book chapters were fabricated, and often the containing book did not exist either.<\/li>\n<\/ul>\n<h4>Sample prompts: the risky ask and the safer ask<\/h4>\n<h5>RISKY<\/h5>\n<p>Write a literature review on gut microbiome and depression with 10 peer-reviewed citations.<\/p>\n<h5>SAFER<\/h5>\n<p>I have attached 8 PDFs. Write a literature review on the link between gut microbiota alterations and antidepressant use, using only these papers. Cite by the filename and page number. If a claim I need is not supported by these 8 papers, say &#8220;not covered by supplied sources&#8221; instead of citing anything else.<\/p>\n<h3><span class=\"ez-toc-section\" id=\"Prediction_optimizes_for_plausibility_not_truth\"><\/span>Prediction optimizes for plausibility, not truth<span class=\"ez-toc-section-end\"><\/span><\/h3>\n<p>A language model is trained to make text likely, not to make it correct. When the training objective is &#8220;predict what comes next,&#8221; a false-but-typical continuation scores as well as a true one, provided both look like the kind of thing that gets written.<\/p>\n<p>Researchers at OpenAI formalized this in 2025, showing that hallucinations arise as ordinary statistical errors during pretraining rather than as a mysterious glitch. When a fact cannot be reliably distinguished from a plausible non-fact given the training signal, some rate of fabrication is mathematically expected.<\/p>\n<h3><span class=\"ez-toc-section\" id=\"Training_data_gaps_staleness_and_long-tail_topics\"><\/span>Training data gaps, staleness, and long-tail topics<span class=\"ez-toc-section-end\"><\/span><\/h3>\n<ul>\n<li>Every model has a training cutoff. Anything after it is unknown, and the model may generate confident content about it anyway.<\/li>\n<li>Niche subfields, non-English scholarship, gray literature, and paywalled corpora are thinly represented, so the model interpolates.<\/li>\n<li>Recently retracted or superseded findings often remain in the model&#8217;s training data and get repeated as current.<\/li>\n<li>Contradictory sources in training data produce blended outputs that match no actual source.<\/li>\n<li>Rare proper nouns are especially fragile: obscure author names, small journals, and regional institutions get merged or invented.<\/li>\n<\/ul>\n<h3><span class=\"ez-toc-section\" id=\"Evaluation_and_training_that_reward_guessing_over_abstention\"><\/span>Evaluation and training that reward guessing over abstention<span class=\"ez-toc-section-end\"><\/span><\/h3>\n<p>This cause is the least intuitive and arguably the most important. Most benchmarks score a wrong answer and an &#8220;I don&#8217;t know&#8221; identically, at zero. Under that scoring, guessing is always the optimal strategy, exactly as it is for a student on a multiple-choice exam with no penalty for wrong answers.<\/p>\n<p>Models are therefore trained and selected to behave like confident test-takers. The proposed fix is to change the scoring so that incorrect answers are penalized more than abstentions, which would make calibrated uncertainty the winning strategy. Until benchmarks change, confident guessing remains rewarded.<\/p>\n<p>There is a second, related pressure. Reinforcement learning from human feedback trains on human ratings, and humans reliably prefer confident answers to hedged ones. Confidence gets reinforced whether or not it is warranted.<\/p>\n<h3><span class=\"ez-toc-section\" id=\"Ambiguous_prompts_and_leading_questions\"><\/span>Ambiguous prompts and leading questions<span class=\"ez-toc-section-end\"><\/span><\/h3>\n<p>Users generate a large share of their own hallucinations. A prompt that presupposes a fact will usually get that fact confirmed, complete with supporting evidence that does not exist.<\/p>\n<h4>Sample prompts: leading vs. neutral<\/h4>\n<h5>LEADING, PRESUPPOSES THE FINDING<\/h5>\n<p>Summarize the studies showing that remote work reduces employee productivity.<\/p>\n<h5>NEUTRAL, ALLOWS A NULL RESULT<\/h5>\n<p>What does the empirical literature find about remote work and employee productivity? Report findings in both directions, note where evidence is mixed or weak, and state explicitly if you are uncertain whether specific studies exist.<\/p>\n<h4>Tips for framing your own prompts<\/h4>\n<ul>\n<li>Do not name a conclusion you want supported.<\/li>\n<li>Do not ask &#8220;find studies proving X.&#8221; Ask &#8220;what does the evidence say about X.&#8221;<\/li>\n<li>Avoid asking for a fixed number of sources. &#8220;Give me 10 citations&#8221; pressures the model to fill quota with inventions.<\/li>\n<li>Do not push back repeatedly until the model changes its answer; that trains the conversation, not the truth.<\/li>\n<\/ul>\n<p>&nbsp;<\/p>\n<h2><span class=\"ez-toc-section\" id=\"How_Hallucination_in_AI_Models_Works_in_Practice\"><\/span>How Hallucination in AI Models Works in Practice<span class=\"ez-toc-section-end\"><\/span><\/h2>\n<p>Knowing the mechanism helps you predict where a given tool will fail, which is more useful than a general warning to &#8220;be careful.&#8221;<\/p>\n<h3><span class=\"ez-toc-section\" id=\"Hallucination_in_AI_models_intrinsic_vs_extrinsic\"><\/span>Hallucination in AI models: intrinsic vs. extrinsic<span class=\"ez-toc-section-end\"><\/span><\/h3>\n<p>The standard taxonomy splits hallucinations by their relationship to a provided source. This distinction determines whether grounding will help you.<\/p>\n<table>\n<thead>\n<tr>\n<td><strong>Type<\/strong><\/td>\n<td><strong>Definition<\/strong><\/td>\n<td><strong>Example<\/strong><\/td>\n<\/tr>\n<\/thead>\n<tbody>\n<tr>\n<td>Intrinsic<\/td>\n<td>The output contradicts a source that was supplied to the model.<\/td>\n<td>You upload a paper reporting no significant effect; the summary reports a significant effect.<\/td>\n<\/tr>\n<tr>\n<td>Extrinsic<\/td>\n<td>The output makes claims that cannot be checked against any supplied source.<\/td>\n<td>The model adds a citation to a 2021 meta-analysis that was never in your document set.<\/td>\n<\/tr>\n<tr>\n<td>Grounding helps?<\/td>\n<td>Partially: contradictions become detectable by comparison.<\/td>\n<td>Yes for detection, no for prevention: the model can still add unsupported material.<\/td>\n<\/tr>\n<\/tbody>\n<\/table>\n<h3><span class=\"ez-toc-section\" id=\"Does_retrieval-augmented_generation_RAG_reduce_AI_hallucinations\"><\/span>Does retrieval-augmented generation (RAG) reduce AI hallucinations?<span class=\"ez-toc-section-end\"><\/span><\/h3>\n<p>RAG reduces hallucination substantially. It does not remove it, and researchers who assume otherwise get caught by a narrower but subtler set of failures.<\/p>\n<ul>\n<li><strong>Retrieval can miss.<\/strong> If the search step returns nothing relevant, most systems answer anyway from parametric memory rather than declining.<\/li>\n<li><strong>Retrieval can return the wrong passage.<\/strong> Semantic search matches topic, not claim. A passage about the same subject can support the opposite conclusion.<\/li>\n<li><strong>Synthesis still hallucinates.<\/strong> Combining 5 real passages into one summary is a generative step, and the connective tissue between them is invented.<\/li>\n<li><strong>Citations can be misattached.<\/strong> The tool cites a real, retrieved document for a sentence that document does not actually support. This is the hardest failure to catch, because the link works.<\/li>\n<li><strong>The corpus itself may be contaminated.<\/strong> Preprint servers and repositories now contain AI-generated text with fabricated references, which retrieval will faithfully surface.<\/li>\n<\/ul>\n<h3><span class=\"ez-toc-section\" id=\"Settings_and_conditions_that_change_the_risk_profile\"><\/span>Settings and conditions that change the risk profile<span class=\"ez-toc-section-end\"><\/span><\/h3>\n<table>\n<thead>\n<tr>\n<td><strong>Condition<\/strong><\/td>\n<td><strong>Effect on hallucination risk<\/strong><\/td>\n<td><strong>What to do<\/strong><\/td>\n<\/tr>\n<\/thead>\n<tbody>\n<tr>\n<td>High temperature<\/td>\n<td>Increases variability and invention.<\/td>\n<td>Use the lowest setting available for factual work.<\/td>\n<\/tr>\n<tr>\n<td>Long context or long conversation<\/td>\n<td>Earlier instructions and sources lose influence.<\/td>\n<td>Start a fresh session per document or per question.<\/td>\n<\/tr>\n<tr>\n<td>Obscure or non-English topic<\/td>\n<td>Sharply increases risk.<\/td>\n<td>Supply sources yourself; do not rely on recall.<\/td>\n<\/tr>\n<tr>\n<td>Reasoning or extended-thinking mode<\/td>\n<td>Generally reduces factual errors.<\/td>\n<td>Enable it for anything analytic.<\/td>\n<\/tr>\n<tr>\n<td>Web search enabled<\/td>\n<td>Reduces fabrication of sources; does not eliminate misreading.<\/td>\n<td>Still open every link and confirm the claim.<\/td>\n<\/tr>\n<tr>\n<td>Requests for a fixed quantity<\/td>\n<td>Pressures the model to fill quota.<\/td>\n<td>Ask for &#8220;however many are well-supported.&#8221;<\/td>\n<\/tr>\n<\/tbody>\n<\/table>\n<p>&nbsp;<\/p>\n<h2><span class=\"ez-toc-section\" id=\"Examples_of_AI_Hallucination_Across_Subject_Areas\"><\/span>Examples of AI Hallucination Across Subject Areas<span class=\"ez-toc-section-end\"><\/span><\/h2>\n<p>Hallucination is not a quirk of one discipline. What changes across fields is the form the fabrication takes and who gets hurt when it survives review.<\/p>\n<h3><span class=\"ez-toc-section\" id=\"Examples_of_AI_hallucination_in_law_the_Mata_v_Avianca_sanction\"><\/span>Examples of AI hallucination in law: the Mata v. Avianca sanction<span class=\"ez-toc-section-end\"><\/span><\/h3>\n<p>This is the reference case. In a personal injury suit against the airline Avianca, an attorney used ChatGPT to research a brief and submitted 6 court decisions that did not exist. Judge P. Kevin Castel of the Southern District of New York sanctioned the attorneys and their firm $5,000 in June 2023.<\/p>\n<p>The instructive detail is not that the model invented cases. It is what happened next.<\/p>\n<ul>\n<li>The fabricated opinions named real, sitting federal judges as their authors, and cited docket numbers belonging to unrelated real cases.<\/li>\n<li>Each fake opinion contained internal citations to further cases that also did not exist.<\/li>\n<li>When the attorney asked ChatGPT whether the cases were real, it confirmed that they were and named Westlaw and LexisNexis as places to find them. That confirmation was itself a hallucination.<\/li>\n<li>The court was explicit that using AI is not improper. Failing to verify is.<\/li>\n<\/ul>\n<p>The pattern has repeated many times since across multiple jurisdictions, and later courts have treated the passage of time since this ruling as removing any excuse of novelty.<\/p>\n<h3><span class=\"ez-toc-section\" id=\"Medicine_and_biomedical_research\"><\/span>Medicine and biomedical research<span class=\"ez-toc-section-end\"><\/span><\/h3>\n<p>Medicine has the largest body of measured evidence, partly because reference accuracy is verifiable through PubMed.<\/p>\n<ul>\n<li><a href=\"https:\/\/pmc.ncbi.nlm.nih.gov\/articles\/PMC10277170\/\">An observational study<\/a> of 30 short medical papers generated by ChatGPT found that of 115 generated references, 47% were fabricated, 46% were real but inaccurate, and only 7% were both real and accurate.<\/li>\n<li>In the same study, an incorrect PubMed ID appeared in 93% of papers, making the identifier itself an unreliable signal of authenticity.<\/li>\n<li><a href=\"https:\/\/www.mcpdigitalhealth.org\/article\/S2949-7612(23)00036-6\/fulltext\">A study of ChatGPT responses to medical questions<\/a> found 41 of 59 references, or 69%, were fabricated while appearing entirely credible.<\/li>\n<li><a href=\"https:\/\/pmc.ncbi.nlm.nih.gov\/articles\/PMC10484980\/\">Another multidisciplinary study<\/a> found 64% of 343 citations generated by ChatGPT-3.5 and ChatGPT-4 could not be located in PubMed or on the open web.<\/li>\n<li>Fabricated medical references typically pair real, topically appropriate authors with a real journal and an invented title, which defeats casual plausibility checks.<\/li>\n<\/ul>\n<h3><span class=\"ez-toc-section\" id=\"Psychology_and_mental_health_research\"><\/span>Psychology and mental health research<span class=\"ez-toc-section-end\"><\/span><\/h3>\n<p><a href=\"https:\/\/mental.jmir.org\/2025\/1\/e80371\">Linardon et al. (2025)<\/a> asked a model to produce 6 literature reviews on mental health topics and audited all 176 resulting citations.<\/p>\n<table>\n<thead>\n<tr>\n<td><strong>Outcome<\/strong><\/td>\n<td><strong>Share of 176 citations<\/strong><\/td>\n<td><strong>Note<\/strong><\/td>\n<\/tr>\n<\/thead>\n<tbody>\n<tr>\n<td>Real and accurate<\/td>\n<td>About 44%<\/td>\n<td>Fewer than half were usable as given.<\/td>\n<\/tr>\n<tr>\n<td>Contained errors<\/td>\n<td>About 45%<\/td>\n<td>Real work, wrong details.<\/td>\n<\/tr>\n<tr>\n<td>Fully fabricated<\/td>\n<td>About 20%<\/td>\n<td>No such work exists.<\/td>\n<\/tr>\n<tr>\n<td>Fabricated but with a DOI<\/td>\n<td>Over 94% of fabrications<\/td>\n<td>The identifier looked legitimate.<\/td>\n<\/tr>\n<tr>\n<td>Fake DOI resolving to a real, unrelated paper<\/td>\n<td>About 64% of fake DOIs<\/td>\n<td>Clicking through was the only way to catch it.<\/td>\n<\/tr>\n<\/tbody>\n<\/table>\n<p>That last row is the finding researchers should internalize. A DOI that resolves is not proof of anything. You have to read what it resolves to.<\/p>\n<p>Another important finding was that <a href=\"https:\/\/www.editage.com\/blog\/ai-accuracy-in-research-writing-how-reliable-is-ai-for-researchers\/\">AI accuracy<\/a> was higher on well-known, well-researched topics, compared to niche, emerging, or under-researched topics.<\/p>\n<h3><span class=\"ez-toc-section\" id=\"Economics_statistics_and_quantitative_social_science\"><\/span>Economics, statistics, and quantitative social science<span class=\"ez-toc-section-end\"><\/span><\/h3>\n<p>Here the fabrication moves from the reference list into the results section, which is considerably more dangerous.<\/p>\n<ul>\n<li><a href=\"https:\/\/www.nature.com\/articles\/s41598-023-41032-5\">In an audit of 84 model-generated papers across policy, economics, and other social sciences<\/a>, 5 of the &#8220;literature reviews&#8221; were structured as empirical studies with invented methods and results.<\/li>\n<li>Among them was a paper containing fabricated correlation coefficients, regression coefficients, and p values, all internally consistent and plausible in magnitude.<\/li>\n<li>Models routinely invent survey sample sizes, response rates, and confidence intervals when asked to summarize a study they cannot access.<\/li>\n<li>Asked to interpret a real table, models transpose columns, misread footnotes, and silently convert nominal to real figures.<\/li>\n<li>Aggregate statistics such as &#8220;roughly 40% of firms&#8221; are generated with no underlying source and are difficult to trace precisely because they sound like common knowledge.<\/li>\n<\/ul>\n<h3><span class=\"ez-toc-section\" id=\"Sample_prompt_extracting_numbers_with_a_hallucination_trap_built_in\"><\/span>Sample prompt: extracting numbers with a hallucination trap built in<span class=\"ez-toc-section-end\"><\/span><\/h3>\n<p><em>From the attached PDF, extract every figure I asked about into a table with 3 columns: the value, the exact page number, and the verbatim sentence it came from. If a value does not appear in the document, write NOT PRESENT. Do not calculate, infer, or supply any number from your own knowledge.<\/em><\/p>\n<p>Requiring the verbatim source sentence makes fabrication far more difficult for the model and instantly checkable for you.<\/p>\n<h3><span class=\"ez-toc-section\" id=\"Computer_science_and_software_engineering\"><\/span>Computer science and software engineering<span class=\"ez-toc-section-end\"><\/span><\/h3>\n<p>Code is where hallucination becomes an attack surface rather than merely an error.<\/p>\n<ul>\n<li><a href=\"https:\/\/www.usenix.org\/conference\/usenixsecurity25\/presentation\/spracklen\">A large study of 16 code-generating models across 576,000 samples<\/a> found 19.6% of recommended software packages did not exist, with roughly 5% for commercial models and 22% for open-source ones.<\/li>\n<li>The study catalogued 205,474 unique hallucinated package names, and only 0.17% corresponded to real packages that had been deleted. The rest were pure invention.<\/li>\n<li>Hallucinated names repeat deterministically, which means an attacker can farm them, register the popular ones, and wait. This attack is now called <a href=\"https:\/\/www.trendaisecurity.com\/en\/resources-insights\/deep-research\/slopsquatting-when-ai-agents-hallucinate-malicious-packages\">slopsquatting<\/a>.<\/li>\n<li>Models also invent function signatures, command-line flags, and configuration keys for real libraries, which is harder to spot than a missing package because the library is genuine.<\/li>\n<li>More recent measurements on frontier models show that hallucination rates have reduced to roughly 5% to 6%, which is lower but not resolved.<\/li>\n<\/ul>\n<h3><span class=\"ez-toc-section\" id=\"History_and_the_humanities\"><\/span>History and the humanities<span class=\"ez-toc-section-end\"><\/span><\/h3>\n<p>Humanities research has less quantified evidence, mostly because archival claims are expensive to audit. The failure modes are nonetheless well documented anecdotally and follow directly from the mechanism.<\/p>\n<ul>\n<li>Invented archival references: a plausible collection name, box number, and folder at a real repository.<\/li>\n<li>Misattributed quotations, where a real aphorism is assigned to a more famous figure, mirroring the same error already common in training data.<\/li>\n<li>Fabricated translations of passages from texts the model has not memorized, often stylistically convincing.<\/li>\n<li>Confident dating of undated manuscripts, and invented provenance chains for objects.<\/li>\n<li>Blended historiography, where the positions of 2 or 3 real scholars are merged into a single school of thought that no one actually holds.<\/li>\n<\/ul>\n<p>&nbsp;<\/p>\n<h2><span class=\"ez-toc-section\" id=\"ChatGPT_Hallucination_Rates_What_the_Studies_Actually_Measure\"><\/span>ChatGPT Hallucination Rates: What the Studies Actually Measure<span class=\"ez-toc-section-end\"><\/span><\/h2>\n<p>ChatGPT is measured more than other systems because of its adoption, not because it is uniquely unreliable. Every model in this class hallucinates. Treat the numbers below as evidence about a category, not a verdict on one product.<\/p>\n<h3><span class=\"ez-toc-section\" id=\"ChatGPT_hallucination_rates_reported_in_citation_studies\"><\/span>ChatGPT hallucination rates reported in citation studies<span class=\"ez-toc-section-end\"><\/span><\/h3>\n<table>\n<thead>\n<tr>\n<td><strong>\u00a0<\/strong><\/p>\n<table width=\"563\">\n<thead>\n<tr>\n<td><strong>Study context<\/strong><\/td>\n<td width=\"71\"><strong>Citations audited<\/strong><\/td>\n<td><strong>Fabricated<\/strong><\/td>\n<td width=\"283\"><strong>DOI \/ URL<\/strong><\/td>\n<\/tr>\n<\/thead>\n<tbody>\n<tr>\n<td>Medical content, 30 papers, GPT-3.5<\/td>\n<td width=\"71\">115<\/td>\n<td>47%<\/td>\n<td width=\"283\"><a href=\"https:\/\/doi.org\/10.7759\/cureus.39238\">https:\/\/doi.org\/10.7759\/cureus.39238<\/a><\/td>\n<\/tr>\n<tr>\n<td>Medical questions, GPT-3.5<\/td>\n<td width=\"71\">59<\/td>\n<td>69%<\/td>\n<td width=\"283\"><a href=\"https:\/\/doi.org\/10.1016\/j.mcpdig.2023.05.004\">https:\/\/doi.org\/10.1016\/j.mcpdig.2023.05.004<\/a><\/td>\n<\/tr>\n<tr>\n<td>Clinical radiology, GPT-3<\/td>\n<td width=\"71\">343<\/td>\n<td>63.8%<\/td>\n<td width=\"283\"><a href=\"https:\/\/doi.org\/10.1177\/08465371231171125\">https:\/\/doi.org\/10.1177\/08465371231171125<\/a><\/td>\n<\/tr>\n<tr>\n<td>Multidisciplinary reviews, GPT-3.5<\/td>\n<td width=\"71\">222<\/td>\n<td>55%<\/td>\n<td width=\"283\"><a href=\"https:\/\/doi.org\/10.1038\/s41598-023-41032-5\">https:\/\/doi.org\/10.1038\/s41598-023-41032-5<\/a><\/td>\n<\/tr>\n<tr>\n<td>Multidisciplinary reviews, GPT-4<\/td>\n<td width=\"71\">414<\/td>\n<td>18%<\/td>\n<td width=\"283\"><a href=\"https:\/\/doi.org\/10.1038\/s41598-023-41032-5\">https:\/\/doi.org\/10.1038\/s41598-023-41032-5<\/a><\/td>\n<\/tr>\n<tr>\n<td>Mental health reviews, GPT-4o<\/td>\n<td width=\"71\">176<\/td>\n<td>19.9%<\/td>\n<td width=\"283\"><a href=\"https:\/\/doi.org\/10.2196\/80371\">https:\/\/doi.org\/10.2196\/80371<\/a><\/td>\n<\/tr>\n<tr>\n<td>Pooled across 6 early studies<\/td>\n<td width=\"71\">732<\/td>\n<td>51%<\/td>\n<td width=\"283\"><a href=\"https:\/\/doi.org\/10.1038\/s41598-023-41032-5\">https:\/\/doi.org\/10.1038\/s41598-023-41032-5<\/a><\/td>\n<\/tr>\n<\/tbody>\n<\/table>\n<\/td>\n<td><strong>\u00a0<\/strong><\/td>\n<td><strong>\u00a0<\/strong><\/td>\n<\/tr>\n<\/thead>\n<\/table>\n<p>The trend across model generations is real and substantial. <a href=\"https:\/\/doi.org\/10.1038\/s41598-023-41032-5\">In one study<\/a>, fabrication fell from 55% to 18% between GPT-3.5 and GPT-4, and substantive errors in the remaining real citations fell from 43% to 24%. Improvement is not the same as reliability: 18% is still roughly 1 fabricated source in every 5.<\/p>\n<h3><span class=\"ez-toc-section\" id=\"Why_published_rates_will_not_predict_your_rate\"><\/span>Why published rates will not predict your rate<span class=\"ez-toc-section-end\"><\/span><\/h3>\n<ul>\n<li>Discipline matters. Fabrication rates ran higher in geography than in medicine in the same comparison.<\/li>\n<li>Obscurity matters more than discipline. Well-covered topics produce more real citations; niche or emerging ones produce more fabricated citations.<\/li>\n<li>Prompt design shifts results measurably. Asking for a set number of sources raises fabrication.<\/li>\n<li>Tool configuration dominates. A model with live search and one without are effectively different systems.<\/li>\n<li>Studies age quickly. Most published figures test models that are now 2 or 3 generations old.<\/li>\n<li>Definitions differ. Some studies count a wrong volume number as an error, others as a fabrication.<\/li>\n<\/ul>\n<h3><span class=\"ez-toc-section\" id=\"Sample_prompt_measure_your_own_rate_before_trusting_a_workflow\"><\/span>Sample prompt: measure your own rate before trusting a workflow<span class=\"ez-toc-section-end\"><\/span><\/h3>\n<p><em>Give me 5 peer-reviewed sources on [your exact research topic], with full bibliographic details and DOIs.<\/em><\/p>\n<ul>\n<li>Then check all 5 yourself in Scopus, Web of Science, or PubMed.<\/li>\n<li>Repeat across 4 topics for a 20-citation sample.<\/li>\n<li>The fabrication rate <strong><em>you find<\/em><\/strong>, not a published benchmark, is your real risk number.<\/li>\n<\/ul>\n<p>&nbsp;<\/p>\n<h2><span class=\"ez-toc-section\" id=\"The_Risks_of_AI_Hallucination_for_Researchers_and_Institutions\"><\/span>The Risks of AI Hallucination for Researchers and Institutions<span class=\"ez-toc-section-end\"><\/span><\/h2>\n<p>The consequences scale outward: from an individual embarrassment, to a corrupted literature, to institutional liability.<\/p>\n<h3><span class=\"ez-toc-section\" id=\"Personal_and_professional_risk\"><\/span>Personal and professional risk<span class=\"ez-toc-section-end\"><\/span><\/h3>\n<ul>\n<li>Retraction, which is permanent and publicly indexed.<\/li>\n<li>Desk rejection and reviewer distrust that carries into future submissions.<\/li>\n<li>Academic misconduct proceedings, since fabricated citations can be treated as fabrication regardless of intent.<\/li>\n<li>Loss of grant funding and, for students, thesis failure or degree revocation.<\/li>\n<li>Reputational damage that outlasts the correction, as the sanctioned attorneys in the Avianca case discovered.<\/li>\n<\/ul>\n<h3><span class=\"ez-toc-section\" id=\"Contamination_of_the_scholarly_record\"><\/span>Contamination of the scholarly record<span class=\"ez-toc-section-end\"><\/span><\/h3>\n<p>The systemic risk is worse than the individual one, and it compounds.<\/p>\n<ul>\n<li>Fabricated references are more likely to survive in <a href=\"https:\/\/www.editage.com\/blog\/what-is-a-preprint\/\">preprints<\/a>, <a href=\"https:\/\/researcher.life\/blog\/article\/thesis-structure-outline-writing-tips\/\">theses<\/a>, and institutional repositories than in fully copyedited journal articles.<\/li>\n<li>Once a fake reference is cited, later authors may cite it secondhand without ever attempting retrieval, creating a citation cascade.<\/li>\n<li>Some fabricated references point to journals that have been identified as predatory, laundering low-quality journals into legitimate ones.<\/li>\n<li>Contaminated repositories then become retrieval corpora for the next generation of AI tools, closing the loop.<\/li>\n<li><a href=\"https:\/\/www.editage.com\/blog\/ai-for-systematic-reviews-how-to-use-ai-in-the-systematic-review-process\/\">Systematic reviews<\/a> and <a href=\"https:\/\/www.editage.com\/blog\/ai-for-meta-analysis-how-researchers-can-use-ai-in-evidence-synthesis\/\">meta-analyses<\/a> are especially exposed, because they aggregate at scale and rarely re-verify every included reference.<\/li>\n<\/ul>\n<h3><span class=\"ez-toc-section\" id=\"Legal_regulatory_and_research-integrity_exposure\"><\/span>Legal, regulatory, and research-integrity exposure<span class=\"ez-toc-section-end\"><\/span><\/h3>\n<table>\n<thead>\n<tr>\n<td><strong>Setting<\/strong><\/td>\n<td><strong>Exposure<\/strong><\/td>\n<\/tr>\n<\/thead>\n<tbody>\n<tr>\n<td>Litigation and legal filings<\/td>\n<td>Sanctions, fee-shifting, dismissal, and mandatory disclosure to affected judges.<\/td>\n<\/tr>\n<tr>\n<td>Clinical and public health guidance<\/td>\n<td>Patient harm, professional liability, and regulatory action.<\/td>\n<\/tr>\n<tr>\n<td>Regulatory and compliance submissions<\/td>\n<td>Findings of misrepresentation; penalties independent of intent.<\/td>\n<\/tr>\n<tr>\n<td>Grant applications and reporting<\/td>\n<td>Findings of research misconduct; funder debarment.<\/td>\n<\/tr>\n<tr>\n<td>Journalism and expert testimony<\/td>\n<td>Defamation exposure and loss of expert credibility.<\/td>\n<\/tr>\n<\/tbody>\n<\/table>\n<h3><span class=\"ez-toc-section\" id=\"Automation_bias_the_reason_plausible_errors_survive\"><\/span>Automation bias: the reason plausible errors survive<span class=\"ez-toc-section-end\"><\/span><\/h3>\n<p>We do not manually recheck the output of statistical software, and that habit of trust transfers to AI tools even though the underlying task is fundamentally different. A statistics package computes; a language model generates.<\/p>\n<p>Hallucinations are unusually hard to catch for 3 reasons. They are fluent, so nothing prompts a second look. They are confident, and people prefer confident answers, which is why models are trained toward confidence. And they are topically apt, so they confirm rather than disrupt your expectations.<\/p>\n<p>The practical implication: your attention is drawn to output that looks wrong, but the dangerous output looks right. When you\u2019re <a href=\"https:\/\/www.editage.com\/blog\/how-to-check-for-hallucinations-in-ai-text-examples-and-checklist-for-researchers-and-students\/\">verifying AI output<\/a>, you need to follow a defined workflow rather than \u201ceyeball\u201d or trust your intuition.<\/p>\n<p>&nbsp;<\/p>\n<h2><span class=\"ez-toc-section\" id=\"How_to_Detect_and_Reduce_Hallucinations_in_Your_Workflow\"><\/span>How to Detect and Reduce Hallucinations in Your Workflow<span class=\"ez-toc-section-end\"><\/span><\/h2>\n<p>None of this requires abandoning AI tools. It requires treating their output as an unverified draft rather than as evidence.<\/p>\n<h3><span class=\"ez-toc-section\" id=\"A_verification_checklist_for_every_AI-assisted_claim\"><\/span>A verification checklist for every AI-assisted claim<span class=\"ez-toc-section-end\"><\/span><\/h3>\n<table>\n<thead>\n<tr>\n<td><strong>Step<\/strong><\/td>\n<td><strong>What to do<\/strong><\/td>\n<td><strong>Why it catches things<\/strong><\/td>\n<\/tr>\n<\/thead>\n<tbody>\n<tr>\n<td>1. Resolve the DOI<\/td>\n<td>Paste it into doi.org and open the target.<\/td>\n<td>Fake DOIs frequently resolve to a real but unrelated paper.<\/td>\n<\/tr>\n<tr>\n<td>2. Search title and author<\/td>\n<td>Query Scopus, Web of Science, or PubMed independently.<\/td>\n<td>Confirms the work exists as described, not just that something similar does.<\/td>\n<\/tr>\n<tr>\n<td>3. Open the source<\/td>\n<td>Retrieve the full text, never just the abstract<\/td>\n<td>Real paper, invented finding is a common failure.<\/td>\n<\/tr>\n<tr>\n<td>4. Locate the claim<\/td>\n<td>Find the specific sentence, table, or figure cited.<\/td>\n<td>Catches misattached citations, the hardest RAG failure.<\/td>\n<\/tr>\n<tr>\n<td>5. Check the numbers<\/td>\n<td>Verify volume, issue, pages, and year separately.<\/td>\n<td>Numeric fields fail most often even in real references.<\/td>\n<\/tr>\n<tr>\n<td>6. Check the chapter<\/td>\n<td>For book chapters, confirm the containing book exists.<\/td>\n<td>Chapters have the highest fabrication rate of any type.<\/td>\n<\/tr>\n<tr>\n<td>7. Log what was AI-assisted<\/td>\n<td>Keep a record of which claims came from a model.<\/td>\n<td>Makes disclosure and later audit possible.<\/td>\n<\/tr>\n<\/tbody>\n<\/table>\n<p>The one step that does not belong on this list is asking the model to check itself. In the Avianca case that step produced a confident confirmation of 6 nonexistent cases, and studies consistently find models defend their own fabrications.<\/p>\n<h3><span class=\"ez-toc-section\" id=\"Prompting_practices_that_measurably_lower_risk\"><\/span>Prompting practices that measurably lower risk<span class=\"ez-toc-section-end\"><\/span><\/h3>\n<table>\n<thead>\n<tr>\n<td><strong>Goal<\/strong><\/td>\n<td><strong>Sample prompt<\/strong><\/td>\n<\/tr>\n<\/thead>\n<tbody>\n<tr>\n<td>Force abstention<\/td>\n<td>If you are not confident a specific source exists, write &#8220;I do not know of a specific source&#8221; instead of producing a citation. I would rather have 2 real references than 10 uncertain ones.<\/td>\n<\/tr>\n<tr>\n<td>Ground in your own corpus<\/td>\n<td>Answer using only the attached documents. For each claim, quote the sentence you relied on and give its page number. If the documents do not address the question, say so.<\/td>\n<\/tr>\n<tr>\n<td>Surface uncertainty<\/td>\n<td>After your answer, list every factual claim you made that you are less than 90% confident about, and say what would need to be checked.<\/td>\n<\/tr>\n<tr>\n<td>Get the opposing case<\/td>\n<td>Now argue the opposite position as strongly as you can, and identify the strongest evidence against what you just told me.<\/td>\n<\/tr>\n<tr>\n<td>Separate recall from reasoning<\/td>\n<td>Do not cite anything. Explain the mechanisms and the main debates. I will supply the sources and come back to you.<\/td>\n<\/tr>\n<tr>\n<td>Cross-check a suspect claim<\/td>\n<td>Here is a claim I found. Without assuming it is true, tell me what evidence would confirm or disconfirm it and where I would look.<\/td>\n<\/tr>\n<\/tbody>\n<\/table>\n<h3><span class=\"ez-toc-section\" id=\"Tools_that_help_and_where_each_stops_helping\"><\/span>Tools that help, and where each stops helping<span class=\"ez-toc-section-end\"><\/span><\/h3>\n<table>\n<thead>\n<tr>\n<td><strong>Tool type<\/strong><\/td>\n<td><strong>What it catches<\/strong><\/td>\n<td><strong>What it misses<\/strong><\/td>\n<\/tr>\n<\/thead>\n<tbody>\n<tr>\n<td>Reference managers<\/td>\n<td>DOIs that fail to resolve; metadata mismatches.<\/td>\n<td>Real DOIs attached to the wrong claim.<\/td>\n<\/tr>\n<tr>\n<td>Search-enabled AI<\/td>\n<td>Wholesale invention of sources.<\/td>\n<td>Misreading of retrieved sources; misattached citations.<\/td>\n<\/tr>\n<tr>\n<td>Citation audit scripts<\/td>\n<td>Bulk existence checks across a bibliography.<\/td>\n<td>Whether the source supports what you said it does.<\/td>\n<\/tr>\n<tr>\n<td><a href=\"https:\/\/www.editage.com\/blog\/ai-detection-in-academic-writing-ai-detectors-accuracy-false-positives-and-researcher-concerns\/\">AI text detectors<\/a><\/td>\n<td>Little that is reliable; false positives are common.<\/td>\n<td>Not a verification method. Do not depend on these.<\/td>\n<\/tr>\n<tr>\n<td>A second model<\/td>\n<td>Some self-inconsistent fabrications.<\/td>\n<td>Shared errors, since models share training data and biases.<\/td>\n<\/tr>\n<tr>\n<td>A professional human librarian<\/td>\n<td>Nearly everything above, plus database strategy.<\/td>\n<td>Underused. Subject librarians are the highest-value check available.<\/td>\n<\/tr>\n<\/tbody>\n<\/table>\n<h3><span class=\"ez-toc-section\" id=\"Disclosure_obligations\"><\/span>Disclosure obligations<span class=\"ez-toc-section-end\"><\/span><\/h3>\n<ul>\n<li><a href=\"https:\/\/www.editage.com\/blog\/is-ai-allowed-in-journal-submissions-journal-and-publisher-ai-policies-and-how-much-ai-is-acceptable-in-a-research-paper\/\">Most major publishers now require disclosure of generative AI use<\/a> in the methods or acknowledgments.<\/li>\n<li><a href=\"https:\/\/www.editage.com\/blog\/first-author-vs-corresponding-author-in-a-research-paper-how-to-decide-authorship-with-examples-templates\/#Can_AI_Be_Listed_as_an_Author_on_a_Research_Paper\">AI tools cannot be listed as authors<\/a>, since they cannot take responsibility for the work. This is now near-universal policy.<\/li>\n<li><a href=\"https:\/\/www.editage.com\/blog\/is-ai-allowed-in-journal-submissions-journal-and-publisher-ai-policies-and-how-much-ai-is-acceptable-in-a-research-paper\/#The_3_tiers_of_AI_use\">Requirements differ between using AI for language editing and using it for content generation<\/a> or analysis. Check the specific journal.<\/li>\n<li>Funders and institutions increasingly have separate policies from publishers. Both apply.<\/li>\n<li>Policies are changing quickly, so verify against the current author guidelines rather than what applied to your last submission.<\/li>\n<\/ul>\n<p>&nbsp;<\/p>\n<h2><span class=\"ez-toc-section\" id=\"Frequently_Asked_Questions\"><\/span>Frequently Asked Questions<span class=\"ez-toc-section-end\"><\/span><\/h2>\n<h3><span class=\"ez-toc-section\" id=\"How_do_I_check_if_an_AI-generated_citation_is_real\"><\/span>How do I check if an AI-generated citation is real?<span class=\"ez-toc-section-end\"><\/span><\/h3>\n<p>Search the exact title in Google Scholar, Scopus, Web of Science, or PubMed, then resolve the DOI separately and confirm the target matches the title and authors given. Both steps are necessary: a real-looking DOI can resolve to an unrelated paper, and a real title can be paired with fabricated publication details. Finally, open the source and locate the specific claim. Never ask the model to verify itself.<\/p>\n<h3><span class=\"ez-toc-section\" id=\"Why_does_ChatGPT_make_up_DOIs_that_link_to_real_papers\"><\/span>Why does ChatGPT make up DOIs that link to real papers?<span class=\"ez-toc-section-end\"><\/span><\/h3>\n<p>A DOI is a short, highly patterned string, so the model generates one that looks structurally correct without any lookup. Some of those strings happen to be live identifiers belonging to other papers. In one 2025 audit, more than 94% of fabricated citations carried a DOI, and about 64% of those resolved to real but unrelated work. A link that opens is not evidence; only the content at the other end is.<\/p>\n<h3><span class=\"ez-toc-section\" id=\"Can_I_use_ChatGPT_for_a_literature_review\"><\/span>Can I use ChatGPT for a literature review?<span class=\"ez-toc-section-end\"><\/span><\/h3>\n<p>Use it for scoping, terminology, and structure, not for finding sources. It is genuinely good at explaining a concept, generating search terms, and polishing language. It is unreliable at producing the reference list, which is where measured fabrication rates run from 18% to 69%. Find sources in a real database always.<\/p>\n<h3><span class=\"ez-toc-section\" id=\"Do_newer_AI_models_still_hallucinate\"><\/span>Do newer AI models still hallucinate?<span class=\"ez-toc-section-end\"><\/span><\/h3>\n<p>Yes, at lower rates. Fabrication in one multidisciplinary study fell from 55% to 18% across a single model generation, and reasoning modes reduce factual errors further. But the cause is structural rather than a fixable defect, and <a href=\"https:\/\/arxiv.org\/html\/2509.04664v1\">OpenAI&#8217;s own researchers<\/a> describe hallucination as an expected statistical outcome of current training and evaluation practice. Assume a reduced rate, not a zero rate.<\/p>\n<h3><span class=\"ez-toc-section\" id=\"What_is_the_difference_between_an_AI_hallucination_and_a_lie\"><\/span>What is the difference between an AI hallucination and a lie?<span class=\"ez-toc-section-end\"><\/span><\/h3>\n<p>A lie requires knowing the truth and choosing to misstate it. A language model has no internal representation of truth to depart from; it produces the most plausible continuation, and truth is incidental. This is why some philosophers argue the accurate label is indifference to truth rather than deception. The distinction matters practically: there is no dishonesty to detect, so behavioral cues will not help you.<\/p>\n<h3><span class=\"ez-toc-section\" id=\"Do_I_need_to_disclose_AI_use_when_submitting_to_a_journal\"><\/span>Do I need to disclose AI use when submitting to a journal?<span class=\"ez-toc-section-end\"><\/span><\/h3>\n<p>Almost certainly yes, though the threshold varies. Most major publishers require disclosure of generative AI use in the manuscript, and none permit AI tools to be listed as authors. Some distinguish between language polishing and substantive content generation. Check the specific journal&#8217;s current author guidelines, along with your funder and institutional policies, since these are separate and all apply.<\/p>\n<h3><span class=\"ez-toc-section\" id=\"Can_AI_hallucinations_be_eliminated_completely\"><\/span>Can AI hallucinations be eliminated completely?<span class=\"ez-toc-section-end\"><\/span><\/h3>\n<p>Not with current architectures. The behavior follows from generating text by prediction rather than retrieval, and from evaluation practices that score confident guessing above admitting uncertainty. Grounding, retrieval, and reasoning modes reduce the rate substantially. Changing benchmark scoring to reward calibrated abstention would help further. Complete elimination would require a fundamentally different design.<\/p>\n<h3><span class=\"ez-toc-section\" id=\"Which_AI_tool_hallucinates_the_least_for_academic_research\"><\/span>Which AI tool hallucinates the least for academic research?<span class=\"ez-toc-section-end\"><\/span><\/h3>\n<p>The honest answer is that published rankings go stale within months, and configuration matters more than brand. A model with live search and grounding in your own documents will outperform a more capable model answering from memory. Rather than choosing on reputation, run a test on your own topics and manually verify 20 citations. You\u2019ll then be able to gauge the actual likelihood of citation hallucination if you continue using AI in your research.<\/p>\n","protected":false},"excerpt":{"rendered":"Key Takeaways Hallucination is inherent in generative AI. Large language models generate statistically plausible text; they do not retrieve verified facts. A fabricated citation and a real one are produced by exactly the same process. Citations are the highest-risk output. Peer-reviewed studies report fabrication rates from 18% to 69%, varying by model, discipline, and prompt [&hellip;]","protected":false},"author":3,"featured_media":1998,"comment_status":"open","ping_status":"open","sticky":false,"template":"","format":"standard","meta":{"_ayudawp_aiss_exclude":false,"_ayudawp_aiss_summary":"The typical fabricated reference names real authors, a real journal, and a working DOI that points to an unrelated paper. An observational study of 30 short medical papers generated by ChatGPT found that of 115 generated references, 47% were fabricated, 46% were real but inaccurate, and only 7% were both real and accurate. Fabricated medical references typically pair real, topically appropriate authors with a real journal and an invented title, which defeats casual plausibility checks.","_ayudawp_aiss_summary_provider":"manual","_ayudawp_aiss_summary_hash":"5c8655a455117275a15796d0d01a8fc412c438df"},"categories":[4],"tags":[],"yoast_head":"<!-- This site is optimized with the Yoast SEO plugin v20.6 - https:\/\/yoast.com\/wordpress\/plugins\/seo\/ -->\n<title>What Are AI Hallucinations in Research: Causes, Examples, and Risks - Educational Articles For Researchers, Students And Authors - Editage Blog<\/title>\n<meta name=\"robots\" content=\"index, follow, max-snippet:-1, max-image-preview:large, max-video-preview:-1\" \/>\n<link rel=\"canonical\" href=\"https:\/\/www.editage.com\/blog\/what-are-ai-hallucinations-in-research-causes-examples-and-risks\/\" \/>\n<meta property=\"og:locale\" content=\"en_US\" \/>\n<meta property=\"og:type\" content=\"article\" \/>\n<meta property=\"og:title\" content=\"What Are AI Hallucinations in Research: Causes, Examples, and Risks - Educational Articles For Researchers, Students And Authors - Editage Blog\" \/>\n<meta property=\"og:description\" content=\"Key Takeaways Hallucination is inherent in generative AI. Large language models generate statistically plausible text; they do not retrieve verified facts. A fabricated citation and a real one are produced by exactly the same process. Citations are the highest-risk output. Peer-reviewed studies report fabrication rates from 18% to 69%, varying by model, discipline, and prompt [&hellip;]\" \/>\n<meta property=\"og:url\" content=\"https:\/\/www.editage.com\/blog\/what-are-ai-hallucinations-in-research-causes-examples-and-risks\/\" \/>\n<meta property=\"og:site_name\" content=\"Educational Articles For Researchers, Students And Authors - Editage Blog\" \/>\n<meta property=\"article:published_time\" content=\"2026-08-21T02:21:42+00:00\" \/>\n<meta property=\"article:modified_time\" content=\"2026-08-19T03:52:30+00:00\" \/>\n<meta property=\"og:image\" content=\"https:\/\/www.editage.com\/blog\/wp-content\/uploads\/2026\/08\/AI-hallucinations-causes-examples.jpg\" \/>\n\t<meta property=\"og:image:width\" content=\"1024\" \/>\n\t<meta property=\"og:image:height\" content=\"559\" \/>\n\t<meta property=\"og:image:type\" content=\"image\/jpeg\" \/>\n<meta name=\"author\" content=\"Marisha Rodrigues\" \/>\n<meta name=\"twitter:card\" content=\"summary_large_image\" \/>\n<meta name=\"twitter:label1\" content=\"Written by\" \/>\n\t<meta name=\"twitter:data1\" content=\"Marisha Rodrigues\" \/>\n\t<meta name=\"twitter:label2\" content=\"Est. reading time\" \/>\n\t<meta name=\"twitter:data2\" content=\"21 minutes\" \/>\n<script type=\"application\/ld+json\" class=\"yoast-schema-graph\">{\"@context\":\"https:\/\/schema.org\",\"@graph\":[{\"@type\":\"Article\",\"@id\":\"https:\/\/www.editage.com\/blog\/what-are-ai-hallucinations-in-research-causes-examples-and-risks\/#article\",\"isPartOf\":{\"@id\":\"https:\/\/www.editage.com\/blog\/what-are-ai-hallucinations-in-research-causes-examples-and-risks\/\"},\"author\":{\"name\":\"Marisha Rodrigues\",\"@id\":\"https:\/\/www.editage.com\/blog\/#\/schema\/person\/60d7626072744221b2260692486b6ff1\"},\"headline\":\"What Are AI Hallucinations in Research: Causes, Examples, and Risks\",\"datePublished\":\"2026-08-21T02:21:42+00:00\",\"dateModified\":\"2026-08-19T03:52:30+00:00\",\"mainEntityOfPage\":{\"@id\":\"https:\/\/www.editage.com\/blog\/what-are-ai-hallucinations-in-research-causes-examples-and-risks\/\"},\"wordCount\":4676,\"commentCount\":0,\"publisher\":{\"@id\":\"https:\/\/www.editage.com\/blog\/#organization\"},\"articleSection\":[\"AI in Research\"],\"inLanguage\":\"en-US\",\"potentialAction\":[{\"@type\":\"CommentAction\",\"name\":\"Comment\",\"target\":[\"https:\/\/www.editage.com\/blog\/what-are-ai-hallucinations-in-research-causes-examples-and-risks\/#respond\"]}]},{\"@type\":\"WebPage\",\"@id\":\"https:\/\/www.editage.com\/blog\/what-are-ai-hallucinations-in-research-causes-examples-and-risks\/\",\"url\":\"https:\/\/www.editage.com\/blog\/what-are-ai-hallucinations-in-research-causes-examples-and-risks\/\",\"name\":\"What Are AI Hallucinations in Research: Causes, Examples, and Risks - Educational Articles For Researchers, Students And Authors - Editage Blog\",\"isPartOf\":{\"@id\":\"https:\/\/www.editage.com\/blog\/#website\"},\"datePublished\":\"2026-08-21T02:21:42+00:00\",\"dateModified\":\"2026-08-19T03:52:30+00:00\",\"breadcrumb\":{\"@id\":\"https:\/\/www.editage.com\/blog\/what-are-ai-hallucinations-in-research-causes-examples-and-risks\/#breadcrumb\"},\"inLanguage\":\"en-US\",\"potentialAction\":[{\"@type\":\"ReadAction\",\"target\":[\"https:\/\/www.editage.com\/blog\/what-are-ai-hallucinations-in-research-causes-examples-and-risks\/\"]}]},{\"@type\":\"BreadcrumbList\",\"@id\":\"https:\/\/www.editage.com\/blog\/what-are-ai-hallucinations-in-research-causes-examples-and-risks\/#breadcrumb\",\"itemListElement\":[{\"@type\":\"ListItem\",\"position\":1,\"name\":\"Home\",\"item\":\"https:\/\/www.editage.com\/blog\/\"},{\"@type\":\"ListItem\",\"position\":2,\"name\":\"What Are AI Hallucinations in Research: Causes, Examples, and Risks\"}]},{\"@type\":\"WebSite\",\"@id\":\"https:\/\/www.editage.com\/blog\/#website\",\"url\":\"https:\/\/www.editage.com\/blog\/\",\"name\":\"Educational Articles For Researchers, Students And Authors - Editage Blog\",\"description\":\"Get insightful educational articles from the world of academia for researchers, students and authors. Visit Editage Blog for helpful content and tips on getting published and writing articles that are up to international journal publication standards. Click here to find out more!\",\"publisher\":{\"@id\":\"https:\/\/www.editage.com\/blog\/#organization\"},\"potentialAction\":[{\"@type\":\"SearchAction\",\"target\":{\"@type\":\"EntryPoint\",\"urlTemplate\":\"https:\/\/www.editage.com\/blog\/?s={search_term_string}\"},\"query-input\":\"required name=search_term_string\"}],\"inLanguage\":\"en-US\"},{\"@type\":\"Organization\",\"@id\":\"https:\/\/www.editage.com\/blog\/#organization\",\"name\":\"Educational Articles For Researchers, Students And Authors - Editage Blog\",\"url\":\"https:\/\/www.editage.com\/blog\/\",\"logo\":{\"@type\":\"ImageObject\",\"inLanguage\":\"en-US\",\"@id\":\"https:\/\/www.editage.com\/blog\/#\/schema\/logo\/image\/\",\"url\":\"https:\/\/www.editage.com\/blog\/wp-content\/uploads\/2022\/08\/editage-logo.png\",\"contentUrl\":\"https:\/\/www.editage.com\/blog\/wp-content\/uploads\/2022\/08\/editage-logo.png\",\"width\":394,\"height\":82,\"caption\":\"Educational Articles For Researchers, Students And Authors - Editage Blog\"},\"image\":{\"@id\":\"https:\/\/www.editage.com\/blog\/#\/schema\/logo\/image\/\"}},{\"@type\":\"Person\",\"@id\":\"https:\/\/www.editage.com\/blog\/#\/schema\/person\/60d7626072744221b2260692486b6ff1\",\"name\":\"Marisha Rodrigues\",\"image\":{\"@type\":\"ImageObject\",\"inLanguage\":\"en-US\",\"@id\":\"https:\/\/www.editage.com\/blog\/#\/schema\/person\/image\/\",\"url\":\"https:\/\/secure.gravatar.com\/avatar\/3fdb8ce7f366f83f27047a1644a5ff30?s=96&d=mm&r=g\",\"contentUrl\":\"https:\/\/secure.gravatar.com\/avatar\/3fdb8ce7f366f83f27047a1644a5ff30?s=96&d=mm&r=g\",\"caption\":\"Marisha Rodrigues\"},\"description\":\"A BELS-certified editor with 15+ years of experience in academic publishing and author education\",\"url\":\"https:\/\/www.editage.com\/blog\/author\/marishar\/\"}]}<\/script>\n<!-- \/ Yoast SEO plugin. -->","yoast_head_json":{"title":"What Are AI Hallucinations in Research: Causes, Examples, and Risks - Educational Articles For Researchers, Students And Authors - Editage Blog","robots":{"index":"index","follow":"follow","max-snippet":"max-snippet:-1","max-image-preview":"max-image-preview:large","max-video-preview":"max-video-preview:-1"},"canonical":"https:\/\/www.editage.com\/blog\/what-are-ai-hallucinations-in-research-causes-examples-and-risks\/","og_locale":"en_US","og_type":"article","og_title":"What Are AI Hallucinations in Research: Causes, Examples, and Risks - Educational Articles For Researchers, Students And Authors - Editage Blog","og_description":"Key Takeaways Hallucination is inherent in generative AI. Large language models generate statistically plausible text; they do not retrieve verified facts. A fabricated citation and a real one are produced by exactly the same process. Citations are the highest-risk output. Peer-reviewed studies report fabrication rates from 18% to 69%, varying by model, discipline, and prompt [&hellip;]","og_url":"https:\/\/www.editage.com\/blog\/what-are-ai-hallucinations-in-research-causes-examples-and-risks\/","og_site_name":"Educational Articles For Researchers, Students And Authors - Editage Blog","article_published_time":"2026-08-21T02:21:42+00:00","article_modified_time":"2026-08-19T03:52:30+00:00","og_image":[{"width":1024,"height":559,"url":"https:\/\/www.editage.com\/blog\/wp-content\/uploads\/2026\/08\/AI-hallucinations-causes-examples.jpg","type":"image\/jpeg"}],"author":"Marisha Rodrigues","twitter_card":"summary_large_image","twitter_misc":{"Written by":"Marisha Rodrigues","Est. reading time":"21 minutes"},"schema":{"@context":"https:\/\/schema.org","@graph":[{"@type":"Article","@id":"https:\/\/www.editage.com\/blog\/what-are-ai-hallucinations-in-research-causes-examples-and-risks\/#article","isPartOf":{"@id":"https:\/\/www.editage.com\/blog\/what-are-ai-hallucinations-in-research-causes-examples-and-risks\/"},"author":{"name":"Marisha Rodrigues","@id":"https:\/\/www.editage.com\/blog\/#\/schema\/person\/60d7626072744221b2260692486b6ff1"},"headline":"What Are AI Hallucinations in Research: Causes, Examples, and Risks","datePublished":"2026-08-21T02:21:42+00:00","dateModified":"2026-08-19T03:52:30+00:00","mainEntityOfPage":{"@id":"https:\/\/www.editage.com\/blog\/what-are-ai-hallucinations-in-research-causes-examples-and-risks\/"},"wordCount":4676,"commentCount":0,"publisher":{"@id":"https:\/\/www.editage.com\/blog\/#organization"},"articleSection":["AI in Research"],"inLanguage":"en-US","potentialAction":[{"@type":"CommentAction","name":"Comment","target":["https:\/\/www.editage.com\/blog\/what-are-ai-hallucinations-in-research-causes-examples-and-risks\/#respond"]}]},{"@type":"WebPage","@id":"https:\/\/www.editage.com\/blog\/what-are-ai-hallucinations-in-research-causes-examples-and-risks\/","url":"https:\/\/www.editage.com\/blog\/what-are-ai-hallucinations-in-research-causes-examples-and-risks\/","name":"What Are AI Hallucinations in Research: Causes, Examples, and Risks - Educational Articles For Researchers, Students And Authors - Editage Blog","isPartOf":{"@id":"https:\/\/www.editage.com\/blog\/#website"},"datePublished":"2026-08-21T02:21:42+00:00","dateModified":"2026-08-19T03:52:30+00:00","breadcrumb":{"@id":"https:\/\/www.editage.com\/blog\/what-are-ai-hallucinations-in-research-causes-examples-and-risks\/#breadcrumb"},"inLanguage":"en-US","potentialAction":[{"@type":"ReadAction","target":["https:\/\/www.editage.com\/blog\/what-are-ai-hallucinations-in-research-causes-examples-and-risks\/"]}]},{"@type":"BreadcrumbList","@id":"https:\/\/www.editage.com\/blog\/what-are-ai-hallucinations-in-research-causes-examples-and-risks\/#breadcrumb","itemListElement":[{"@type":"ListItem","position":1,"name":"Home","item":"https:\/\/www.editage.com\/blog\/"},{"@type":"ListItem","position":2,"name":"What Are AI Hallucinations in Research: Causes, Examples, and Risks"}]},{"@type":"WebSite","@id":"https:\/\/www.editage.com\/blog\/#website","url":"https:\/\/www.editage.com\/blog\/","name":"Educational Articles For Researchers, Students And Authors - Editage Blog","description":"Get insightful educational articles from the world of academia for researchers, students and authors. Visit Editage Blog for helpful content and tips on getting published and writing articles that are up to international journal publication standards. Click here to find out more!","publisher":{"@id":"https:\/\/www.editage.com\/blog\/#organization"},"potentialAction":[{"@type":"SearchAction","target":{"@type":"EntryPoint","urlTemplate":"https:\/\/www.editage.com\/blog\/?s={search_term_string}"},"query-input":"required name=search_term_string"}],"inLanguage":"en-US"},{"@type":"Organization","@id":"https:\/\/www.editage.com\/blog\/#organization","name":"Educational Articles For Researchers, Students And Authors - Editage Blog","url":"https:\/\/www.editage.com\/blog\/","logo":{"@type":"ImageObject","inLanguage":"en-US","@id":"https:\/\/www.editage.com\/blog\/#\/schema\/logo\/image\/","url":"https:\/\/www.editage.com\/blog\/wp-content\/uploads\/2022\/08\/editage-logo.png","contentUrl":"https:\/\/www.editage.com\/blog\/wp-content\/uploads\/2022\/08\/editage-logo.png","width":394,"height":82,"caption":"Educational Articles For Researchers, Students And Authors - Editage Blog"},"image":{"@id":"https:\/\/www.editage.com\/blog\/#\/schema\/logo\/image\/"}},{"@type":"Person","@id":"https:\/\/www.editage.com\/blog\/#\/schema\/person\/60d7626072744221b2260692486b6ff1","name":"Marisha Rodrigues","image":{"@type":"ImageObject","inLanguage":"en-US","@id":"https:\/\/www.editage.com\/blog\/#\/schema\/person\/image\/","url":"https:\/\/secure.gravatar.com\/avatar\/3fdb8ce7f366f83f27047a1644a5ff30?s=96&d=mm&r=g","contentUrl":"https:\/\/secure.gravatar.com\/avatar\/3fdb8ce7f366f83f27047a1644a5ff30?s=96&d=mm&r=g","caption":"Marisha Rodrigues"},"description":"A BELS-certified editor with 15+ years of experience in academic publishing and author education","url":"https:\/\/www.editage.com\/blog\/author\/marishar\/"}]}},"jetpack_featured_media_url":"https:\/\/www.editage.com\/blog\/wp-content\/uploads\/2026\/08\/AI-hallucinations-causes-examples.jpg","_links":{"self":[{"href":"https:\/\/www.editage.com\/blog\/wp-json\/wp\/v2\/posts\/1996"}],"collection":[{"href":"https:\/\/www.editage.com\/blog\/wp-json\/wp\/v2\/posts"}],"about":[{"href":"https:\/\/www.editage.com\/blog\/wp-json\/wp\/v2\/types\/post"}],"author":[{"embeddable":true,"href":"https:\/\/www.editage.com\/blog\/wp-json\/wp\/v2\/users\/3"}],"replies":[{"embeddable":true,"href":"https:\/\/www.editage.com\/blog\/wp-json\/wp\/v2\/comments?post=1996"}],"version-history":[{"count":3,"href":"https:\/\/www.editage.com\/blog\/wp-json\/wp\/v2\/posts\/1996\/revisions"}],"predecessor-version":[{"id":2143,"href":"https:\/\/www.editage.com\/blog\/wp-json\/wp\/v2\/posts\/1996\/revisions\/2143"}],"wp:featuredmedia":[{"embeddable":true,"href":"https:\/\/www.editage.com\/blog\/wp-json\/wp\/v2\/media\/1998"}],"wp:attachment":[{"href":"https:\/\/www.editage.com\/blog\/wp-json\/wp\/v2\/media?parent=1996"}],"wp:term":[{"taxonomy":"category","embeddable":true,"href":"https:\/\/www.editage.com\/blog\/wp-json\/wp\/v2\/categories?post=1996"},{"taxonomy":"post_tag","embeddable":true,"href":"https:\/\/www.editage.com\/blog\/wp-json\/wp\/v2\/tags?post=1996"}],"curies":[{"name":"wp","href":"https:\/\/api.w.org\/{rel}","templated":true}]}}