{"product_id":"9783032363206","title":"Recent Advances in Multimodal Hallucination","description":"\u003ch1\u003eRecent Advances in Multimodal Hallucination\u003c\/h1\u003e\u003ch3\u003eLiqiang Jing | Yue Zhang | Xinya Du\u003c\/h3\u003e\u003cdiv\u003e\u003cb\u003eComputers \/ Artificial Intelligence \/ General\u003c\/b\u003e\u003c\/div\u003e\u003cbr\u003e\u003cdiv\u003e\n\u003cp\u003e\u003cstrong\u003eHallucination in Multimodal Models\u003c\/strong\u003e is the first comprehensive research monograph dedicated to the growing challenge of \u003cem\u003ehallucinations\u003c\/em\u003e in large-scale multimodal AI systems, particularly vision-language models (VLMs) and multimodal large language models (MLLMs). The book systematically defines, categorizes, evaluates, and mitigates hallucinations — cases where models generate content that is factually inconsistent, visually unsupported, or commonsensically implausible. These hallucinations have become increasingly problematic in real-world applications of AI, including robotics, autonomous systems, and AI-generated media, where the consequences of inaccurate outputs can be severe.\u003c\/p\u003e\r\n\u003cp\u003eThe purpose of this book is threefold:\u003c\/p\u003e\r\n\u003cp\u003e(1) to formalize the types and causes of hallucination in multimodal models;\u003c\/p\u003e\r\n\u003cp\u003e(2) to present state-of-the-art evaluation frameworks, such as \u003cstrong\u003eFaithScore\u003c\/strong\u003e and \u003cstrong\u003eFIHA\u003c\/strong\u003e, for quantifying hallucination at a fine-grained level; and\u003c\/p\u003e\r\n\u003cp\u003e(3) to introduce a unified mitigation framework.\u003c\/p\u003e\r\n\u003cp\u003e\u003cspan data-olk-copy-source=\"MessageBody\"\u003eThe book presents recent research results, including FaithScore, FIHA, FIFA, FGAIF, and Dentist, together with empirical studies across leading large vision-language models. Of particular interest are its fine-grained approaches to atomic fact verification, semantic dependency modeling, unified text-video hallucination evaluation, reward-based alignment, and training-free hallucination mitigation. The book also examines how these methods can improve the faithfulness and reliability of multimodal model outputs, offering practical tools for both researchers and engineers.\u003c\/span\u003e\u003c\/p\u003e\r\n\u003cp\u003eThis book complements and extends the existing literature on multimodal model evaluation (e.g., MME, SEED-Bench, LAMM) by moving beyond surface-level metrics to offer deeper, interpretable, and automated hallucination analysis. Unlike survey papers or benchmarks that only diagnose the problem, our monograph provides \u003cstrong\u003ea cohesive solution path\u003c\/strong\u003e from diagnosis to mitigation, built upon novel technical contributions and real-world implementations.\u003c\/p\u003e\r\n\u003cp\u003eAs this is the first edition, it introduces original theoretical frameworks, algorithms, benchmarks, and design paradigms. It is intended to serve as a reference for graduate students, academic researchers, and industry practitioners working in natural language processing, computer vision, embodied AI, and trustworthy AI systems.\u003c\/p\u003e\n\u003c\/div\u003e\u003cdiv\u003e\n\u003cp\u003eLiqiang Jing is a PhD candidate in the Computer Science Department at the University of Texas at Dallas. He was a research intern at Alibaba Damo Acadamy and Tencent AI Lab (Seattle). His research interests include multimodal learning and natural language processing. He has published over 15 papers in leading NLP and AI conferences, such as ICLR, ACL, EMNLP, ACM SIGIR, ACM MM, AAAI, etc. In addition, he has served as a reviewer for many top conferences and journals, such as NeurIPS, ICLR, ICML, and EMNLP.\u003c\/p\u003e\r\n\u003cp\u003eYue Zhang is s a PhD candidate in the Computer Science Department at the University of Texas at Dallas. Her research interests include multimodal learning and natural language processing.\u003c\/p\u003e\r\n\u003cp\u003eXinya Du is a tenure-track Assistant Professor at the University of Texas at Dallas, Computer Science Department. He earned a Ph.D. degree from Cornell University and was a Postdoctoral Research Associate at UIUC. His research is on NLP, Knowledge and Reasoning, and LLMs. The goal is to build intelligent machines with both faithful knowledge \\\u0026amp; reasoning capabilities. His work has been published in leading NLP and ML conferences (ACL, EMNLP, ICLR). His work was included in the list of Most Influential ACL Papers by Paper Digest and has been covered by major media like New Scientist. He was named a Spotlight Rising Star in Data Science and was selected for the New Faculty Highlights program by AAAI. He is the recipient of the 2024 Amazon Research Award and the 2024 NSF CAREER Award\u003c\/p\u003e\n\u003c\/div\u003e\u003cbr\u003e\u003ctable\u003e\n\u003ctr\u003e\n\u003ctd\u003ePublication Date: \u003c\/td\u003e\n\u003ctd\u003e14 November 2026\u003c\/td\u003e\n\u003c\/tr\u003e\n\u003ctr\u003e\n\u003ctd\u003ePublisher: \u003c\/td\u003e\n\u003ctd\u003eSpringer Nature Switzerland\u003c\/td\u003e\n\u003c\/tr\u003e\n\u003ctr\u003e\n\u003ctd\u003eImprint: \u003c\/td\u003e\n\u003ctd\u003eSpringer\u003c\/td\u003e\n\u003c\/tr\u003e\n\u003ctr\u003e\n\u003ctd\u003eISBN-13: \u003c\/td\u003e\n\u003ctd\u003e9783032363206\u003c\/td\u003e\n\u003c\/tr\u003e\n\u003ctr\u003e\n\u003ctd\u003eFormat: \u003c\/td\u003e\n\u003ctd\u003eHardback\u003c\/td\u003e\n\u003c\/tr\u003e\n\u003c\/table\u003e","brand":"Springer Nature Switzerland","offers":[{"title":"Default Title","offer_id":51562258661516,"sku":"9783032363206","price":49.49,"currency_code":"USD","in_stock":true}],"url":"https:\/\/fh90cf-fv.myshopify.com\/products\/9783032363206","provider":"Late Knight Books and Services, LLC","version":"1.0","type":"link"}