لطفا منتظر بمانید ...
0% Complete
صفحه اصلی
/
ششمین کنفرانس بین المللی محاسبات نرم
When Autocomplete Lies - An Evaluation On LLM Hallucination
نویسندگان :
Reza Shabankhah
1
Amirhossein Moradi
2
Abdorreza Hesam mohseni
3
1- دانشگاه گیلان - دانشکده فنی و مهندسی شرق گیلان
2- دانشکده فنی مهندسی شرق گیلان/گیلان
3- دانشکده فنی مهندسی شرق گیلان/گیلان
کلمات کلیدی :
Hallucination Detection،Retrieval-Augmented Generation،Model Trustworthiness،Factuality Verification,،Mitigation Strategies
چکیده :
Large Language Models (LLMs) have changed how we process natural language. Still, they have a major problem: hallucination. This means they can create believable text that is wrong or not based on evidence. This paper looks at this issue by sorting hallucinations into factual, logical, intrinsic, and extrinsic types. We look at the reasons why this happens. These reasons include the statistics used to train these models, weaknesses that allow data poisoning, and how tests reward guessing instead of showing uncertainty. The paper looks at ways to spot hallucinations, such as self-checking, checking against known information, and using neuro-symbolic methods. It also reviews ways to reduce hallucinations, including using retrieval-augmented generation, prompt engineering, parameter-efficient fine-tuning, and specific decoding techniques. Our study shows that while theory says we can't completely get rid of hallucinations, we can greatly reduce them by using many methods together. This will help create more reliable AI systems.
لیست مقالات
لیست مقالات بایگانی شده
A Comparative Analysis of SDN Controllers for Industrial Internet of Things Applications
Ahmad Jalili - Habibollah Agh Atabay
Temporal–Spatial Graph- Integrated Framework for EEG- Based Emotion Recognition Using LSTM–GCN Architecture
Zahra Amiri - Abdorreza Hesam Mohseni
مروری بر کاربرد مدلهای هوش مصنوعی در پیشبینی شاخصهای خشکسالی هواشناسی
سید محمد سجاد عمرانی - محمد نجف زاده - صدیقه انوری
مدل سازی معکوس داده های ژئوالکتریکی با استفاده از روش MCMC
زهرا تفقد خباز - رضا قناتی - سید محمود طاهری - سید مرتضی امینی
تشخیص کیفیت برگ سبز چای بهکمک یادگیری عمیق
علی اسدی - میلاد بهنیا - کامراد خوشحال رودپشتی - محسن فلاح راد
برنامه نویسی ژنتیک چند منظوره جهت تشخیص به موقع بیماری سکته مغزی
سحر فقیهی راد - سیده نفیسه آل محمد
تولیدات پایدار در زنجیرهتامین با در نظرگیری چندین کارخانه
ندا کریمی
Investigation of Different Artificial Intelligence Algorithms in Predicting the Bending Strength of Reinforced Concrete Beams
Shokoohozaman Chamanzari - Aliaskar Dorostkar - Abolfazl Yosefi - HamidReza Nasseri
مروری بر کاربرد تحلیل درخت عیب فازی در مهندسی ایمنی و قابلیت اطمینان
روح اله رمضانی - محمدرضا ربیعی
مدلسازی فازی برای زنجیره بحرانی پروژه با ارائه الگوریتم محاسبه بافر با توجه به محدودیتهای منابع
عرفان کریم خانی میانجی - سید میثم موسوی - محمد فرهمندمهر
بیشتر
ثمین همایش، سامانه مدیریت کنفرانس ها و جشنواره ها - نگارش 44.9.4