Search

Tag: #hallucinations34 results

South Africa AI Policy Has Fake Citations

South Africa AI Policy Has Fake Citations

South Africa’s Department of Communications and Digital Technologies used AI to draft its national AI policy over months. The draft proposes a National AI Commission, AI Ethics Board, and other bodies across five governance pillars. However, the policy included fabricated citations.

The Next Web (TNW)MediaApr 28#ai-policy#fake-citations#government
Steering Multimodal AI Hallucination Verifiability

Steering Multimodal AI Hallucination Verifiability

Researchers built a dataset from 4,470 human responses to categorize MLLM hallucinations as obvious or elusive based on verifiability. They developed activation-space interventions using separate probes for each type, enabling precise control over hallucination detectability. Results show effective tuning for diverse security and usability needs.

ArXiv AIResearchApr 10#hallucinations#verifiability
Geometric Taxonomy of LLM Hallucinations

Geometric Taxonomy of LLM Hallucinations

Researchers propose a geometric taxonomy classifying LLM hallucinations into three types: unfaithfulness, confabulation, and factual error. Benchmark hallucinations show strong domain-local detection but fail cross-domain, while human-crafted confabulations enable a single global detection direction. Factual errors remain undetectable via embeddings due to distributional encoding limits.

ArXiv AIResearchFeb 17#research#llms#hallucinations
Page 1 of 4