AI reduces sensory hallucinations, even at night or in smoke
Multimodal large language models (MLLMs), which process multiple types of sensory information, such as text, images and audio, at the same time, are rapidly expanding the range of applications for artificial intelligence (AI). However, in real-world environments, these models can misinterpret the physical characteristics of sensors, mistakenly identify objects or claim to hear sounds that are not actually present simply because a certain object appears in a video. These errors are known as hallucinations.
