AI agents are significantly more dangerous than chatbots because they act autonomously; new detection methods like Finch-Zk and LettuceDetect show improvements but cannot fully prevent hallucinations.
NeuroCogMap maps the internal representations of LLMs onto functional systems, mechanistically identifies failure patterns such as hallucinations and bias, and simultaneously improves prediction of human brain activity.