A recent experiment revealed that current AI models are unreliable for identifying edible mushrooms, posing a significant risk to users. Piotr Migdał tested 16 AI models, including ChatGPT, Qwen, and Gemini-3.8-flash, using a dataset of Danish and Polish mushrooms. Even the best-performing model, Gemini-3.8-flash, correctly identified species only 65% of the time on its first guess, with a 35% chance of error. Some models, like Qwen3.8-27b, performed much worse, misidentifying poisonous mushrooms as edible with a 36% false positive rate, highlighting the potential for life-threatening mistakes. AI
IMPACT AI models are currently unreliable for critical identification tasks, posing safety risks and highlighting the need for caution with AI-generated advice.
RANK_REASON Research paper detailing an experiment on AI model capabilities. [lever_c_demoted from research: ic=1 ai=1.0]
- chanterelle
- ChatGPT
- death cap mushroom
- fatal dapperling
- fool's funnel
- Meta*
- Piotr Migdał
- Quesma
- Qwen
- Qwen3.8-27b
- WebCapsule
AI-generated summary · Google Gemini · from 1 sources. How we write summaries →