A study conducted by Mass General Brigham in Massachusetts, published in Jama Network Open, highlighted that common chatbots, such as ChatGPT and Gemini, make incorrect diagnoses in over 80% of cases when they do not have sufficient information about patients. The study tested 21 language models, including those developed by OpenAI, Anthropic, Google, xAI, and DeepSeek, using 29 clinical vignettes based on reference medical texts. Even when they had access to complete information, the chatbots had an error rate of over 40%, although in some cases they managed to provide the correct diagnosis for 90% of patients. The conclusion of the experiments is that the performance of these models significantly depends on the volume of available information, and hallucinations, that is, the invention of information, remain a major problem in generating correct responses.
Sources
Latest News
20:22
Railway traffic was temporarily halted between Simeria and Orăștie, after the discovery of improvised electrical installations.
20:02
Greece has moved a Patriot anti-aircraft battery to Crete for military defense against the risks of Iranian attacks on American installations.
19:41
Conflict at Amzei Pastry Shop: A man in a state of intoxication attempted to assault a woman, who defended herself with pepper spray.
19:18
Lando Norris won the Dutch Grand Prix, bringing a third consecutive victory for McLaren
19:00
China postpones the Chang'e-7 mission for the exploration of the Moon's south pole.
See more news