A study conducted by Mass General Brigham in Massachusetts, published in Jama Network Open, highlighted that common chatbots, such as ChatGPT and Gemini, make incorrect diagnoses in over 80% of cases when they do not have sufficient information about patients. The study tested 21 language models, including those developed by OpenAI, Anthropic, Google, xAI, and DeepSeek, using 29 clinical vignettes based on reference medical texts. Even when they had access to complete information, the chatbots had an error rate of over 40%, although in some cases they managed to provide the correct diagnosis for 90% of patients. The conclusion of the experiments is that the performance of these models significantly depends on the volume of available information, and hallucinations, that is, the invention of information, remain a major problem in generating correct responses.
Sources
Latest News
07:33
Vincent Pastore, an actor known for his role in 'The Sopranos', has died at the age of 80.
07:18
Donald Trump: The USA and Israel are ready to cancel attacks if Iran completely opens the Strait of Hormuz and renounces the nuclear threat.
07:17
Sever Voinescu: Temptation and tendency
22:50
The Danube is approaching a new historic minimum. Romanian Waters warns: the next two weeks are critical.
22:10
Devastating explosion in a cafe in Moscow: three dead and 15 injured
See more news