A new study tasked AIs with tackling the 'Stroop' test GPT and Claude performed very poorly compared to humans There are nuances here, but broadly, the researchers argue that improving this side of ...
AI text generators understandably fail a classic test from psychology and cannot correctly name colored color words when the two do not match. This was discovered by a US research team that had GPT-4o ...
Some results have been hidden because they may be inaccessible to you
Show inaccessible results