1 / 2
The new model GPT-o1, dedicated to advanced reasoning, also fails at my "three card blind" task (skipping some spots in a very lengthy response, where it carefully weighs each option and improperly evaluates the correct option, coming to the wrong conclusion)
13 likes 2 replies
?