Nick Huntington-Klein @nickchk.com · Sep 13

The new model GPT-o1, dedicated to advanced reasoning, also fails at my "three card blind" task (skipping some spots in a very lengthy response, where it carefully weighs each option and improperly evaluates the correct option, coming to the wrong conclusion)

13 likes 2 replies

?

Replies

post malone ergo propter malone · Sep 13

how's it do with Forchard/Brand?

Nick Huntington-Klein · Sep 13

It does manage to get this (much easier) one now though!