Researchers are gathering today to tackle AI deception. The workshop builds on Chris Cundy's finding: high-quality lie detectors can cut deception ~50%, but weak ones backfire. Models can learn to evade rather than become honest.๐
0 likes 1 replies
?