Benchmarks keep going up but the models don't feel better. @steveklabnik.com and @skriptble.com on why "10% improvement" means nothing if using it still feels bad, and how some open models post frontier numbers by overtraining on the benchmarks. fallthrough.fm/ep/70
0 likes 0 replies
?