Flashes Home Search Notifications Sign in Post
Zsolt Ero @hyperknot.com · Nov 10

I've just read about the new math benchmark for AIs called FrontierMath and found this interesting graph on their homepage. Basically, it's saying that Sonnet 3.5 was much better compared to o1-preview, even though this area (solving hard math problems) is the specialty area of o1-preview.

1 likes 1 replies

?

Replies

Zsolt Ero · Nov 10

https://epochai.org/frontiermath/the-benchmark

Legal

Privacy Policy Terms and Conditions

Contact

FAQs Feedback

Follow

Bluesky Instagram Threads
Flashes for Bluesky Get it on the App Store

Feeds

  • Global
  • Trending
  • Blacksky Images
  • Photography
  • Discover
Browse all feeds
Feedback Ideas, bugs, and what ships next
Privacy Policy Terms and Conditions FAQs Bluesky Instagram Threads
Get it on the App Store