OSS Safety Models

by @julietshen.online

a feed to learn, discuss, and rant about open source safety models

Pull to refresh
Ethan Mollick @emollick.bsky.social · 1d
3

Randomized trials with GPT-4o: "AI access raises test scores, and a smaller gain persists a week later. Gains remain for students who use AI as a tutor (“augmentation”) and fade for students who have AI write for them (“automation”)" Across all experiments, positive effects arxiv.org/abs/2607.08849

Ai2 @ai2.bsky.social · 1d
2

We’re sharing the engineering work behind Ai2’s GPU scheduler—our new approach to allocating compute across research teams. On our largest H100 cluster, median queue wait fell from 5 minutes to 24 seconds. 🧵 buff.ly/4UsJEYe

Ethan Mollick @emollick.bsky.social · 2d
3

From The Lancet: in an urgent care setting, the advice of the obsolete Gemini 2.5 Pro & Gemini 2.5 Flash (without access to patient medical records) were rated of similar quality to doctors by other physicians. There was no safety issues spotted. Models have gotten significantly better since.

Ethan Mollick @emollick.bsky.social · 2d
6

As a researcher who did early some work on the productivity impacts of AI chatbots using RCTs, I’d note a lack of similar studies since the dawn of true agents last fall Partially that is newness & partially research design challenges, but I suspect we are missing some large & important effects

Ethan Mollick @emollick.bsky.social · 3d
4

Some early first-hand accounts of the experience of encountering a narrow superhuman intelligence as mathematicians grapple with the hundreds of big AI proofs released by OpenAI. Problems solved in inhuman ways that make us wonder what it means to actually know things... scottaaronson.blog?p=10169