Von Werra urges frontier AI labs to share small models and alignment recipes
Original titleIt's frustrating that the discussion about safety and pacing of AI progress is once again led by a handful of people when the implication...
AISummary
Hugging Face's Leandro von Werra argues that frontier AI labs should release small variants of their models, share core parts of their alignment recipe, and publish tech reports with more than evaluations.
He says these steps would let the wider community test model behavior and verify safety claims, rather than leaving the safety agenda to a few labs.
He also calls for independent verification of alarming internal findings, with sensitive details disclosed first to an independent team.
Source: Leandro von Werra · x.comPublished · added here