RL needs more long-horizon tasks beyond math and coding. @niloofar_mire's team at CMU built an environment for drug design.
SMDD-Bench comprises 502 small-molecule design tasks with RDKit, ADMET-AI and Boltz-2 in the loop. The challenges go beyond chemistry: long-horizon planning, exploration, and learning from imperfect feedback are also open problems for RL/ML!
SMDD-Bench is available in our Environments Hub, ready to train with prime-rl. Thank you for sharing with the community!
