TY - RPRT TI - A closed-loop reinforcement learning framework for rapid compound directed optimization AU - Wang, H. AU - Lu, D. AU - Lyu, W. AU - Xiu, S. AU - Shi, C. AU - Zhou, X. AU - Xi, B. AU - Feng, W. AU - Xiao, Y. AU - Chen, Y. AU - Zhang, H. AU - Li, Q. AU - Huang, B. AU - Liu, Z. PY - 2026 DO - 10.64898/2026.08.24.745890 UR - https://www.biorxiv.org/content/10.64898/2026.08.24.745890v1 ID - 10.64898/2026.08.24.745890 ER -