RLVR-World
Collection
14 items
•
Updated
•
1
See https://github.com/thuml/RLVR-World for examples for using this model.
@article{wu2025rlvr,
title={RLVR-World: Training World Models with Reinforcement Learning},
author={Jialong Wu and Shaofeng Yin and Ningya Feng and Mingsheng Long},
journal={arXiv preprint arXiv:2505.13934},
year={2025},
}
Base model
deepseek-ai/DeepSeek-R1-Distill-Qwen-1.5B