Safetensors
qwen2
web-agent
webarena
world-model
rlvr

See https://github.com/thuml/RLVR-World for examples for using this model.

Citation

@article{wu2025rlvr,
    title={RLVR-World: Training World Models with Reinforcement Learning}, 
    author={Jialong Wu and Shaofeng Yin and Ningya Feng and Mingsheng Long},
    journal={arXiv preprint arXiv:2505.13934},
    year={2025},
}
Downloads last month
3
Safetensors
Model size
1.78B params
Tensor type
BF16
·
Inference Providers NEW
This model isn't deployed by any Inference Provider. 🙋 Ask for provider support

Model tree for thuml/webarena-world-model-rlvr

Dataset used to train thuml/webarena-world-model-rlvr

Collection including thuml/webarena-world-model-rlvr