File size: 488 Bytes
17a500d 6121292 17a500d 6121292 |
1 2 3 4 5 6 7 8 9 10 11 12 13 14 15 16 17 18 19 20 21 22 |
---
license: mit
tags:
- text-game
- world-model
datasets:
- thuml/bytesized32-world-model-cot
base_model:
- deepseek-ai/DeepSeek-R1-Distill-Qwen-1.5B
---
See https://github.com/thuml/RLVR-World for examples for using this model.
## Citation
```
@article{wu2025rlvr,
title={RLVR-World: Training World Models with Reinforcement Learning},
author={Jialong Wu and Shaofeng Yin and Ningya Feng and Mingsheng Long},
journal={arXiv preprint arXiv:2505.13934},
year={2025},
} |