Command Palette
Search for a command to run...
R1-Onevision Multimodal Reasoning Dataset
Date
Size
Publish URL
Paper URL
License
Apache 2.0
Tags
The R1-Onevision dataset was released by Zhejiang University in 2025. It aims to give the model advanced multimodal reasoning capabilities. The relevant blog is "R1-Onevision: Open-Source Multimodal Large Language Model with Reasoning AbilityIt bridges the gap between visual and textual understanding by enabling rich, context-aware reasoning tasks in multiple domains, including natural scenes, science, math problems, OCR-based content, and complex diagrams.

Citation
If you find this code useful for your research, please use the following BibTeX entry. python @article{yang2025r1onevision, title={R1-Onevision: Advancing Generalized Multimodal Reasoning through Cross-Modal Formalization}, author={Yi Yang and Xiaoxuan He and Hongkun Pan and Xiyan Jiang and Yan Deng and Xingtao Yang and Haoyu Lu and Dacheng Yin and Fengyun Rao and Minfeng Zhu and Bo Zhang and Wei Chen}, journal={arXiv preprint arXiv:2503.10615}, year={2025}, }
Build AI with AI
From idea to launch — accelerate your AI development with free AI co-coding, out-of-the-box environment and best price of GPUs.