← all repositories

EvolvingLMMs-Lab/open-r1-multimodal

A fork of open-r1 that adds multimodal RL training support for vision-language models using the GRPO algorithm.

1.6k stars Python Language ModelsML Frameworks
open-r1-multimodal
Not currently ranked — collecting fresh signals.
star history

This repository extends the open-r1 project to support multimodal reasoning model training. It implements the GRPO (Group Relative Policy Optimization) algorithm for training vision-language models like Qwen2-VL and Aria-MoE on math reasoning tasks. The project provides open-sourced training datasets with reasoning paths and verifiable answers, trained model checkpoints, and scripts for creating custom multimodal reasoning data.

Frequently asked

What is EvolvingLMMs-Lab/open-r1-multimodal?
A fork of open-r1 that adds multimodal RL training support for vision-language models using the GRPO algorithm.
Is open-r1-multimodal open source?
Yes — EvolvingLMMs-Lab/open-r1-multimodal is open source, released under the Apache-2.0 license.
What language is open-r1-multimodal written in?
EvolvingLMMs-Lab/open-r1-multimodal is primarily written in Python.
How popular is open-r1-multimodal?
EvolvingLMMs-Lab/open-r1-multimodal has 1.6k stars on GitHub.
Where can I find open-r1-multimodal?
EvolvingLMMs-Lab/open-r1-multimodal is on GitHub at https://github.com/EvolvingLMMs-Lab/open-r1-multimodal.

heatdrop uses Google Analytics to see which pages get read — nothing else. Your call. How we handle data.