TencentARC/GRPO-CARE

🤗 Hugging Face 来源video-text-to-textapache-2.050 GBsafetensors✓ 18 个校验和今天更新
一条命令提交

在你的模型文件夹旁边运行它。它会制作种子、将文件与 Hugging Face 比对,然后提交。你只需开始做种,并粘贴你账户中的密钥。它只读取你的文件,绝不修改。如果愿意,可以先阅读脚本。

curl -fsSL https://pirateface.co/package.sh | bash -s -- --repo TencentARC/GRPO-CARE ./model-folder
需要做种者 →

This repository contains the GRPO-CARE model, presented in the paper GRPO-CARE: Consistency-Aware Reinforcement Learning for Multimodal Reasoning.

Code released at GRPO-CARE.

Citation

@misc{chen2025grpocareconsistencyawarereinforcementlearning,
  title={GRPO-CARE: Consistency-Aware Reinforcement Learning for Multimodal Reasoning}, 
  author={Yi Chen and Yuying Ge and Rui Wang and Yixiao Ge and Junhao Cheng and Ying Shan and Xihui Liu},
  year={2025},
  eprint={2506.16141},
  archivePrefix={arXiv},
  primaryClass={cs.CV},
  url={https://arxiv.org/abs/2506.16141}, 
}