AI Resonance · 阅读最新日报 · AI 入门推荐

2026-10-01 · America/Los_Angeles · 论文 · #7

UniEvo-VL: An On-policy Self-Distillation Training Recipe for Multimodal Model Self-improvement

Modern multimodal models bring generation and understanding into a single unified system, which enables them to provide and learn from their own feedback. Motivated by this unified capacity, we introduce UniEvo-VL, a self-evolving framework for multimodal models to learn from this constructive self-correction feedback during test-time compute. Instead of relying on a separate, often larger, teacher, we leverage their self-critiques as privileged information and ask a single multimodal model to act as both teacher and student with different contexts. The student only sees the vanilla…

UniEvo-VL: An On-policy Self-Distillation Training Recipe for Multimodal Model Self-improvement

热度 55.9 / 100;排名与评分保留该期记录。