Opening book details…
Can I read Self-Improvement in Multimodal Large Language Models: A Survey on EtoBox?
Self-Improvement in Multimodal Large Language Models: A Survey by Shijian Deng & Kai Wang & Tianyu Yang & Harsh Singh & Yapeng Tian is a book available to read on EtoBox.
What is Self-Improvement in Multimodal Large Language Models: A Survey about?
AbstractarXiv:2510.02665v1 [cs.CL] 3 Oct 2025Recent advancements in self-improvement forLarge Language Models (LLMs) have efficiently enhanced model capabilities withoutsignificantly increasing costs, particularly interms of human effort. While this area isstill relatively young, its extension to the multimodal domain holds immense potential forleveraging diverse data sources and developing more general self-improving models. Thissurvey is the first to provide a comprehensiveFiguroverview of self-improvement in MultimodalmodaLLMs (MLLMs). We provide a structuredselectoverview of the current literature and discussit intomethods from three perspectives: 1) data colcolleclection, 2) data organization, and 3) model opthroutimization, to facilitate the further developmentrecurof self-improvement in MLLMs. We also include commonly used evaluations and downstream applications. Finally, we conclude byoutlining open challenges and future researchsingldirections.self-iet al.1 Introductionin texeralSelf-improvement aims to enable models to colet allect and organize data required to build a betternatiogeneration of themselves, which offers a path tosive sovercome the costly scaling iss
- Author
- Shijian Deng & Kai Wang & Tianyu Yang & Harsh Singh & Yapeng Tian
- Published
- 2025
- Language
- EN
More by Shijian Deng & Kai Wang & Tianyu Yang & Harsh Singh & Yapeng Tian
Browse all works by Shijian Deng & Kai Wang & Tianyu Yang & Harsh Singh & Yapeng Tian
Similar books
- Document Intelligence in the Era of Large Language Models: A Survey — Weishi Wang & Hengchang Hu & Zhijie Zhang & Zhaochen Li & Hongxin Shao & Daniel Dahlmeier (2025)
- Understanding Large Language Models : Learning Their Underlying Concepts and Technologies — Thimira Amaratunga (2023)
- A Survey on Agentic Multimodal Large Language Models — Huanjin Yao & Ruifei Zhang & Jiaxing Huang & Jingyi Zhang & Yibo Wang & Bo Fang & Ruolin Zhu & Yongcheng Jing & Shunyu Liu & Guanbin Li & Dacheng Tao (2025)
- A Survey of Reinforcement Learning for Large Reasoning Models — Kaiyan Zhang & Yuxin Zuo & Bingxiang He & Youbang Sun & Runze Liu & Che Jiang & Yuchen Fan & Kai Tian & Guoli Jia & Pengfei Li & Yu Fu & Xingtai Lv & Yuchen Zhang & Sihang Zeng & Shang Qu & Haozhan Li & Shijie Wang & Yuru Wang & Xinwei Long & Fangfu... (2025)
- Multimodal Large Language Models Meet Multimodal Emotion Recognition and Reasoning: A Survey — Yuntao Shou & Tao Meng & Wei Ai & Keqin Li (2025)
- Visual Cognition in Multimodal Large Language Models — Luca M. Schulze Buschoff & Elif Akata & Matthias Bethge & Eric Schulz (2025)