Visionllm: Large language model is also an open-ended decoder for vision-centric tasks

Wenhai Wang, Zhe Chen, Xiaokang Chen, Jiannan Wu, Xizhou Zhu, Gang Zeng, Ping Luo, Tong Lu, Jie Zhou, Yu Qiao, Jifeng Dai

September 2023

PDF

Type

Conference paper

Publication

Thirty-seventh Annual Conference on Neural Information Processing Systems (NeurIPS) 2023