Papers
Communities
Events
Blog
Pricing
Search
Open menu
Home
Papers
2408.01120
Cited By
An Efficient and Effective Transformer Decoder-Based Framework for Multi-Task Visual Grounding
2 August 2024
Wei Chen
Mahdieh Hatamian
Yu Wu
Re-assign community
ArXiv
PDF
HTML
Papers citing
"An Efficient and Effective Transformer Decoder-Based Framework for Multi-Task Visual Grounding"
4 / 4 papers shown
Title
LAVT: Language-Aware Vision Transformer for Referring Image Segmentation
Zhao Yang
Jiaqi Wang
Yansong Tang
Kai-xiang Chen
Hengshuang Zhao
Philip H. S. Torr
133
306
0
04 Dec 2021
Open-vocabulary Object Detection via Vision and Language Knowledge Distillation
Xiuye Gu
Tsung-Yi Lin
Weicheng Kuo
Yin Cui
VLM
ObjD
223
897
0
28 Apr 2021
Multi-task Collaborative Network for Joint Referring Expression Comprehension and Segmentation
Gen Luo
Yiyi Zhou
Xiaoshuai Sun
Liujuan Cao
Chenglin Wu
Cheng Deng
Rongrong Ji
ObjD
159
286
0
19 Mar 2020
A Real-Time Cross-modality Correlation Filtering Method for Referring Expression Comprehension
Yue Liao
Si Liu
Guanbin Li
Fei-Yue Wang
Yanjie Chen
Chao Qian
Bo-wen Li
ObjD
62
199
0
16 Sep 2019
1