Papers
Communities
Events
Blog
Pricing
Search
Open menu
Home
Papers
2307.02280
Cited By
Interactive Image Segmentation with Cross-Modality Vision Transformers
5 July 2023
Kun Li
G. Vosselman
M. Yang
ViT
Re-assign community
ArXiv
PDF
HTML
Papers citing
"Interactive Image Segmentation with Cross-Modality Vision Transformers"
3 / 3 papers shown
Title
PseudoClick: Interactive Image Segmentation with Click Imitation
Qin Liu
Meng Zheng
Benjamin Planche
Srikrishna Karanam
Terrence Chen
Marc Niethammer
Ziyan Wu
VLM
43
57
0
12 Jul 2022
Pyramid Vision Transformer: A Versatile Backbone for Dense Prediction without Convolutions
Wenhai Wang
Enze Xie
Xiang Li
Deng-Ping Fan
Kaitao Song
Ding Liang
Tong Lu
Ping Luo
Ling Shao
ViT
316
3,625
0
24 Feb 2021
Unified Vision-Language Pre-Training for Image Captioning and VQA
Luowei Zhou
Hamid Palangi
Lei Zhang
Houdong Hu
Jason J. Corso
Jianfeng Gao
MLLM
VLM
252
927
0
24 Sep 2019
1