Mobile QR Code QR CODE

2025

Reject Ratio

81.5%

References

1 
Z. Chen , Y. Duan , W. Wang , J. He , T. Lu , J. Dai , Y. Qiao , Vision transformer adapter for dense predictions, arXiv preprint arXiv:2205.08534, 2022DOI
2 
M. Cordts , M. Omran , S. Ramos , T. Rehfeld , M. Enzweiler , R. Benenson , U. Franke , S. Roth , B. Schiele , The Cityscapes dataset for semantic urban scene understanding, Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition, pp. 3213-3223, 2016DOI
3 
J.-Y. Jung , S.-H. Lee , J.-O. Kim , Knowledge transfer based spatial embedding network for plant leaf instance segmentation, IEIE Transactions on Smart Processing and Computing, Vol. 12, No. 2, pp. 162-170, 2023DOI
4 
W. Kim , B. Son , I. Kim , ViLT: Vision-and-language transformer without convolution or region supervision, Proceedings of the 38th International Conference on Machine Learning, Vol. 139, pp. 5583-5594, 2021DOI
5 
A. Kirillov , E. Mintun , N. Ravi , H. Mao , C. Rolland , L. Gustafson , T. Xiao , S. Whitehead , A. C. Berg , W.-Y. Lo , P. Dollár , R. Girshick , Segment anything, Proceedings of the IEEE/CVF International Conference on Computer Vision, pp. 4015-4026, 2023DOI
6 
B. Li , K. Q. Weinberger , S. Belongie , V. Koltun , R. Ranftl , Language-driven semantic segmentation, arXiv preprint arXiv:2201.03546, 2022DOI
7 
J. Li , D. Li , C. Xiong , S. C. H. Hoi , BLIP: Bootstrapping language-image pre-training for unified vision-language understanding and generation, Proceedings of the 39th International Conference on Machine Learning, Vol. 162, pp. 12888-12900, 2022DOI
8 
J. Li , R. R. Selvaraju , A. D. Gotmare , S. Joty , C. Xiong , S. C. H. Hoi , Align before fuse: Vision and language representation learning with momentum distillation, Advances in Neural Information Processing Systems, Vol. 34, pp. 9694-9705, 2021DOI
9 
Z. Liu , H. Mao , C.-Y. Wu , C. Feichtenhofer , T. Darrell , S. Xie , A ConvNet for the 2020s, Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition, pp. 11976-11986, 2022DOI
10 
T. Lüddecke , A. Ecker , Image segmentation using text and image prompts, Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition, pp. 7086-7096, 2022DOI
11 
G.-Y. Moon , J.-O. Kim , RoI-attention network for small disease segmentation in crop images, IEEE Access, Vol. 12, pp. 63725-63734, 2024DOI
12 
A. Radford , J. W. Kim , C. Hallacy , A. Ramesh , G. Goh , S. Agarwal , G. Sastry , A. Askell , P. Mishkin , J. Clark , G. Krueger , I. Sutskever , Learning transferable visual models from natural language supervision, Proceedings of the 38th International Conference on Machine Learning, Vol. 139, pp. 8748-8763, 2021DOI
13 
Y. Rao , W. Zhao , G. Chen , Y. Tang , Z. Zhu , G. Huang , J. Zhou , J. Lu , DenseCLIP: Language-guided dense prediction with context-aware prompting, Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition, pp. 18082-18091, 2022DOI
14 
W. Wang , J. Dai , Z. Chen , Z. Huang , Z. Li , X. Zhu , X. Hu , T. Lu , L. Lu , H. Li , X. Wang , Y. Qiao , InternImage: Exploring large-scale vision foundation models with deformable convolutions, Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition, pp. 14408-14419, 2023DOI
15 
T. Xiao , Y. Liu , B. Zhou , Y. Jiang , J. Sun , Unified perceptual parsing for scene understanding, Proceedings of the European Conference on Computer Vision, pp. 418-434, 2018DOI
16 
E. Xie , W. Wang , Z. Yu , A. Anandkumar , J. M. Alvarez , P. Luo , SegFormer: Simple and efficient design for semantic segmentation with transformers, Advances in Neural Information Processing Systems, Vol. 34, pp. 12077-12090, 2021DOI
17 
W. Yu , P. Zhou , S. Yan , X. Wang , InceptionNeXt: When inception meets ConvNeXt, Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition, pp. 5672-5683, 2024DOI
18 
K. Zhou , J. Yang , C. C. Loy , Z. Liu , Conditional prompt learning for vision-language models, Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition, pp. 16816-16825, 2022DOI
19 
K. Zhou , J. Yang , C. C. Loy , Z. Liu , Learning to prompt for vision-language models, International Journal of Computer Vision, Vol. 130, No. 9, pp. 2337-2348, 2022DOI