Seuraa
Jianfeng Wang
Jianfeng Wang
Vahvistettu sähköpostiosoite verkkotunnuksessa microsoft.com
Nimike
Viittaukset
Viittaukset
Vuosi
Florence: A new foundation model for computer vision
L Yuan, D Chen, YL Chen, N Codella, X Dai, J Gao, H Hu, X Huang, B Li, ...
arXiv preprint arXiv:2111.11432, 2021
8902021
Improving Image Generation with Better Captions
J Betker, G Goh, L Jing, T Brooks, J Wang, L Li, L Ouyang, J Zhuang, ...
6622023
Multimedia cloud computing
W Zhu, C Luo, J Wang, S Li
IEEE Signal Processing Magazine 28 (3), 59-69, 2011
5992011
End-to-end semi-supervised object detection with soft teacher
M Xu, Z Zhang, H Hu, J Wang, L Wang, F Wei, X Bai, Z Liu
Proceedings of the IEEE/CVF international conference on computer vision …, 2021
5492021
Git: A generative image-to-text transformer for vision and language
J Wang, Z Yang, X Hu, L Li, K Lin, Z Gan, Z Liu, C Liu, L Wang
arXiv preprint arXiv:2205.14100, 2022
5292022
The dawn of lmms: Preliminary explorations with gpt-4v (ision)
Z Yang, L Li, K Lin, J Wang, CC Lin, Z Liu, L Wang
arXiv preprint arXiv:2309.17421 9 (1), 1, 2023
5102023
An empirical study of gpt-3 for few-shot knowledge-based vqa
Z Yang, Z Gan, J Wang, X Hu, Y Lu, Z Liu, L Wang
Proceedings of the AAAI conference on artificial intelligence 36 (3), 3081-3089, 2022
4042022
Mm-vet: Evaluating large multimodal models for integrated capabilities
W Yu, Z Yang, L Li, J Wang, K Lin, Z Liu, X Wang, L Wang
arXiv preprint arXiv:2308.02490, 2023
3972023
An empirical study of training end-to-end vision-and-language transformers
ZY Dou, Y Xu, Z Gan, J Wang, S Wang, L Wang, C Zhu, P Zhang, L Yuan, ...
Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern …, 2022
3852022
Mm-react: Prompting chatgpt for multimodal reasoning and action
Z Yang, L Li, J Wang, K Lin, E Azarnasab, F Ahmed, Z Liu, C Liu, M Zeng, ...
arXiv preprint arXiv:2303.11381, 2023
3182023
Scaling up vision-language pre-training for image captioning
X Hu, Z Gan, J Wang, Z Yang, Z Liu, Y Lu, L Wang
Proceedings of the IEEE/CVF conference on computer vision and pattern …, 2022
2882022
Prompting gpt-3 to be reliable
C Si, Z Gan, Z Yang, S Wang, J Wang, J Boyd-Graber, L Wang
arXiv preprint arXiv:2210.09150, 2022
2362022
Generalized decoding for pixel, image, and language
X Zou, ZY Dou, J Yang, Z Gan, L Li, C Li, X Dai, H Behl, J Wang, L Yuan, ...
Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern …, 2023
2272023
Seed: Self-supervised distillation for visual representation
Z Fang, J Wang, L Wang, L Zhang, Y Yang, Z Liu
arXiv preprint arXiv:2101.04731, 2021
2232021
Aligning large multi-modal model with robust instruction tuning
F Liu, K Lin, L Li, J Wang, Y Yacoob, L Wang
arXiv preprint arXiv:2306.14565, 2023
1892023
Tap: Text-aware pre-training for text-vqa and text-caption
Z Yang, Y Lu, J Wang, X Yin, D Florencio, L Wang, C Zhang, L Zhang, ...
Proceedings of the IEEE/CVF conference on computer vision and pattern …, 2021
1732021
Reco: Region-controlled text-to-image generation
Z Yang, J Wang, Z Gan, L Li, K Lin, C Wu, N Duan, Z Liu, C Liu, M Zeng, ...
Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern …, 2023
1292023
Anchor box optimization for object detection
Y Zhong, J Wang, J Peng, L Zhang
Proceedings of the IEEE/CVF Winter Conference on Applications of Computer …, 2020
1292020
Order preserving hashing for approximate nearest neighbor search
J Wang, J Wang, N Yu, S Li
Proceedings of the 21st ACM international conference on Multimedia, 133-142, 2013
1272013
Optimized cartesian k-means
J Wang, J Wang, J Song, XS Xu, HT Shen, S Li
IEEE Transactions on Knowledge and Data Engineering 27 (1), 180-192, 2014
1252014
Järjestelmä ei voi suorittaa toimenpidettä nyt. Yritä myöhemmin uudelleen.
Artikkelit 1–20