Target Detection and Segmentation Technology for Zero-shot Learning

Authors

  • Zongzhi Lou
  • Linlin Chen
  • Tian Guo
  • Zhizhong Wang
  • Yuxuan Qiu
  • Jinyang Liang

DOI:

https://doi.org/10.54097/v7tbh549

Keywords:

Zero-shot Learning, Target Detection, Image Segmentation, Reinforcement Learning, Computer Vision

Abstract

Zero-shot learning (ZSL) in the field of computer vision refers to enabling the model to recognize and understand categories that have not been encountered during the training phase. It is particularly critical for object detection and segmentation tasks, because these tasks require the model to have good generalization capabilities to unknown categories. Object detection requires the model to determine the location of the object, while segmentation further requires the precise demarcation of the object's boundaries. In ZSL research, knowledge representation and transfer are core issues. Researchers have tried to use semantic attributes as a knowledge bridge to connect categories seen during the training phase and categories not seen during the testing phase. These attributes may be color, shape, etc., but this method requires accurate attribute annotation, which is often not easy to achieve in practice. Therefore, researchers have begun to explore the use of non-visual information such as knowledge maps and text descriptions to enrich the recognition capabilities of models, but this also introduces the challenge of information integration and alignment. At present, ZSL has made certain progress in target detection and segmentation tasks, but there is still a significant gap compared with traditional supervised learning. This is mainly due to the limited ability of ZSL models to generalize to new categories. To this end, researchers have begun to explore combining ZSL with other technologies, such as generative adversarial networks (GANs) and reinforcement learning, to enhance the model's detection and segmentation capabilities for new categories. Future research needs to focus on several aspects. The first is how to design a more effective knowledge representation and transfer mechanism so that the model can better utilize existing knowledge. The second step is to develop new algorithms to improve the performance of ZSL in complex environments. In addition, research should focus on how to reduce the dependence on computing resources so that the ZSL method can run effectively in resource-limited environments. In summary, the research on target detection and segmentation technology of zero-shot learning is a cutting-edge topic in the field of computer vision. Despite the challenges, with the deepening of research, we expect these technologies to contribute to improving the generalization ability and intelligence level of computer vision systems.

Downloads

Download data is not yet available.

References

Ren, W., Tang, Y., Sun, Q., Zhao, C., & Han, Q. L. (2023). Visual semantic segmentation based on few/zero-shot learning: An overview. IEEE/CAA Journal of Automatica Sinica.

Lu, X., Wang, W., Shen, J., Crandall, D., & Luo, J. (2020). Zero-shot video object segmentation with co-attention siamese networks. IEEE transactions on pattern analysis and machine intelligence, 44(4), 2228-2242.

Dong, Y., Jiang, X., Zhou, H., Lin, Y., & Shi, Q. (2021). SR2CNN: Zero-shot learning for signal recognition. IEEE Transactions on Signal Processing, 69, 2316-2329.

Bian, C., Yuan, C., Ma, K., Yu, S., Wei, D., & Zheng, Y. (2021). Domain adaptation meets zero-shot learning: an annotation-efficient approach to multi-modality medical image segmentation. IEEE Transactions on Medical Imaging, 41(5), 1043-1056.

Li, P., Wei, Y., & Yang, Y. (2020). Consistent structural relation learning for zero-shot segmentation. Advances in Neural Information Processing Systems, 33, 10317-10327.

Lv, F., Liu, H., Wang, Y., Zhao, J., & Yang, G. (2020). Learning unbiased zero-shot semantic segmentation networks via transductive transfer. IEEE Signal Processing Letters, 27, 1640-1644.

Gu, Z., Zhou, S., Niu, L., Zhao, Z., & Zhang, L. (2022). From pixel to patch: Synthesize context-aware features for zero-shot semantic segmentation. IEEE Transactions on Neural Networks and Learning Systems.

Zhang, Z., Liu, Q., Qiu, S., Zhou, S., & Zhang, C. (2020). Unknown attack detection based on zero-shot learning. IEEE Access, 8, 193981-193991.

Liu, R., Wu, Z., Yu, S., & Lin, S. (2021). The emergence of objectness: Learning zero-shot segmentation from videos. Advances in Neural Information Processing Systems, 34, 13137-13152.

Zhou, T., Li, J., Wang, S., Tao, R., & Shen, J. (2020). Matnet: Motion-attentive transition network for zero-shot video object segmentation. IEEE Transactions on Image Processing, 29, 8326-8338.

Xie, G. S., Zhang, Z., Xiong, H., Shao, L., & Li, X. (2022). Towards zero-shot learning: A brief review and an attention-based embedding network. IEEE Transactions on Circuits and Systems for Video Technology.

Li, H., Feng, C. M., Xu, Y., Zhou, T., Yao, L., & Chang, X. (2023). Zero-shot camouflaged object detection. IEEE Transactions on Image Processing.

Li, C., Ye, X., Cao, D., Hou, J., & Yang, H. (2021). Zero shot objects classification method of side scan sonar image based on synthesis of pseudo samples. Applied Acoustics, 173, 107691.

Xi, J., Ye, X., & Li, C. (2022). Sonar Image Target Detection Based on Style Transfer Learning and Random Shape of Noise under Zero Shot Target. Remote Sensing, 14(24), 6260.

Li, A., Qiu, C., Kloft, M., Smyth, P., Rudolph, M., & Mandt, S. (2024). Zero-shot anomaly detection via batch normalization. Advances in Neural Information Processing Systems, 36.

Shi, P., Qiu, J., Abaxi, S. M. D., Wei, H., Lo, F. P. W., & Yuan, W. (2023). Generalist vision foundation models for medical imaging: A case study of segment anything model on zero-shot medical segmentation. Diagnostics, 13(11), 1947.

Shin, G., Xie, W., & Albanie, S. (2022). Reco: Retrieve and co-segment for zero-shot transfer. Advances in Neural Information Processing Systems, 35, 33754-33767.

Zhou, T., Wang, S., Zhou, Y., Yao, Y., Li, J., & Shao, L. (2020, April). Motion-attentive transition for zero-shot video object segmentation. In Proceedings of the AAAI conference on artificial intelligence (Vol. 34, No. 07, pp. 13066-13073).

Downloads

Published

11-03-2024

Issue

Section

Articles

How to Cite

Lou, Z., Chen, L., Guo, T., Wang, Z., Qiu, Y., & Liang, J. (2024). Target Detection and Segmentation Technology for Zero-shot Learning. Frontiers in Computing and Intelligent Systems, 7(2), 38-42. https://doi.org/10.54097/v7tbh549