Page 70 - 《软件学报》2026年第5期
P. 70

李泽超 等: 基于   CLIP  引导标签优化的弱监督图像哈希                                                1949


                     31st Int’l Conf. on Neural Information Processing Systems. Long Beach: Curran Associates Inc., 2017. 6000–6010.
                 [24]   Huiskes MJ, Lew MS. The MIR Flickr retrieval evaluation. In: Proc. of the 1st ACM Int’l Conf. on Multimedia Information Retrieval.
                     Vancouver: ACM, 2008. 39–43. [doi: 10.1145/1460096.1460104]
                 [25]   Chua TS, Tang JH, Hong RC, Li HJ, Luo ZP, Zheng YT. NUS-WIDE: A real-world Web image database from National University of
                     Singapore. In: Proc. of the 2009 ACM Int’l Conf. on Image and Video Retrieval. Santorini: ACM, 2009. 48. [doi: 10.1145/1646396.
                     1646452]
                 [26]   Indyk P, Motwani R. Approximate nearest neighbors: Towards removing the curse of dimensionality. In: Proc. of the 30th Annual ACM
                     Symp. on Theory of Computing. Dallas: ACM, 1998. 604–613. [doi: 10.1145/276698.276876]
                 [27]   Weiss Y, Torralba A, Fergus R. Spectral hashing. In: Proc. of the 22nd Int’l Conf. on Neural Information Processing Systems. Vancouver:
                     Curran Associates Inc., 2008. 1753–1760.
                 [28]   Gong YC, Lazebnik S, Gordo A, Perronnin F. Iterative quantization: A procrustean approach to learning binary codes for large-scale
                     image retrieval. IEEE Trans. on Pattern Analysis and Machine Intelligence, 2013, 35(12): 2916–2929. [doi: 10.1109/TPAMI.2012.193]

                 附中文参考文献
                  [1]   黄小燕, 孙彬, 杨展源, 朱映映, 田奇. 面向视觉搜索的空间局部敏感哈希方法. 中国图象图形学报, 2021, 26(7): 1568–1582. [doi:
                     10.11834/jig.200534]
                  [2]   李志欣, 凌锋, 唐振军, 马慧芳, 施智平. 基于多头注意力网络的无监督跨媒体哈希检索. 中国科学: 信息科学, 2021, 51(12):
                     2053–2068. [doi: 10.1360/SSI-2020-0264]
                 [21]   殷炯, 张哲东, 高宇涵, 杨智文, 李亮, 肖芒, 孙垚棋, 颜成钢. 视觉语言预训练综述. 软件学报, 2023, 34(5): 2000–2023. http://www.jos.
                     org.cn/1000-9825/6774.htm [doi: 10.13328/j.cnki.jos.006774]
                 [22]   张浩宇, 王天保, 李孟择, 赵洲, 浦世亮, 吴飞. 视觉语言多模态预训练综述. 中国图象图形学报, 2022, 27(9): 2652–2682. [doi:
                     10.11834/jig.220173]

                 作者简介
                 李泽超, 博士, 教授, 博士生导师, CCF  杰出会员, 主要研究领域为海量媒体智能分析, 图像视频理解, 多模态大模型.
                 金露, 博士, 副教授, 主要研究领域为多媒体分析与理解, 多模态检索, 多模态大模型.
                 王浩骅, 硕士生, 主要研究领域为图像检索.
                 唐金辉, 博士, 教授, 博士生导师, CCF  杰出会员, 主要研究领域为智能媒体分析与理解, 图像视频理解与生成, 多模态大模型.
   65   66   67   68   69   70   71   72   73   74   75