Page 156 - 《软件学报》2026年第3期
P. 156

江宇轩 等: 权重残差向量量化: 向量压缩与分层索引结构                                                    1119


                     inference data. In: Proc. of the 2017 Conf. on Empirical Methods in Natural Language Processing. Copenhagen: ACL, 2017. 670–680.
                     [doi: 10.18653/v1/D17-1070]
                  [9]   Reimers N, Gurevych I. Sentence-BERT: Sentence embeddings using siamese BERT-networks. In: Proc. of the 2019 Conf. on Empirical
                     Methods  in  Natural  Language  Processing  and  the  9th  Int’l  Joint  Conf.  on  Natural  Language  Processing.  Hong  Kong:  ACL,  2019.
                     3982–3992. [doi: 10.18653/v1/D19-1410]
                 [10]   Muennighoff N, Tazi N, Magne L, Reimers N. MTEB: Massive text embedding benchmark. In: Proc. of the 17th Conf. of the European
                     Chapter of the Association for Computational Linguistics. Dubrovnik: ACL, 2022. 2014–2037. [doi: 10.18653/v1/2023.eacl-main.148]
                 [11]   Gao T, Yao X, Chen D. SimCSE: Simple contrastive learning of sentence embeddings. In: Proc. of the 2021 Conf. on Empirical Methods
                     in Natural Language Processing (EMNLP). Punta Cana: ACL, 2021. 6894–6910. [doi: 10.18653/v1/2021.emnlp-main.552]
                 [12]   Brown TB, Mann B, Ryder N, Subbiah M, Kaplan J, Dhariwal P, Neelakantan A, Shyam P, Sastry G, Askell A, Agarwal S, Herbert-Voss
                     A, Krueger G, Henighan T, Child R, Ramesh A, Ziegler DM, Wu J, Winter C, Hesse C, Chen M, Sigler E, Litwin M, Gray S, Chess B,
                     Clark J, Berner C, McCandlish S, Radford A, Sutskever I, Amodei D. Language models are few-shot learners. In: Proc. of the 34th Int’l
                     Conf. on Neural Information Processing Systems. Vancouver: Curran Associates Inc., 2020. 159.
                 [13]   Radford A, Kim JW, Hallacy C, Ramesh A, Goh G, Agarwal S, Sastry G, Askell A, Mishkin P, Clark J, Krueger G, Sutskever I. Learning
                     transferable visual models from natural language supervision. In: Proc. of the 38th Int’l Conf. on Machine Learning. 2021. 8784–8763.
                 [14]   Jie Tao, Lei Zhang, Ying Li, Zhi Chen. Optimizing LLM embeddings via prompting and data-driven tuning. In: Proc. of the 62nd Annual
                     Meeting of the Association for Computational Linguistics, Vol. 1 (Long Papers). Bangkok: ACL, 2024. 1234–1245.
                 [15]   Xiao ST, Liu Z, Zhang PT, Muennighoff N, Lian DF, Nie JY. C-Pack: Packed resources for general Chinese embeddings. In: Proc. of the
                     47th Int’l ACM SIGIR Conf. on Research and Development in Information Retrieval. Washington: ACM, 2024. 641–649. [doi: 10.1145/
                     3626772.3657878]
                 [16]   Yang W, Li T, Fang G, Wei H. PASE: PostgreSQL ultra-high-dimensional approximate nearest neighbor search extension. In: Proc. of
                     the 2020 ACM SIGMOD Int’l Conf. on Management of Data. Portland: ACM, 2020. 2241–2253. [doi: 10.1145/3318464.3386131]
                 [17]   Wei  CX,  Wu  B,  Wang  S,  Lou  RJ,  Zhan  CQ,  Li  FF,  Cai  YZ.  AnalyticDB-V:  A  hybrid  analytical  engine  towards  query  fusion  for
                     structured and unstructured data. Proc. of the VLDB Endowment, 2020, 13(12): 3152–3165. [doi: 10.14778/3415478.3415541]
                 [18]   Johnson J, Douze M, Jégou H. Billion-scale similarity search with GPUs. IEEE Trans. on Big Data, 2021, 7(3): 535–547. [doi: 10.1109/
                     TBDATA.2019.2921572]
                 [19]   Wang JG, Yi XM, Guo RT, Jin H, Xu P, Li SJ, Wang XY, Guo XZ, Li CM, Xu XH, Yu K, Yuan YX, Zou YH, Long JQ, Cai YD, Li ZX,
                     Zhang ZF, Mo YH, Gu J, Jiang RY, Wei Y, Xie C. Milvus: A purpose-built vector data management system. In: Proc. of the 2021 Int’l
                     Conf. on Management of Data. ACM, 2021. 2614–2627. [doi: 10.1145/3448016.3457550]
                 [20]   Chroma is the open-source search and retrieval database for AI applications. 2025. https://www.trychroma.com
                 [21]   Stonebraker M, Cetintemel U. “One size fits all”: An idea whose time has come and gone. In: Proc. of the 21st Int’l Conf. on Data
                     Engineering. Tokyo: IEEE, 2005. 2–11. [doi: 10.1109/ICDE.2005.1]
                 [22]   Dittrich J, Jindal A. Towards a one size fits all database architecture. In: Proc. of the 5th Biennial Conf. on Innovative Data Systems
                     Research. 2011. 195–198.
                 [23]   Zhang YN, Liu SG, Wang JG. Are there fundamental limitations in supporting vector data management in relational databases? A case
                     study of PostgreSQL. In: Proc. of the 40th IEEE Int’l Conf. on Data Engineering (ICDE). Utrecht: IEEE, 2024. 3640–3653. [doi: 10.1109/
                     ICDE60146.2024.00280]
                 [24]   Beis JS, Lowe DG. Shape indexing using approximate nearest-neighbour search in high-dimensional spaces. In: Proc. of the 1997 IEEE
                     Computer Society Conf. on Computer Vision and Pattern Recognition. San Juan: IEEE, 1997. 1000–1006. [doi: 10.1109/CVPR.1997.
                     609451]
                 [25]   Muja  M,  Lowe  DG.  Fast  approximate  nearest  neighbors  with  automatic  algorithm  configuration.  In:  Proc.  of  the  4th  Int’l  Conf.  on
                     Computer Vision Theory and Applications. 2009. 861–864.
                 [26]   Jégou H, Tavenard R, Douze M, Amsaleg L. Searching in one billion vectors: Re-rank with source coding. In: Proc. of the 2011 IEEE Int’l
                     Conf. on Acoustics, Speech and Signal Processing. Prague: IEEE, 2011. 861–864. [doi: 10.1109/ICASSP.2011.5946540]
                 [27]   Jégou  H,  Douze  M,  Schmid  C.  Product  quantization  for  nearest  neighbor  search.  IEEE  Trans.  on  Pattern  Analysis  and  Machine
                     Intelligence, 2011, 33(1): 117–128. [doi: 10.1109/TPAMI.2010.57]
                 [28]   Malkov YA, Yashunin DA. Efficient and robust approximate nearest neighbor search using hierarchical navigable small world graphs.
                     IEEE Trans. on Pattern Analysis and Machine Intelligence, 2020, 42(4): 824–836. [doi: 10.1109/TPAMI.2018.2889473]
                 [29]   Zhang  HL,  Zhao  PH,  Miao  XP,  Shao  YX,  Liu  ZR,  Yang  T,  Cui  B.  Experimental  analysis  of  large-scale  learnable  vector  storage
                     compression. Proc. of the VLDB Endowment, 2023, 17(4): 808–822. [doi: 10.14778/3636218.3636234]
   151   152   153   154   155   156   157   158   159   160   161