Page 374 - 《软件学报》2026年第3期
P. 374
张犬俊 等: 检索增强生成在软件工程中的应用综述 1337
IEEE/ACM Int’l Conf. on Software Engineering. Lisbon: IEEE, 2024. 1–12. [doi: 10.1145/3597503.3623326]
[72] Shao YC, Huang YH, Shen JW, Ma L, Su T, Wan CC. Vortex under ripplet: An empirical study of RAG-enabled applications.
arXiv:2407.05138v1, 2024.
[73] Guo YC, Li ZX, Jin XL, Liu YT, Zeng YT, Liu WX, Li X, Yang P, Bai L, Guo JF, Cheng XQ. Retrieval-augmented code generation for
universal information extraction. In: Proc. of the 13th National CCF Conf. on Natural Language Processing and Chinese Computing.
Hangzhou: Springer, 2025. 30–42. [doi: 10.1007/978-981-97-9434-8_3]
[74] Yoo J, Han H, Lee Y, Kim J, Hwang SW. PERC: Plan-as-query example retrieval for underrepresented code generation. In: Proc. of the
31st Int’l Conf. on Computational Linguistics. Abu Dhabi: ACL, 2025. 7982–7997.
[75] Liu ZH, Zeng RN, Wang DX, Peng GY, Wang JY, Liu Q, Liu PY, Wang WH. Agents4PLC: Automating closed-loop PLC code
generation and verification in industrial control systems using LLM-based agents. arXiv:2410.14209, 2024.
[76] Zan DG, Chen B, Lin ZQ, Guan B, Wang YJ, Lou JG. When language model meets private library. In: Findings of the Association for
Computational Linguistics: EMNLP 2022. Abu Dhabi: ACL, 2022. 277–288. [doi: 10.18653/v1/2022.findings-emnlp.21]
[77] Zhang KC, Zhang HZ, Li G, Li J, Li Z, Jin Z. ToolCoder: Teach code generation models to use API search tools. arXiv:2305.04032,
2023.
3
[78] Liao DS, Pan SD, Sun XY, Ren XW, Huang Q, Xing ZC, Jin H, Li QY. A -CodGen: A repository-level code generation framework for
code reuse with local-aware, global-aware, and third-party-library-aware. IEEE Trans. on Software Engineering, 2024, 50(12):
3369–3384. [doi: 10.1109/TSE.2024.3486195]
[79] Zhang KC, Li J, Li G, Shi XJ, Jin Z. CodeAgent: Enhancing code generation with tool-integrated agent systems for real-world repo-
level coding challenges. In: Proc. of the 62nd Annual Meeting of the Association for Computational Linguistics. Bangkok: ACL, 2024.
13643–13658. [doi: 10.18653/v1/2024.acl-long.737]
[80] Tan HZ, Luo Q, Jiang L, Zhan ZZ, Li J, Zhang HT, Zhang YQ. Prompt-based code completion via multi-retrieval augmented
generation. ACM Trans. on Software Engineering and Methodology, 2025. [doi: 10.1145/3725812]
[81] Liu W, Yu AL, Zan DG, Shen B, Zhang W, Zhao HY, Jin Z, Wang QX. GraphCoder: Enhancing repository-level code completion via
coarse-to-fine retrieval based on code context graph. In: Proc. of the 39th IEEE/ACM Int’l Conf. on Automated Software Engineering.
Sacramento: ACM, 2024. 570–581. [doi: 10.1145/3691620.3695054]
[82] Zhu TW, Li Z, Pan MX, Shi CX, Zhang T, Pei Y, Li XD. Deep is better? An empirical comparison of information retrieval and deep
learning approaches to code summarization. ACM Trans. on Software Engineering and Methodology, 2024, 33(3): 67. [doi: 10.1145/
3631975]
[83] Ahmed T, Pai KS, Devanbu P, Barr E. Automatic semantic augmentation of language model prompts (for code summarization). In:
Proc. of the 46th IEEE/ACM Int’l Conf. on Software Engineering. Lisbon: ACM, 2024. 220. [doi: 10.1145/3597503.3639183]
[84] Shi ES, Wang YL, Tao W, Du L, Zhang HY, Han S, Zhang DM, Sun HB. RACE: Retrieval-augmented commit message generation. In:
Proc. of the 2022 Conf. on Empirical Methods in Natural Language Processing. Abu Dhabi: ACL, 2022. 5520–5530. [doi: 10.18653/v1/
2022.emnlp-main.372]
[85] Ji XY, Parameswaran A, Hulsebos M. TARGET: Benchmarking table retrieval for generative tasks. In: Proc. of the 3rd Table
Representation Learning Workshop. Vancouver, 2024. 1–8.
[86] Li HY, Zhang J, Li CP, Chen H. RESDSQL: Decoupling schema linking and skeleton parsing for text-to-SQL. In: Proc. of the 37th
AAAI Conf. on Artificial Intelligence. Washington: AAAI Press, 2023. 13067–13075. [doi: 10.1609/aaai.v37i11.26535]
[87] Zhang K, Lin XX, Wang YZ, Zhang X, Sun F, Cen JH, Tan HX, Jiang XH, Shen HW. ReFSQL: A retrieval-augmentation framework
for text-to-SQL generation. In: Findings of the Association for Computational Linguistics: EMNLP 2023. Singapore: ACL, 2023.
664–673. [doi: 10.18653/v1/2023.findings-emnlp.48]
[88] Poesia G, Polozov O, Le V, Tiwari A, Soares G, Meek C, Gulwani S. Synchromesh: Reliable code generation from pre-trained language
models. In: Proc. of the 10th Int’l Conf. on Learning Representations. OpenReview.net, 2022.
[89] Shi P, Zhang R, Bai H, Lin J. XRICL: Cross-lingual retrieval-augmented in-context learning for cross-lingual text-to-SQL semantic
parsing. In: Findings of the Association for Computational Linguistics: EMNLP 2022. Abu Dhabi: ACL, 2022. 5248–5259. [doi: 10.
18653/v1/2022.findings-emnlp.384]
[90] Chang SC, Fosler-Lussier E. Selective demonstrations for cross-domain text-to-SQL. In: Findings of the Association for Computational
Linguistics: EMNLP 2023. Singapore: ACL, 2023. 14174–14189. [doi: 10.18653/v1/2023.findings-emnlp.944]
[91] Zhang XL, Wang DZR, Dou LX, Zhu QF, Che WX. MURRE: Multi-hop table retrieval with removal for open-domain text-to-SQL. In:
Proc. of the 31st Int’l Conf. on Computational Linguistics. Abu Dhabi: ACL, 2025. 5789–5806.
[92] Deng X, Ye W, Xie R, Zhang SK. Survey of source code bug detection based on deep learning. Ruan Jian Xue Bao/Journal of Software,

