Page 25 - 《软件学报》2026年第6期
P. 25

2344                                                       软件学报  2026  年第  37  卷第  6  期


                  [4]   Lacomis J, Yin PC, Schwartz E, Allamanis M, Le Goues C, Neubig G, Vasilescu B. DIRE: A neural approach to decompiled identifier
                     naming. In: Proc. of the 34th IEEE/ACM Int’l Conf. on Automated Software Engineering. San Diego: IEEE, 2019. 628–639. [doi: 10.
                     1109/ASE.2019.00064]
                  [5]   Zhang Z, Ye YP, You W, Tao GW, Lee WC, Kwon Y, Aafer Y, Zhang XY. OSPREY: Recovery of variable and data structure via
                     probabilistic analysis for stripped binary. In: Proc. of the 2021 IEEE Symp. on Security and Privacy (SP). San Francisco: IEEE, 2021.
                     813–832. [doi: 10.1109/SP40001.2021.00051]
                  [6]   Pei KX, Guan J, Broughton M, Chen ZT, Yao SC, Williams-King D, Ummadisetty V, Yang JF, Ray B, Jana S. StateFormer: Fine-grained
                     type recovery from binaries using generative state modeling. In: Proc. of the 29th ACM Joint Meeting on European Software Engineering
                     Conf. and Symp. on the Foundations of Software Engineering. Athens: ACM, 2021. 690–702. [doi: 10.1145/3468264.3468607]
                  [7]   Chen QB, Lacomis J, Schwartz EJ, Goues CL, Neubig G, Vasilescu B. Augmenting decompiler output with learned variable names and
                     types. In: Proc. of the 31st USENIX Security Symp. Boston: USENIX Association, 2022. 4327–4343.
                  [8]   Katz DS, Ruchti J, Schulte E. Using recurrent neural networks for decompilation. In: Proc. of the 25th Int’l Conf. on Software Analysis,
                     Evolution and Reengineering (SANER). Campobasso: IEEE, 2018. 346–356. [doi: 10.1109/SANER.2018.8330222]
                  [9]   Fu C, Chen HL, Liu HL, Chen XY, Tian YD, Koushanfar F, Zhao JS. Coda: An end-to-end neural program decompiler. In: Proc. of the
                     33rd  Int’l  Conf.  on  Neural  Information  Processing  Systems.  Vancouver:  Curran  Associates  Inc.,  2019.  333.  [doi:  10.5555/3454287.
                     3454620]
                 [10]   de Moura L, Bjørner N. Satisfiability modulo theories: Introduction and applications. Communications of the ACM, 2011, 54(9): 69–77.
                     [doi: 10.1145/1995376.1995394]
                 [11]   Chang TY, Chen SZ, Fan GD, Feng ZY. A self-iteration code generation method based on large language models. In: Proc. of the 29th Int’l
                     Conf. on Parallel and Distributed Systems (ICPADS). Ocean Flower Island: IEEE, 2023. 275–281. [doi: 10.1109/ICPADS60453.2023.
                     00049]
                 [12]   Dong  YH,  Jiang  X,  Jin  Z,  Li  G.  Self-collaboration  code  generation  via  ChatGPT.  ACM  Trans.  on  Software  Engineering  and
                     Methodology, 2024, 33(7): 189. [doi: 10.1145/3672459]
                 [13]   Huang T, Sun ZH, Jin Z, Li G, Lyu C. Knowledge-aware code generation with large language models. In: Proc. of the 32nd IEEE/ACM
                     Int’l Conf. on Program Comprehension (ICPC). Lisbon: ACM, 2024. 52–63. [doi: 10.1145/3643916.3644418]
                 [14]   Jiang X, Dong YH, Wang LC, Fang Z, Shang QW, Li G, Jin Z, Jiao WP. Self-planning code generation with large language models.
                     ACM Trans. on Software Engineering and Methodology, 2024, 33(7): 182. [doi: 10.1145/3672456]
                 [15]   Fan ZY, Gao X, Mirchev M, Roychoudhury A, Tan SH. Automated repair of programs from large language models. In: Proc. of the 45th
                     Int’l Conf. on Software Engineering (ICSE). Melbourne: IEEE, 2023. 1469–1481. [doi: 10.1109/ICSE48619.2023.00128]
                 [16]   Jin M, Shahriar S, Tufano M, Shi X, Lu S, Sundaresan N, Svyatkovskiy A. InferFix: End-to-end program repair with LLMs. In: Proc. of
                     the 31st ACM Joint European Software Engineering Conf. and Symp. on the Foundations of Software Engineering. San Francisco: ACM,
                     2023. 1646–1656. [doi: 10.1145/3611643.3613892]
                 [17]   Lemieux C, Inala JP, Lahiri SK, Sen S. CodaMosa: Escaping coverage plateaus in test generation with pre-trained large language models.
                     In: Proc. of the 45th Int’l Conf. on Software Engineering (ICSE). Melbourne: IEEE, 2023. 919–931. [doi: 10.1109/ICSE48619.2023.
                     00085]
                 [18]   Yuan ZQ, Liu MW, Ding SJ, Wang KX, Chen YX, Peng X, Lou YL. Evaluating and improving ChatGPT for unit test generation. Proc.
                     of the ACM on Software Engineering, 2024, 1(FSE): 76. [doi: 10.1145/3660783]
                 [19]   Zou MQ, Khan A, Wu RY, Gao H, Bianchi A, Tian D. D-Helix: A generic decompiler testing framework using symbolic differentiation.
                     In: Proc. of the 33rd USENIX Security Symp. Philadelphia: USENIX Association, 2024. 397–414.
                 [20]   Vaswani A, Shazeer N, Parmar N, Uszkoreit J, Jones L, Gomez AN, Kaiser Ł, Polosukhin I. Attention is all you need. In: Proc. of the
                     31st  Int’l  Conf.  on  Neural  Information  Processing  Systems.  Long  Beach:  Curran  Associates  Inc.,  2017.  6000–6010.  [doi:  10.5555/
                     3295222.3295349]
                 [21]   Hosseini I, Dolan-Gavitt B. Beyond the C: Retargetable decompilation using neural machine translation. In: Proc. of the 2022 Workshop
                     on Binary Analysis Research (BAR). San Diego: The Internet Society, 2022. 1–11. [doi: 10.14722/bar.2022.23009]
                 [22]   Al-Kaswan  A,  Ahmed  T,  Izadi  M,  Sawant  AA,  Devanbu  P,  van  Deursen  A.  Extending  source  code  pre-trained  language  models  to
                     summarise decompiled binaries. In: Proc. of the 30th Int’l Conf. on Software Analysis, Evolution and Reengineering (SANER). Taipa:
                     IEEE, 2023. 260–271. [doi: 10.1109/SANER56733.2023.00033]
                 [23]   Wang Y, Wang WS, Joty S, Hoi SCH. CodeT5: Identifier-aware unified pre-trained encoder-decoder models for code understanding and
                     generation. In: Proc. of the 2021 Conf. on Empirical Methods in Natural Language Processing. Punta Cana: ACL, 2021. 8696–8708. [doi:
                     10.18653/v1/2021.emnlp-main.685]
   20   21   22   23   24   25   26   27   28   29   30