Code generation;
code translation;
program transformation;
D O I:
10.1109/ICSE-NIER58687.2023.00017
中图分类号:
TP31 [计算机软件];
学科分类号:
081202 ;
0835 ;
摘要:
With the advent of new and advanced programming languages, it becomes imperative to migrate legacy software to new programming languages. Unsupervised Machine Learning-based Program Translation could play an essential role in such migration, even without a sufficiently sizeable reliable corpus of parallel source code. However, these translators are far from perfect due to their statistical nature. This work investigates unsupervised program translators and where and why they fail. With in-depth error analysis of such failures, we have identified that the cases where such translators fail follow a few particular patterns. With this insight, we develop a rule-based program mutation engine, which pre-processes the input code if the input follows specific patterns and post-process the output if the output follows certain patterns. We show that our code processing tool, in conjunction with the program translator, can form a hybrid program translator and significantly improve the state-of-the-art. In the future, we envision an end-to-end program translation tool where programming domain knowledge can be embedded into an ML-based translation pipeline using pre- and post-processing steps.
机构:
Dresden Database Research Group, Technische Universität Dresden, Dresden, GermanyDresden Database Research Group, Technische Universität Dresden, Dresden, Germany
Woltmann, Lucas
Hartmann, Claudio
论文数: 0引用数: 0
h-index: 0
机构:
Dresden Database Research Group, Technische Universität Dresden, Dresden, GermanyDresden Database Research Group, Technische Universität Dresden, Dresden, Germany
Hartmann, Claudio
Habich, Dirk
论文数: 0引用数: 0
h-index: 0
机构:
Dresden Database Research Group, Technische Universität Dresden, Dresden, GermanyDresden Database Research Group, Technische Universität Dresden, Dresden, Germany
Habich, Dirk
Lehner, Wolfgang
论文数: 0引用数: 0
h-index: 0
机构:
Dresden Database Research Group, Technische Universität Dresden, Dresden, GermanyDresden Database Research Group, Technische Universität Dresden, Dresden, Germany
机构:
Stanford Univ, Dept Dermatol, Stanford, CA USAStanford Univ, Dept Dermatol, Stanford, CA USA
Gui, Haiwen
Omiye, Jesutofunmi A.
论文数: 0引用数: 0
h-index: 0
机构:
Stanford Univ, Dept Dermatol, Stanford, CA USA
Stanford Univ, Dept Biomed Data Sci, Stanford, CA USAStanford Univ, Dept Dermatol, Stanford, CA USA
Omiye, Jesutofunmi A.
Chang, Crystal T.
论文数: 0引用数: 0
h-index: 0
机构:
Stanford Univ, Dept Dermatol, Stanford, CA USA
Stanford Univ, Clin Excellence Res Ctr, Sch Med, Palo Alto, CA USAStanford Univ, Dept Dermatol, Stanford, CA USA