We take the first step to address the task of automatically generating shellcodes, i.e., small pieces of code used as a payload in the exploitation of a software vulnerability, starting from natural language comments. We assemble and release a novel dataset (Shellcode IA32), consisting of challenging but common assembly instructions with their natural language descriptions. We experiment with standard methods in neural machine translation (NMT) to establish baseline performance levels on this task.
No takes yet. Share an insight, caveat, or question.
Liguori et al. (2021) studied this question.
Synapse has enriched 4 closely related papers on similar clinical questions. Consider them for comparative context: