We gratefully acknowledge support from
the Simons Foundation and member institutions.
Full-text links:

Download:

Current browse context:

cs.SE

Change to browse by:

References & Citations

DBLP - CS Bibliography

Bookmark

(what is this?)
CiteULike logo BibSonomy logo Mendeley logo del.icio.us logo Digg logo Reddit logo ScienceWISE logo

Computer Science > Software Engineering

Title: Shellcode_IA32: A Dataset for Automatic Shellcode Generation

Abstract: We take the first step to address the task of automatically generating shellcodes, i.e., small pieces of code used as a payload in the exploitation of a software vulnerability, starting from natural language comments. We assemble and release a novel dataset (Shellcode_IA32), consisting of challenging but common assembly instructions with their natural language descriptions. We experiment with standard methods in neural machine translation (NMT) to establish baseline performance levels on this task.
Comments: Paper accepted to NLP4Prog Workshop 2021 co-located with ACL-IJCNLP 2021. Extended journal version of this work has been published in the Automated Software Engineering journal, Volume 29, Article no. 30, March 2022, DOI: 10.1007/s10515-022-00331-3
Subjects: Software Engineering (cs.SE); Computation and Language (cs.CL)
DOI: 10.18653/v1/2021.nlp4prog-1.7
Cite as: arXiv:2104.13100 [cs.SE]
  (or arXiv:2104.13100v4 [cs.SE] for this version)

Submission history

From: Pietro Liguori [view email]
[v1] Tue, 27 Apr 2021 10:50:47 GMT (5524kb,D)
[v2] Sat, 5 Jun 2021 07:41:21 GMT (5998kb,D)
[v3] Tue, 8 Jun 2021 09:23:08 GMT (5612kb,D)
[v4] Fri, 18 Mar 2022 10:28:57 GMT (5612kb,D)

Link back to: arXiv, form interface, contact.