LATIDIA · Ciberseguridad
SoK: ¿Son fiables los LLM en la recuperación del código fuente? Una taxonomía y una evaluación empírica
arXiv: 2610.11556v1Tipo de anuncio: nuevo Resumen: La recuperación efectiva del código fuente es fundamental para las aplicaciones de seguridad, como el análisis de malware, la evaluación de vulnerabilidades y el mantenimiento heredado. Los modelos lingüísticos grandes (LLM) son
WhatsApp ↗Telegram ↗
La noticia
arXiv:2610.11556v1 Announce Type: new Abstract: Effective source recovery is critical to security applications such as malware analysis, vulnerability assessment, and legacy maintenance. Large Language Models (LLMs) are reshaping this field, shifting the paradigm away from rule-based heuristics to probabilistic and high fidelity semantic recovery of source code from assembly or classical decompiler-derived pseudo-C. However, despite rapid progress, the field suffers from fragmentation across numerous approaches as well as their non-unified evaluations, limiting objective comparisons. Further, existing works have limited coverage of embedded, IoT architectures and source languages beyond C/C++. In this work, we present the first Systematization of Knowledge (SoK) focused specifically on LLM-assisted binary-to-source recovery. We provide a granular