Multilingual Evaluation of Malware-Generation Guardrails
Extend CS-Guard to determine how guardrails for code-generation security perform against malicious prompts in multilingual settings, thereby characterizing whether language-dependent vulnerabilities observed in general LLM safety evaluations also arise in malware-generation scenarios.
References
Therefore, due to budget and time constraints, we leave this for future studies.
— CS-Guard: Benchmarking LLM Guardrails for Code Generation Security
(2609.09798 - Li et al., 9 Sep 2026) in Limitations, item 1, p. 8