From ASR to ASP: Evaluating Prompt Attack Vulnerabilities Against Open-Source LLMs
Researchers have identified new security vulnerabilities in open-source large language models that allow attackers to force the generation of sensitive or harmful content. As these models see increased adoption in sectors like finance and healthcare, these findings highlight the necessity of improving robustness against prompt-based exploitation.
Covered by 1 source
- AarXiv CS.AI↗Jiawen Wang, Pritha Gupta, Eyke H\"ullermeier, Xiaoxue Gao, Nancy F. Chen3d ago