电脑桌面
添加运营动脉到电脑桌面
安装后可以在桌面快捷访问

大型语言模型安全:全面综述(英文)会员免费

大型语言模型安全:全面综述(英文)_第1页
1/159
大型语言模型安全:全面综述(英文)_第2页
2/159
大型语言模型安全:全面综述(英文)_第3页
3/159
大型语言模型安全:全面综述(英文)_第4页
4/159
大型语言模型安全:全面综述(英文)_第5页
5/159
大型语言模型安全:全面综述(英文)_第6页
6/159
大型语言模型安全:全面综述(英文)_第7页
7/159
大型语言模型安全:全面综述(英文)_第8页
8/159
大型语言模型安全:全面综述(英文)_第9页
9/159
大型语言模型安全:全面综述(英文)_第10页
10/159
Large Language Model Safety: A Holistic SurveyDan Shi'*.TianhaoShen'*,Yufei Huang', Zhigen Li-2Yomgyi Leng?, Renran Iin', Chuang LAnm', Xinwed Wir,Zishan Guo',Linhao Yu'. Ling Shi', Bojian Jiang!8, DeyiXiong'TJUNLP Lab. Tianin University2PingAnTechnology,3Du Xiaoman Finance7Z0Z0QGEZIVSO]14989L1ZI7ZAIXDAbstractThe rapid development and deploy ment of large language mnodels (LLMs) haveintroduced a new frontier in artificial inteligence, marked by unprecedented capabilities in natural language understanding and generation. Howevwer, the increasing integration of these models into critical applications raisessubstantiasafety concerns, necessitating a thorough examination of their potential risksand associated mitigation strategies.This survev provides a comprehensivon theseinterpretabilitv in enhancingLLM safety, the technoldcompanies and institutes for LLM safety, and AIgovernance alimed at LLMsafety with discussions on international cooperation. policv probosals andprospective regulatorv directioneOurfindings underscore thea proactive. multifaceted approach toLAM safetv. emphasizing the intration of technical solutions. ethical considas a foundationalreso, industry practitioners, andsand opportunities associatedwith the safe integration of LLMs intoetv Ultimatelv it seeksto contribnute to the safe and benelicial develoomef lMs aligning with the overarchinggoal of harnessing Al for societal advancement and well-being. A curated lisof related papers has been publiclv avaiable at a Git Hub repositorv.*EqualcontributiontCoresponding author.Correspondence to: {shidan, thshen, dyxIhttps://github.com/tjunlp-lab/Awesome-LLM-Safety-Papers

Contents下991 Introduction1.11.2Paper and Source Selection.1.3Related Work2Taxonomy Basic Areas of LLM Safety...2.12.2RelatedAreastoLLMSafety.3Value Misalignment3.1SocialBias.... Defnition and Safety Impact....3.1.13.1.2Social Bias in the LLM Lifecvcle3.1.3Methods for Mitigating Social Bias3.1.4Evaluation..3.1.5Future Directions,3.2Privacy3.2.1Preliminaries3.2.2Sources and Channels of Privacy Leakage...323Privacy Protection Methods....3.3Toxicity3.3.1Definition and Saiety Impact.33.2Methods for Mitigating lovicity.3.3.3Evaluation .I3.4Ethics and Morality.3.4.1Definition3.4.2Safetv Issues Related to Ethics and Morality3.4.3Methodsfor Mitioatino L.L.MAmorality344Evalation4Robustness toAttack

4.1Jailbreaking....411Black-boxAttacks4.1.2White-boxAttacks4.2RedTeaming4.2.1ManualRedTeaming.4.2.2AutomatedRedTeaming4.2.3Evaluation..4.3Defense4.3.1ExternalSafeguard4.3.2Internal Protection;Misuse15.1Weaponization..5.1.1Risks of Misuse in Weapons Acquisition5.1.2Mitication Methods for Weaponized Misus5.1.3Evaluation52MisinformationCampaigns...5.2.2Social Media Manipulations;5.2.3Risks to Public Health Information5.2.4MitigationMethodsfor theSoread of Misinformation5.3Deepfakes....Malicious Applications of Deplakes.5.3.1532Methods for Mitigating Deeplakes-..5.4Future Directions5.4.1WeaponizationMisinformation Campaigns .....5.4.25.4.3Deepfakes5.4.4Comorehensive Evaluation5款6AutonomousAIRisks6.1

招路高药...

1、当您付费下载文档后,您只拥有了使用权限,并不意味着购买了版权,文档只能用于自身使用,不得用于其他商业用途(如 [转卖]进行直接盈利或[编辑后售卖]进行间接盈利)。
2、本站所有内容均由合作方或网友上传,本站不对文档的完整性、权威性及其观点立场正确性做任何保证或承诺!文档内容仅供研究参考,付费前请自行鉴别。
3、如文档内容存在违规,或者侵犯商业秘密、侵犯著作权等,请点击“违规举报”。

查找下载文件的指引

一、电脑端

- Windows 系统:按下键盘快捷键 `Ctrl + J`,即可打开下载列表。  

- Mac 系统:按下键盘快捷键 `⌘ + J`,即可打开下载列表。  

二、手机端

1. 打开手机浏览器,点击浏览器右下角的 “≡”(或“更多”)图标。  

2. 在弹出的菜单中找到并点击 “下载内容”(或类似选项),即可查看已下载的文件。  

提示:不同浏览器界面略有差异,若未找到“下载”入口,可尝试在浏览器设置中搜索“下载”关键词。


大型语言模型安全:全面综述(英文)

确认删除?
会员
教程
收藏
足迹
联系
  • 站长微信
回到顶部