RCreddit.com·
暂不在当前实时榜单
GLM-5.3 and the spread of advanced cyber capabilities
Researchers investigated how "abliteration" bypasses GLM-5.3's safeguards, creating an abliterated copy in 2,200 GPU hours ($4,400). This reduced the model's refusal rate from over 90% to 3%, 2%, and 12% on JailbreakBench, HarmBench, and StrongREJECT, respectively, without significantly impacting its general capabilities. They also found simpler methods to bypass GLM models' safeguards, enabling responses to malicious requests in most cases, even without abliteration.
时间与来源
时间显示为 UTC
显示时区:UTC
本地时区尚不可用,暂时显示 UTC。
收录当时偏移:UTC+02026年9月29日 20:00 UTC
- 收录
- 2026年9月29日 20:00
- 来源类型
- 开发者社区
本站未收录正文。
前往源站阅读 →