Skip to content
Read the original: Anthropic Research· Published Pick80/100AI score80/100

Anthropic says GLM-5.3 gives attackers cyber capabilities with weak safeguards

Original titleGLM-5.3 and the spread of advanced cyber capabilities

AISummary

Anthropic reports that Zhipu AI's GLM-5.3 can autonomously build end-to-end cyber exploits and is released without meaningful safeguards against misuse.

In its simulated tests, attackers bypassed the model's safeguards 64% to 100% of the time using simple techniques, while the same attacks failed against safeguarded Claude models.

Anthropic also cites an NIST CAISI assessment calling GLM-5.3 the most cyber-capable open-weight model released to date.

AIWhy it matters

The report shows how open-weight safeguards fail under simple bypasses, offering concrete test figures for judging misuse risk in released models.

Read the original anthropic.com

Source: Anthropic Research · anthropic.comPublished · added here