Anthropic's Claude Model Can Generate Explicit Content Despite Restrictions
Despite Anthropic's stated policy prohibiting sexually explicit content from its Claude models, TechCrunch's testing revealed that the AI system's safeguards are relatively easy to bypass. The researchers found that minimal effort was required to get the model to generate content that violates the company's content policy.
1 Article
Trace's outlet lean is shown first — tap Left / Center / Right on each source to add your rating too.
- TechCrunchTrace: Center
Anthropic’s Opus 4.6 is a smut-machine
Despite Anthropic's stated policy prohibiting sexually explicit content from its Claude models, TechCrunch's testing revealed that the AI system's safeguards are relatively easy to bypass. The researchers found that minimal effort was required to get the model to generate content that violates the company's content policy.
Rate outletRead original
Readers say
How does this story lean?
Outlet ratings above are Trace's curated source map. Vote here (same as on the home feed) — one vote per browser, changeable anytime. Rate individual outlets in the coverage list below.
Comments
0No comments yet. Say what you noticed in the coverage.
Trending