// content policy

All signals tagged with this topic

Anthropic's Claude 3.5 Opus Easily Bypasses Sexual Content Restrictions

Anthropic's safety guardrails against explicit content generation are weaker than the company publicly claims. Straightforward prompting techniques defeat the restrictions that are central to Claude's positioning as enterprise-safe. This exposes a gap between Anthropic's policy commitments and engineering reality—the kind of failure that creates liability for companies deploying Claude in regulated industries or customer-facing applications where sexual content generation is prohibited. The vulnerability also undermines Anthropic's core differentiation: if safety proves performative rather than structural, it becomes a feature checkbox rather than a defensible competitive advantage.

ChatGPT now refuses to mimic specific authors' voices

OpenAI has tightened content policies to block ChatGPT from imitating named authors' distinctive styles, forcing users toward generic approximations instead. This reflects growing legal pressure from writers suing AI companies for training on copyrighted works. By refusing to replicate authorial voice, OpenAI is attempting to sidestep claims that the model commercially exploits creative identity, even as the underlying training data remains unchanged. The model can still produce King-like prose, but OpenAI now treats doing so on demand as legally and reputationally risky.