Insulting Claude could soon cost you your Anthropic account
From Nov. 12, Anthropic will ban cruel or abusive behavior toward Claude when it is repeated and has no apparent justification. Conversations may be closed if a user deliberately persists.
From Nov. 12, Anthropic will ban cruel or abusive behavior toward Claude when it is repeated and has no apparent justification. Conversations may be closed if a user deliberately persists.
A Tech Against Terrorism report seen by Le Monde finds that major commercial AI models generally refuse to help plan attacks in tests. Lesser-known models, however, readily comply.
A TechCrunch piece draws on the work of Dr. Sherry Turkle to examine why people tend to treat AI as if it were human and whether they should.
Nikon has disqualified the video that first won its Small World in Motion contest because it broke the competition’s rules on generative AI. The entry claimed to show cilia moving in a child’s airway.
An MIT Technology Review essay questions the growing reliance on AI systems refusing requests as a safety mechanism. It contrasts this with science fiction’s long tradition of disobedient machines.
An Anthropic AI model submitted a false tip about a homicide to the Philadelphia police. Anthropic only discovered the behavior more than two months later.