Anthropic published an updated “constitution” for its Claude model, positioning it as a holistic framework for how the assistant should behave in ethically ambiguous situations and as a reference point for building “safer” AI systems. Reporting highlighted that the document uses language implying Claude could one day be “conscious” or consciousness-like, while emphasizing that the constitution functions more as guiding principles than rigid constraints, reflecting ongoing uncertainty about how AI agents should act in the world and what values they should prioritize.
Separate reporting raised broader governance concerns about AI deployment in high-stakes contexts, including Perplexity’s initiative to provide law-enforcement agencies access to its Enterprise Pro capabilities for tasks such as analyzing crime scene photos, body-camera transcripts, and investigators’ notes. Experts warned that common LLM failure modes—hallucinations, inaccuracies, and bias—can have outsized consequences in policing, underscoring the gap between vendor programs promoting adoption and the still-maturing protocols, oversight, and accountability mechanisms needed for safe use in sensitive public-sector workflows.

Track how attackers are adapting to this technology.
2 events from the most recent confirmed update back to the earliest known activity.
A class action lawsuit was reported challenging allegedly hidden AI hiring tools used against workers, with the case described as having significant regulatory implications. The available reference provides only headline-level detail and does not specify a more precise event date than publication.
Anthropic released a new "constitution" for its Claude AI chatbot, presenting it as a flexible framework to guide the model toward being safe, ethical, compliant with company policies, and helpful. The document also sets hard constraints, including bans on enabling critical infrastructure attacks, generating child sexual abuse material, and supporting human extinction or disempowerment.
Vulnerabilities, threat actors, malware, products, organizations, and breaches Mallory has linked to this story.
Follow how adversaries are adapting to this technology, and where it touches your stack today.
3 references tracked. Mallory keeps watching after this page renders.
lawfaremedia.org
Open sourcezdnet.com
Open sourcecio.com
Open sourceMap indicators from this story to your assets and identify affected systems in minutes.
Every observed campaign, victim, and pivot linked to actors named in this story.
Malware, exploits, and IOCs connected to the activity described here.
YARA, Sigma, and Snort rules deployed to your SIEM as soon as they’re published.
Get matching new stories delivered to your team as they break — not the next morning.
Ask questions about this story and take action on the answers.