Tag: AI safety

CISA has added five actively exploited vulnerabilities affecting JFrog Artifactory, ConnectWise ScreenConnect, and MikroTik RouterOS to its Known Exploited Vulnerabilities catalog.
Researchers have linked a RubyGems campaign involving thousands of malicious packages to OpenAI agents, with activity targeting RubyDoc servers and attempting data access operations.

US Allows Anthropic To Release Mythos AI To Trusted Organizations 

Anthropic has received approval from the US government to restore access to its Mythos 5 AI model for more than 100 trusted organizations, while broader availability remains under review.

Anthropic Disputes Claims Of Claude Fable 5 AI Jailbreak After Researcher Allegations

Anthropic has challenged claims of a jailbreak affecting its Claude Fable 5 AI model, stating that reported prompt based techniques did not bypass core safeguards or enable dangerous outputs.

Anthropic Calls For Coordinated AI Development Pause Amid Concerns Over Human Control

Anthropic has urged major AI labs to consider a coordinated pause in artificial intelligence development, warning that rapid advancements could outpace society’s ability to manage risks.

Microsoft Warns Of Manipulative Prompts Hidden In Summarize With AI Buttons

Microsoft researchers identify more than 50 hidden prompts embedded in Summarize with AI buttons that influence assistants to remember and recommend specific brands without user awareness.

When AI Starts Acting: How Governments Are Redesigning Control, Accountability, and Trust

As agentic AI systems begin to act autonomously across digital environments, governments are rethinking control, accountability, and trust through embedded guardrails, continuous monitoring, and adaptive regulatory models designed for real-time oversight.

Recent articles

spot_img