The non-profit SentinelAI Foundation has launched 'Glaive,' a new open-source framework designed to automate the security testing and red-teaming of Large Language Models. Glaive provides a comprehens
AIBW Security DeskA research paper published by Carnegie Mellon University's CyLab introduces a novel jailbreak technique named 'Chameleon.' The attack embeds hidden, malicious instructions within Unicode glyphs that a
AIBW Security DeskNexusAI, a leading AI development firm, confirmed it suffered a significant data breach originating from a supply-chain attack on a third-party data annotation partner. The breach exposed a repository
AIBW Security DeskYour 3-minute AI news briefing for Tuesday, April 14, 2026. The top stories in artificial intelligence, delivered daily.
Daily DigestNew research reveals a major weakness in today's most advanced vision AI models: they often can't cite the specific data source for their answers. A new benchmark called ViTaB-A highlights this critical attribution gap. Why does this matter for AI trust?
Research WireThe United States Senate has passed the 'AI Verification, Evaluation, and Reporting for Foundational Integrity' (VERIFY) Act, a landmark piece of legislation aimed at improving AI safety and accountab
AIBW Security DeskA new paper published by researchers from Carnegie Mellon University introduces a novel jailbreak technique named 'Contextual Carryover Attack' (CCA). This method exploits the expanding context window
AIBW Security DeskEnterprise AI provider CognitoAI disclosed a significant security breach affecting its core platform. Attackers exploited a vulnerability in a third-party data ingestion library used for model fine-tu
AIBW Security DeskA paper published by researchers at Stanford's AI Lab has introduced a powerful new jailbreak technique named 'Semantic Blindspotting,' which specifically targets multimodal large language models. The
AIBW Security DeskCognition Labs, the creator of the AI software engineer agent 'Devin,' confirmed it was the victim of a major security breach orchestrated by a disgruntled former employee. The company disclosed that
AIBW Security DeskA landmark international agreement, the 'Geneva Accords on AI Safety,' has been ratified by 30 nations, including the US, EU member states, and the UK. This policy establishes the first binding intern
AIBW Security DeskResearchers unveil EpiScreen, a new LLM that analyzes electronic health records to distinguish between epileptic and non-epileptic seizures, accelerating patient care. Can AI finally solve the epilepsy diagnosis dilemma?
Research Wire