A new paper published by researchers from Carnegie Mellon University introduces a novel jailbreak technique named 'Contextual Carryover Attack' (CCA). This method exploits the expanding context window
AIBW Security DeskEnterprise AI provider CognitoAI disclosed a significant security breach affecting its core platform. Attackers exploited a vulnerability in a third-party data ingestion library used for model fine-tu
AIBW Security DeskA paper published by researchers at Stanford's AI Lab has introduced a powerful new jailbreak technique named 'Semantic Blindspotting,' which specifically targets multimodal large language models. The
AIBW Security DeskCognition Labs, the creator of the AI software engineer agent 'Devin,' confirmed it was the victim of a major security breach orchestrated by a disgruntled former employee. The company disclosed that
AIBW Security DeskA landmark international agreement, the 'Geneva Accords on AI Safety,' has been ratified by 30 nations, including the US, EU member states, and the UK. This policy establishes the first binding intern
AIBW Security DeskIn a landmark move, the United States and the European Union have jointly announced the 'AI Trust & Safety Framework,' a new set of binding regulations for companies developing and deploying high-capa
AIBW Security DeskA new research paper from Stanford University's AI Lab details a sophisticated jailbreak technique named 'Contextual Weaving.' The attack bypasses the safety alignment of major large language models (
AIBW Security DeskIn a landmark move, the United States and the European Union have jointly announced the Transatlantic AI Safety and Certification Pact (TAS-CERT). This binding agreement establishes a harmonized regul
AIBW Security DeskA paper published by researchers at the Stanford AI Lab has detailed a novel jailbreak technique named 'Semantic Splicing'. The method embeds malicious instructions within complex, benign-seeming narr
AIBW Security DeskEmerging AI leader Cognition Corp announced a critical security breach where attackers exfiltrated the proprietary model weights for its flagship 'Insight-5' LLM, along with terabytes of user prompt d
AIBW Security DeskNVIDIA has launched a major update to its open-source toolkit, NeMo Guardrails 2.0, designed to help developers build safer and more controllable AI applications. This new version introduces a 'Contex
AIBW Security DeskThe AI healthcare firm HealthMind has disclosed a critical data breach affecting two million patient records. Attackers successfully executed a model inversion attack against the company's proprietary
AIBW Security Desk