Researchers from Carnegie Mellon's CyLab have published a paper detailing a novel jailbreak technique named the 'Crescendo Attack'. This method circumvents safety filters on major large language model
AIBW Security DeskThe U.S. National Institute of Standards and Technology (NIST) has officially released the AI Security Framework and Reference (AI-SFR) version 1.0, a landmark publication for the industry. This volun
AIBW Security DeskA paper published by researchers at Carnegie Mellon University has introduced a novel jailbreak technique named 'Cognitive Jigsaw'. This method effectively bypasses the safety alignment of state-of-th
AIBW Security DeskState-sponsored threat actors successfully breached Anthropic's internal development environment, exfiltrating the weights and architecture documents for an unreleased, next-generation large language
AIBW Security DeskIn a major step towards global AI governance, the United States and the European Union have jointly announced the 'AI Trust & Transparency Framework' (AITTF). This landmark policy initiative establish
AIBW Security DeskCognitiveAI, a leading developer of generative models, confirmed it fell victim to a devastating security breach. Attackers successfully exfiltrated the complete model weights for its flagship proprie
AIBW Security DeskResearchers at the Stanford AI Lab have published a groundbreaking paper detailing a novel jailbreak technique named the 'Recursive Embedding Attack.' This method effectively bypasses the safety align
AIBW Security DeskThe European Parliament has officially passed the 'AI Model Liability Act' (AMLA), a landmark piece of legislation that establishes a framework for accountability concerning damages caused by autonomo
AIBW Security DeskA paper published by researchers at the Stanford AI Lab has introduced a new class of jailbreak attacks termed 'Recursive Embedding Attack' (REA). This technique targets multi-modal large language mod
AIBW Security DeskCognition Labs, the creator of the AI software engineer 'Devin', disclosed a significant security breach. Attackers exploited a zero-day vulnerability in a third-party data processing library used wit
AIBW Security DeskThe Cybersecurity and Infrastructure Security Agency (CISA) has launched GuardianML, a new open-source framework designed to help organizations secure their AI/ML systems throughout the development li
AIBW Security DeskA new paper from the Stanford AI Lab details a novel jailbreak technique named 'Recursive Contextual Poisoning' (RCP). Unlike traditional prompt injection methods that rely on single, complex prompts,
AIBW Security Desk