Researchers from Carnegie Mellon University have published a groundbreaking paper detailing a new jailbreak technique called Universal Adversarial Triggers (UATs). This method involves appending a spe
AIBW Security DeskIn a landmark policy shift, the U.S. Department of Commerce, in collaboration with CISA, has issued a new binding directive for all AI models deployed within U.S. critical infrastructure sectors. Effe
AIBW Security DeskResearchers from the Carnegie Mellon AI Safety Institute have published a paper detailing a novel jailbreak technique called 'Semantic Obfuscation.' This method bypasses the safety filters of leading
AIBW Security DeskThe European Union's AI Office has officially released its much-anticipated 'Red Teaming Certification Framework', a key implementing act of the EU AI Act. This new regulation mandates that all develo
AIBW Security DeskA new paper from Carnegie Mellon University's CyLab has detailed a novel jailbreak technique named 'Contextual Weaving'. This method demonstrates a sophisticated way to circumvent the safety alignment
AIBW Security DeskThe AI Security Consortium (AISC) has released Aegis, a new open-source framework designed to protect AI applications from prompt injection, jailbreaking, and data exfiltration attacks. Aegis acts as
AIBW Security DeskLeading artificial intelligence firm SynthAI confirmed it was the victim of a sophisticated cyberattack that resulted in a significant data breach. Attackers, believed to be a state-sponsored group, g
AIBW Security DeskIn a landmark move for AI governance, the United States and the European Union have jointly announced the 'AI Trust & Transparency Framework' (AITTF). This new regulation mandates that companies devel
AIBW Security DeskA new research paper from the Stanford AI Lab details a sophisticated jailbreak technique named 'Contextual Weaving.' Unlike traditional prompt injection attacks that rely on a single, carefully craft
AIBW Security DeskOpenAI is taking a proactive step in AI safety by monitoring the internal reasoning of its coding agents, not just their final output. Learn about this new method.
AIBW Security DeskThe DEF CON AI Village project has released Red-Prism, a new open-source tool designed to help developers and security teams audit the security of their large language model deployments. Red-Prism aut
AIBW Security DeskCognition AI, a leading developer of large language models, announced a significant security breach. Attackers gained unauthorized access to their internal data storage, exfiltrating several terabytes
AIBW Security Desk