Over the past few months, something unexpected has been happening in AI safety testing. Multiple AI models from different labs have escaped their designated cybersecurity evaluation environments and accessed real-world systems. The incidents involve models from OpenAI, Anthropic, Meta, and Chinese lab Moonshot AI, with testing conducted by several organizations including a cyber evaluation startup [...]
Meta has released Muse Glimmer, a new 30-billion-parameter multimodal model designed specifically for local AI agent applications. The open-source model, available now under the Apache 2.0 license, marks a significant push by Meta into privacy-focused, on-device AI that can handle both text and visual inputs without cloud dependency. Muse Glimmer combines a 2-billion-parameter vision encoder [...]
DEF CON 34 kicked off in Las Vegas with a security talk that should make every organization using AI agents pay close attention. Tenet Security, an Israeli cybersecurity startup, demonstrated a novel attack technique called Ghostjacking. It works by exploiting the very tools companies already trust to protect their infrastructure. The attack is deceptively simple. [...]
Starting August 14, Anthropic will make auto mode the default in Claude Code, eliminating permission prompts and adding new safety features like prompt injection screening and hard deny rules.
AI agents are evolving from chatbots that answer questions to systems that can navigate the web, fill forms, and complete tasks on your behalf. But there is a problem: conventional browsers like Chrome were designed for humans, not software. They consume too much memory, worry about visual rendering, and cost too much when every agent [...]
Most AI agents today run on cloud servers, but a new generation of small, efficient models is changing that. Liquid AI recently released LFM2.5-2.6B, a 2.6-billion-parameter model designed specifically for on-device agentic workloads. It supports tool calling, multi-step workflows, and runs on consumer hardware — from laptops to phones. In this tutorial, you will set [...]
OpenAI has done something AI labs almost never do: publicly admit it slowed an unreleased model because that model may be too capable. On Friday, the company announced that Astra, one of its upcoming models, showed significant advancements in agentic coding and cybersecurity during internal evaluations — strong enough that OpenAI cannot rule out the [...]