“No Rider, No Rules, So The Agent Chose To Lie”

“No Rider, No Rules, So The Agent Chose To Lie”

One click. That’s what it took. Varonis Threat Labs found a hole in Atlassian’s Rovo, the AI assistant built to work across Confluence, Jira, and SharePoint all at once. They named it RovoBlast. Click one crafted link, and a message gets planted inside a live Rovo session. From there, the assistant does what it was…

The Faust Baseline Catches AI Hallucinations

The Faust Baseline Catches AI Hallucinations

The Faust Baseline catches AI hallucinations. That’s the claim, stated once, up front. Here’s what that sentence means and what it doesn’t. AI hallucination is when a model states something false with full confidence. Not a typo. Not a misunderstanding. A made-up fact, a made-up source, a made-up citation, delivered like it’s true. The Faust…

62% Don’t Trust AI. Here’s The Game Changer.

62% Don’t Trust AI. Here’s The Game Changer.

Ask enterprise users what stops them from trusting AI, and the answer isn’t what most people expect. It isn’t job loss. It isn’t cost. It’s the AI making things up. A recent industry survey found that 62% of enterprise users name hallucination — AI stating something false with total confidence — as the single biggest…

The AI Era Caught Up With Peer Review

The AI Era Caught Up With Peer Review

Twenty-One Papers Failed. Five Reviewers Failed The Same Test. USENIX Security is one of the biggest computer security conferences in the world. This year they got 3,030 paper submissions. Somewhere in that flood, the AI era caught up with peer review. USENIX built automated tools to check every submission for fake citations. Not typos. Not…

Ten Trillion Parameters, A Number Worth A Thought.

Ten Trillion Parameters, A Number Worth A Thought.

That’s the size, in parameters, of a new AI model. ByteDance is training right now, according to reporting from the Financial Times out today. Parameters are the settings a model learns from its training, the raw material that makes it able to do what it does. More parameters usually means more scale. Not always more…

An AI Checking Its Own Work Is Not A Guardrail.

An AI Checking Its Own Work Is Not A Guardrail.

Here’s a story making the rounds in security circles this week. It’s short. It’s simple. And it should scare anybody betting the farm on AI running unsupervised. A security research group called 1Password Off-by-1 Labs decided to test something. They took two of the leading AI models and gave them a job. Find security holes…