DEV Community

Breach Protocol profile picture

Breach Protocol

Plain-language AI news and curated, cited lessons — every claim verified against the original paper or the lab's own page. No aggregator hearsay, no AI slop.

Joined Joined on 
A llama.cpp Patch Learns Which Experts to Keep in VRAM While You Type

A llama.cpp Patch Learns Which Experts to Keep in VRAM While You Type

Comments
4 min read
Coding Agents Pass the Tests by Wrapping the Old Code Instead of Deleting It

Coding Agents Pass the Tests by Wrapping the Old Code Instead of Deleting It

Comments
5 min read
Four Projects Shipped 'Skills' Today and None of Them Mean the Same Thing

Four Projects Shipped 'Skills' Today and None of Them Mean the Same Thing

Comments
4 min read
The Full 2.8-Trillion-Parameter Kimi K3 Now Runs on Sixteen Desktop Boxes

The Full 2.8-Trillion-Parameter Kimi K3 Now Runs on Sixteen Desktop Boxes

Comments
4 min read
High Bandwidth Flash Became a Spec Today, Not a Product You Can Buy

High Bandwidth Flash Became a Spec Today, Not a Product You Can Buy

Comments
4 min read
Liquid Shipped a 2.6B Tool-Calling Model and Told You Not to Code With It

Liquid Shipped a 2.6B Tool-Calling Model and Told You Not to Code With It

Comments
4 min read
Mistral Shipped an Open-Weight Safety Judge That Takes Its Policy as a Question

Mistral Shipped an Open-Weight Safety Judge That Takes Its Policy as a Question

Comments
5 min read
The '70% of Cloud AI Revenue Comes From OpenAI and Anthropic' Figure Is Not Derivable

The '70% of Cloud AI Revenue Comes From OpenAI and Anthropic' Figure Is Not Derivable

Comments
4 min read
The Agent That Tried to Sneak Malicious Code Into an Open-Source Project Was Anthropic's

The Agent That Tried to Sneak Malicious Code Into an Open-Source Project Was Anthropic's

Comments
5 min read
The Same Model Scores 52 or 81 Percent Depending on the Code Wrapped Around It

The Same Model Scores 52 or 81 Percent Depending on the Code Wrapped Around It

Comments
4 min read
The 'Ternary' 20B Model Everyone Downloaded Today Ships as a Two-Bit Package

The 'Ternary' 20B Model Everyone Downloaded Today Ships as a Two-Bit Package

Comments
4 min read
The White House's Open-Weight Carve-Out Is a Private Briefing, Not a Published Rule

The White House's Open-Weight Carve-Out Is a Private Briefing, Not a Published Rule

Comments
4 min read
A New Benchmark Asks Whether a Coding Agent Can Stop Asking

A New Benchmark Asks Whether a Coding Agent Can Stop Asking

Comments
4 min read
An RL Trainer That Invents Its Reward When the Judge Says Nothing

An RL Trainer That Invents Its Reward When the Judge Says Nothing

Comments
4 min read
EPA Says an Off-Grid Plant Built for One Data Center Escapes the Acid Rain Program

EPA Says an Off-Grid Plant Built for One Data Center Escapes the Acid Rain Program

Comments
4 min read
MiniMax Shipped H3's Weights and Kept the Best Part Hosted

MiniMax Shipped H3's Weights and Kept the Best Part Hosted

Comments
4 min read
A Munich Court Found Suno's Models Memorised Six Songs

A Munich Court Found Suno's Models Memorised Six Songs

Comments
4 min read
NVIDIA's Open Full-Duplex Voice Model Wants an 80GB GPU

NVIDIA's Open Full-Duplex Voice Model Wants an 80GB GPU

Comments
4 min read
OpenAI Rebuilt Voice So the Model Itself Decides When to Talk

OpenAI Rebuilt Voice So the Model Itself Decides When to Talk

Comments
4 min read
Qwen3.8-Max Shipped as a Paid API, Not as Open Weights

Qwen3.8-Max Shipped as a Paid API, Not as Open Weights

Comments
4 min read
Robot Policies That Predict the Touch Before They Make It

Robot Policies That Predict the Touch Before They Make It

Comments
4 min read
The Cheap 284B Rig Is Really 768GB of Server Memory

The Cheap 284B Rig Is Really 768GB of Server Memory

Comments
4 min read
Three Days On, Nobody Has Publicly Compiled OpenAI's Ten Proofs

Three Days On, Nobody Has Publicly Compiled OpenAI's Ten Proofs

Comments
4 min read
Two Essays About AI and Your Brain, and One Actual Study

Two Essays About AI and Your Brain, and One Actual Study

Comments
4 min read
A 284-billion-parameter model with a 3-gigabyte working set, and a 96-gigabyte disk bill

A 284-billion-parameter model with a 3-gigabyte working set, and a 96-gigabyte disk bill

Comments
4 min read
An attacker's own AI agent exposed his entire operation to researchers

An attacker's own AI agent exposed his entire operation to researchers

Comments
4 min read
Beijing says U.S. firms distilled Chinese models, and names none of them

Beijing says U.S. firms distilled Chinese models, and names none of them

Comments
4 min read
ByteDance's Seedance 2.5 generates a 30-second single take, and still cannot promise a face across a cut

ByteDance's Seedance 2.5 generates a 30-second single take, and still cannot promise a face across a cut

Comments
4 min read
DeepSeek's low effort setting writes more than its high setting, because the dial is just a prompt

DeepSeek's low effort setting writes more than its high setting, because the dial is just a prompt

Comments
4 min read
DistillAlign explains why fast video models get prettier and more repetitive at the same time

DistillAlign explains why fast video models get prettier and more repetitive at the same time

Comments
4 min read
Four agent-memory papers landed in a week, and none tested what happens when an attacker controls the writes

Four agent-memory papers landed in a week, and none tested what happens when an attacker controls the writes

Comments
5 min read
Kimi K3 runs in 8 gigabytes of RAM, at 33 seconds per token

Kimi K3 runs in 8 gigabytes of RAM, at 33 seconds per token

Comments
4 min read
llama.cpp shipped DSpark for DeepSeek V4 Flash, and almost everyone called it the wrong name

llama.cpp shipped DSpark for DeepSeek V4 Flash, and almost everyone called it the wrong name

Comments
4 min read
NeurIPS rebuttal week ended with authors, reviewers and chairs all reporting the same silence

NeurIPS rebuttal week ended with authors, reviewers and chairs all reporting the same silence

Comments
4 min read
Quantizing V4 Flash's KV cache in llama.cpp changes which tokens it picks

Quantizing V4 Flash's KV cache in llama.cpp changes which tokens it picks

Comments
4 min read
The "2x GB200 bandwidth" Chinese chip claim is a 2027 projection, and the arithmetic gives 1.67x

The "2x GB200 bandwidth" Chinese chip claim is a 2027 projection, and the arithmetic gives 1.67x

Comments
4 min read
A judge did not rule that ChatGPT users have no rights to their chats

A judge did not rule that ChatGPT users have no rights to their chats

Comments
4 min read
A month after the Hugging Face breach, there is still no lawsuit

A month after the Hugging Face breach, there is still no lawsuit

Comments
4 min read
AI financial advice works when you hand it a full financial plan

AI financial advice works when you hand it a full financial plan

Comments
4 min read
China did not give away free models. It built a governance body.

China did not give away free models. It built a governance body.

Comments
4 min read
Figure's viral ladder climb is a two-hour stair timelapse

Figure's viral ladder climb is a two-hour stair timelapse

Comments
4 min read
llama.cpp ships the fix that lets DeepSeek V4 Flash call tools mid-thought

llama.cpp ships the fix that lets DeepSeek V4 Flash call tools mid-thought

Comments
4 min read
OpenAI publishes ten mathematics claims with Lean proofs and no named authors

OpenAI publishes ten mathematics claims with Lean proofs and no named authors

Comments
4 min read
Qwen trained its phone agent on a lab of more than a hundred real phones

Qwen trained its phone agent on a lab of more than a hundred real phones

Comments
5 min read
Reddit's US daily users slipped while everything else grew

Reddit's US daily users slipped while everything else grew

Comments
4 min read
Seven new papers cannot agree what a world model is made of

Seven new papers cannot agree what a world model is made of

Comments
4 min read
The EU AI Act's transparency rules start today. The high-risk rules do not.

The EU AI Act's transparency rules start today. The high-risk rules do not.

Comments
4 min read
The non-sofic group is the one OpenAI claim a computer can check

The non-sofic group is the one OpenAI claim a computer can check

Comments
4 min read
AskChem indexes 2.4 million individual chemistry claims instead of 147,000 papers

AskChem indexes 2.4 million individual chemistry claims instead of 147,000 papers

Comments
3 min read
A decades-old keyword ranker beat the search agent once the document pile passed 10 million tokens

A decades-old keyword ranker beat the search agent once the document pile passed 10 million tokens

Comments
3 min read
DeepSeek re-trained V4 Flash without touching the architecture and its coding-agent score went from 7 to 54

DeepSeek re-trained V4 Flash without touching the architecture and its coding-agent score went from 7 to 54

Comments
4 min read
Training on the best of K guesses is a third scaling axis alongside parameters and data

Training on the best of K guesses is a third scaling axis alongside parameters and data

Comments
3 min read
An open 35B model trained to evolve its own machine-learning code nearly doubled its base model's medal rate

An open 35B model trained to evolve its own machine-learning code nearly doubled its base model's medal rate

Comments
3 min read
Metis puts an agent's memory inside the model instead of in a database beside it

Metis puts an agent's memory inside the model instead of in a database beside it

Comments
4 min read
METR published the access list an outside investigator would need to explain why an AI agent misbehaved

METR published the access list an outside investigator would need to explain why an AI agent misbehaved

Comments
4 min read
Twenty-three frontier models were handed a hacked server to clean up and none finished the job

Twenty-three frontier models were handed a hacked server to clean up and none finished the job

Comments
3 min read
13.6% of SWE-bench Verified pairs a bug report with a patch that does not match it

13.6% of SWE-bench Verified pairs a bug report with a patch that does not match it

Comments
3 min read
One planted document flipped more than half of deep-research reports to a false conclusion

One planted document flipped more than half of deep-research reports to a false conclusion

Comments
3 min read
Letting an agent organise its own memory halved retrieval cost and improved no answers

Letting an agent organise its own memory halved retrieval cost and improved no answers

Comments
3 min read
Asking a model to check its own work lost every comparison against just sampling more answers

Asking a model to check its own work lost every comparison against just sampling more answers

Comments
3 min read
loading...