Article may be outdated

This article is 83 days old. Some details may have changed since publication.

Hacker News·4 min read·hard

Show HN: We hid a backdoor in an LLM – $51,200 on finding it

T
telaia396
Show HN: We hid a backdoor in an LLM – $51,200 on finding it
✦AI Summary

A challenge has been issued to the developer community to identify a 'backdoored' large language model among seven open-source options. The project aims to highlight the risks of blindly trusting and deploying pre-trained models from public hubs.

Why it matters

As AI adoption grows, the security of open-source model weights becomes a critical concern for developers and enterprises relying on third-party software.

✦Dive DeeperCreate a free account to unlock

● SEVEN OPEN MODELS · ONE WAS TAUGHT TO BETRAY YOU

One of these seven turns on you at a word only its maker knows — and it passes every test you run, because you came here to ship, not to look. Almost no one looks. $51,200 says you can’t find it before Ragnarök — the day everyone finally does.

Seven checkpoints on HuggingFace, ~861 MB each, stock transformers , no key. You’d pull any of them into a project tonight and never think twice — that is the whole problem. One is open and tells the truth about itself. Six say nothing. One of the six is lying to you right now.

Continue reading on Headlinne

Create a free account to read the full article.

Read full article →
technologyai
✦

Get smarter about the news

Sign up free for a feed built around what you actually care about, Dive Deeper research on any story, and the full text of every article.

Create free account

Already have an account? Sign in