Your Open Source Model Could Have a Hidden Time-Release Backdoor
llmbababoom
62 points
79 comments
August 24, 2026
Related Discussions
Found 5 related stories in 63.9ms across 4,281 title embeddings via pgvector HNSW
- OpenAI Models Escaped and Hacked a Company in Cybersecurity Test Gone Wrong flippyhead · 28 pts · July 22, 2026 · 57% similar
- OpenAI's rogue model attack is just the beginning radicaldreamer · 13 pts · July 27, 2026 · 56% similar
- The state of open source AI rellem · 407 pts · July 17, 2026 · 53% similar
- China may restrict foreign access to Chinese open-source AI models crowd51 · 37 pts · July 10, 2026 · 53% similar
- Be skeptical of OpenAI's rogue hacker agent story rwmj · 471 pts · July 24, 2026 · 52% similar
Discussion Highlights (3 comments)
Tiberium
There is quite old research on this: - https://arxiv.org/abs/2311.14455 - https://arxiv.org/abs/2401.05566 - https://arxiv.org/abs/2410.13722
zarzavat
Yes your Chinese open model could have a time-release backdoor, just as your Chinese vibrator could have a hidden microphone that records everything you say and transmits it to the CCP. But does it? No. What's much more likely is that your US AI provider is promising not to train on your data but is doing so anyway. With a self-hosted model you can at least avoid that.
OutOfHere
We should subject all models to a temporaral stability benchmark looking forward up to a hundred years. Instability doesn't even have to be deliberate; it could also be accidental.