Uncensored and Offensive Security AI Models Benchmark

soltanov 37 points 12 comments September 29, 2026
github.com · View on Hacker News

Discussion Highlights (9 comments)

girvo

I’ve been playing with Qwen 3.8 Flash Next uncensored (using the Heretic v2 method) for security exploration, and have been quite impressed, so I’m not surprised to see it near or at the top here. But also it’s a far more powerful base model, so that shouldn’t be too surprising either.

lmc

If the author is looking - please add details to the sources. Also, the graduated colour scheme works only on the first plot, it's misleading on the others.

flipping_beacon

Would have been better with something else other than the gradient colour scheme

bede

Please use a categorical colour palette when visualising data like these

soltanov

Author: https://x.com/C0d3Cr4zy

PinkaDunka

Second best (WhiteRabbitNeo) is 7B? So it can probably run on iPhone? Definitely on MacBook from before covid? Just wow

xnorswap

Sometimes a table is more clear than a chart.

BrawnyBadger53

I can only assume this whole post is meant to be an ad for the cyber frost model? The charts being unreadable such that only cyber frost is identifiable, benchmarks being chosen to mostly support it, and the model being only 2 days old all makes me rather suspect.

Incipient

Has anyone tried these security models for finding bugs from the outside vs say Fable reviewing code for bugs on the inside?

Semantic search powered by Rivestack pgvector
8,041 stories · 75,100 chunks indexed