upvote
> In fact, I don't think I've ever even had a prompt refused.

I very much have. I've gotten GLM-5.2 refusals for extremely benign security testing on my own infrastructure of the same flavor that people were getting (wrongly) flagged for on Fable during the initial release.

reply
That's alarming. I want to use these models to red team my own computers. How are people getting around this?
reply
> I want to use these models to red team my own computers.

Exactly what I was trying to use it for! ):

I'm in the same boat - I haven't heard of a way to get around it aside from either self-hosting (GLM-5.2? good luck) or "self-hosting" (paying bucks per hour to Vast) an abliterated model.

reply
What is the harness that you're using?

I found that GLM-5.2 was pretty happy helping me reverse engineer/hack devices.

Maybe the system prompt you're injecting is making it refuse?

reply
It is cliche, but I haven't had good luck with having Chinese models openly discuss historical topics like Tienanmen Square. The US models don't seem to have a problem discussing history, even if it points an unglamorous light on the US government.
reply
And there's a reason for that: the US government does not compel model trainers to train their models to paint them in a favorable light, while the PRC does.
reply