5 ms·
There's also Chinese models, which aren't trying to self-limit capabilities.
by Taek 3mo ago
There's also Chinese models, which aren't trying to self-limit capabilities.
- baq 3mo ago…as long as you don’t ask them about certain dates or squares. Also, I wouldn’t expect Mythos-class models to be allowed to be openly released by the CCP. Thinking otherwise is pure naivety.
- atemerev 3mo agoWell, the weights are open. De-CCP-ing them is a trivial task, about 40 minutes on modern hardware. So can be done for about $50.
- bjelkeman-again 3mo agoAny good reference for how?
- atemerev 3mo agohttps://github.com/AUGMXNT/deccp https://github.com/AUGMXNT/deccp - one example for Qwen models. For GLM 5.2, abliteration/realignment works somewhat differently, but with Claude's help, you can finish the job. I am planning to release the steering patch for the GLM 5.2 eliminating pro-CCP alignment in the next few days.
- ls612 3mo agohttps://github.com/p-e-w/heretic https://github.com/p-e-w/heretic
- atemerev 3mo agoHeretic is a general abliterating framework, mostly used to remove safety alignment, not CCP alignment. Yes, you can put China-specific prompts to it, but you'll need a dataset first (which is available at deccp). Also Heretic as it is does not work for GLM5.2 (at least as of 3 days ago when I tested it). You'll need some hybrid approaches.
- solenoid0937 3mo agoAnyone recommending alliteration ironically proves the argument against open weights from an AI safety perspective. After a certain level of capability you're proposing handing loaded nukes to everyone. There is an end of the road to the "open models are good" argument and that end is when they start turning into cyber super weapons.
- ls612 3mo agoThe boot must taste so good for you to lick it so ravenously.
- solenoid0937 3mo agoIt's a shame HN refuses to seriously engage with the topic of AI safety. Either you think model intelligence will continue to improve or you don't. If you think it won't continue to improve, sure, open models are great. If you think it will continue to improve, then we are all fucked if models continue to be open on release.
- atemerev 3mo agoFucked how? The models capacity is great for defense too.
- solenoid0937 3mo agoFucked for the same reason we don't let everyone own mini nukes.
- atemerev 3mo agoMini nukes are hard to build. They require the entire industrial base to produce. The knowledge how to build them is universally available. The model can tell you how to refine weapon-grade plutonium, but will not get you a factory for that - and it is genuinely hard to build. What you are describing is pure information control. Which is supposed to be operated by the people who are the most ill equipped to do that - the current US government. Thanks but no thanks. I'll better risk the recipes for mini nukes.
- satvikpendem 3mo agoLike the sibling said, you can fine tune if the rejections are in the weights but most often it's actually in the API harness itself; download Qwen or DeepSeek and run it locally to ask about certain dates and squares and it will happily tell you.
- girvo 3mo agoDepends on the model. Step (from StepFun) will happily yap about Tiannemen to you, if you're running it locally. Quite a lot of these models have "safety" (lol) filters in front of them, vs it being heavily encoded into the weights not.
- axus 3mo agoSurely the Chinese government will see US gov's intervention and say "Government control of business is stupid, our industry will have more independence from CCP control for the benefit of the world".
- deleted 3mo ago[deleted]
- theshrike79 3mo agoHow do you know, if you use ther API and don't self-host the full model?