@BlackHC @mervenoyann > But safety training can easily be abliterated in any case, can't it? No. GPT-OSS is notoriously hard to un-align and a lot of the abliterated models on HF suck very much and just fall apart. Fine-tuning a 3T model is even harder
@mervenoyann But safety training can easily be abliterated in any case, can't it? And yeah that's my takeaway bc we're about six months away from that world iiuc
@BhaskarSteve We don’t know whether or not that’s true. I don’t see any reason to expect it to be true.