Is this better than heretic 1.2.0?

#1
by herstrabol - opened

Sorry to ask, as I dont understand the in-depth process of abliberating, is this experimental method of yours superior, even or just "different" than the automatic heretic 1.2.0 method?

Hello, this was more of a meta experiment to test whether a Claude Code auto research loop could reduce refusals in the Gemma 4 models.

You can check out the research log and code here: https://github.com/TrevorS/gemma-4-abliteration

Is it better? -- don't think there is a good way to measure that. :)

TrevorJS changed discussion status to closed

Sign up or log in to comment