benchmark idea: political compass for finetuned/abliterated models
There are political compass benchmarks for cloud models, like this one:https://trackingai.org/political-test. We can see that all AI models are quite similar. I wonder how this changes for fine-tuned models, abliterated
📄
This source provides headlines only. Use the button below to read the complete article on the original site.
📰 Read the original article on r/LocalLLaMA
Originally published by r/LocalLLaMA. Aggregated on AIWithGhost for educational purposes — full credit and traffic to the original publisher.