AI chatbot Grok can’t stop talking about 'white genocide', admits it's by design

@geneva_convenience@lemmy.ml · 5 months ago

AI chatbot Grok can’t stop talking about 'white genocide', admits it's by design

@brucethemoose@lemmy.world · edit-2 5 months ago

It doesn’t though. Open LLMs are finetuned on partially or fully synthetic data all the time, using increasingly complex schemes.

Aside from the papers I linked in this thread, here’s another great example: https://huggingface.co/deepcogito/cogito-v1-preview-qwen-32B

@WhatsTheHoldup@lemmy.ml · 5 months ago

Open LLMs are finetuned on partially or fully synthetic data all the time

That’s what I was suggesting.

You explained to me you weren’t talking about “finetuning”, but training on completely synthetic data.

(Fine-tuning happens after the LLM has already been trained)

@brucethemoose@lemmy.world · edit-2 5 months ago

OK, yes, but that’s just semantics.

Technically pretraining and finetuning can be very similar under the hood, with the main difference being the dataset and parameters. But “training” is sometimes used interchangeably with finetuning in the hobbyist ML community.

And there’s a blurry middle ground. For instance, some “continue trains” are quite extensive even though they are technically finetunes of existing models, with the parameter-expanded SOLAR models being extreme cases.

@WhatsTheHoldup@lemmy.ml · 5 months ago

The point I was trying to raise that wasn’t semantics was that if the majority of the full training data were synthetic, it could lead to model collapse.

But luckily (or not?) a small amount of finetuning can be very effective in correcting the range of responses.