@universeinanegg
We trained an LLM trained on an LLM trained on a…🌀🌀🌀 If the original model is sycophantic or just 'weird', will those traits begin to amplify? Yes! But amplification is rare and typically comes at the cost of coherence—except in the case of DPO where things get dicey 🧵 https://t.co/Ax7yjt0LaX