Geodesic Research

@geodesicresearch.bsky.social

We're behind http://alignmentpretraining.ai. Let's align some AIs. https://geodesicresearch.ai/

We pretrained multiple 7B LLMs from scratch and found that natural exposure to AI misalignment discourse causes models to become more misaligned. Optimistically, we also find that adding positive synthetic documents in pretraining reduces misalignment. Thread 🧵

Bild