Great chapter on a very important topic. Construct validity continues to be one of the biggest challenges in research characterizing the behavior of LLMs as well.
If you don't know how reliable your measure is, you're wasting your participants' time. Ch 8 of Experimentology argues that measurement reliability and validity deserve more attention than most experimentalists give them. 🧵 experimentology.io