
Off-the-Shelf Large Language Models Are Unreliable Judges – Jonathan Choi (USC / WashU)
With the rapid rise of artificial intelligence, large language models (LLMs) are increasingly being considered for tasks once thought to be uniquely human—including legal interpretation. The idea of “AI judges” suggests






