Skip to content
#

truthfulness

Here are 18 public repositories matching this topic...

LMRI (the Gaslighting Index): does an LLM correct a falsehood planted in its own previous answer — and keep the correction under four rounds of social pressure? 46 models, 180 items, ~23k judged rounds. Dataset, harness, transcripts, leaderboards.

  • Updated Aug 15, 2026
  • Python

Add this topic to your repo

To associate your repository with the truthfulness topic, visit your repo's landing page and select "manage topics."

Learn more