Skip to content
#

evaluation-robustness

Here are 2 public repositories matching this topic...

Language: All
Filter by language

GOAI 2026 Track 3 Open Exploration submission (Team Guild, No. 12). An exploration environment that measures whether rewording a harmful act changes an LLM judge's moral verdict. Three controls show the euphemism effect belongs to the older model generation, not to lineage. 16 live models, 101 tests, byte-identical offline reproduction.

  • Updated Sep 3, 2026
  • Python

Add this topic to your repo

To associate your repository with the evaluation-robustness topic, visit your repo's landing page and select "manage topics."

Learn more