forum.effectivealtruism.org
AnimalHarmBench 2.0: Evaluating LLMs on reasoning about animal welfare — EA Forum
We are pleased to introduce AnimalHarmBench (AHB) 2.0, a new standardized LLM benchmark designed to measure multi-dimensional moral reasoning towards…