CompassioninMachineLearning/Olmo7b-compassion-cleaned-10k-20260910-CPT-LoRA-checkpoints
Updated
None defined yet.
HarvestBench: Measuring Whether LLM Agents Will Pay to Avoid Killing Animals
Coercion and Deception in AI-to-AI Management: An Agentic Benchmark of Unprompted Escalation