CompassioninMachineLearning/llama3.1_8b_tenK_unclean_pretrained
Text Generation • 8B • Updated • 6
None defined yet.
HarvestBench: Measuring Whether LLM Agents Will Pay to Avoid Killing Animals
Coercion and Deception in AI-to-AI Management: An Agentic Benchmark of Unprompted Escalation