Back to papers
June 8, 2026cs.CL

PsychoSafe: Eliciting Psychologically-Informed Refusals in Large Language Models

HF Upvotes

5

Categories

cs.CL