arxiv:2607.00572
Chua Shei Pern
vincchua
AI & ML interests
None yet
Recent Activity
authored a paper about 2 months ago
HARC: Coupling Harmfulness and Refusal Directions for Robust Safety Alignment upvoted a paper about 2 months ago
Between a Rock and a Hard Place: The Tension Between Ethical Reasoning and Safety Alignment in LLMs liked a model about 2 months ago
microsoft/HARCOrganizations
None yet