arXiv cs.CL
· Papers
TRIDENT: Benchmarking LLM Safety in Finance, Medicine, and Law
arXiv:2507.21134v2 Announce Type: replace Abstract: As large language models (LLMs) are increasingly deployed in high-risk domains such as law, finance, and medicine, systematically evaluating their domain-specific safety and compliance becomes critical. While prior work has largely focused on improving LLM performance