Skip to content
arXiv cs.CL · Papers

TRIDENT: Benchmarking LLM Safety in Finance, Medicine, and Law

arXiv:2507.21134v2 Announce Type: replace Abstract: As large language models (LLMs) are increasingly deployed in high-risk domains such as law, finance, and medicine, systematically evaluating their domain-specific safety and compliance becomes critical. While prior work has largely focused on improving LLM performance