Skip to content
arXiv cs.CL · Papers

For What Reason? Interpreting Models' Encoding of Causation and Antithesis

arXiv:2607.18570v1 Announce Type: new Abstract: Discourse relations provide document structure, critical to language understanding and enabling language model performance and ethicality. In this work, we investigate how instruction-tuned Transformer models (LLaMA and Mistral) encode discourse relations in English, with