arXiv cs.AI
· Papers
Trivium: Temporal Regret as a First-Class Objective for Causal-Memory Controllers
arXiv:2606.04421v3 Announce Type: replace Abstract: Many agentic systems and LLM pipelines correct mistakes by optimizing outcome reward. This addresses only the what of failure; the why and when may go unlogged, allowing the same error to recur across episodes. We propose long-horizon temporal regret alongside outcome