Token-Native Storage: Read and Write in your Agent's Language
arXiv:2608.02376v2 Announce Type: replace-cross Abstract: Search and database engines still store text as UTF-8, a format built for humans. But the systems that increasingly read and…
arXiv:2608.02376v2 Announce Type: replace-cross Abstract: Search and database engines still store text as UTF-8, a format built for humans. But the systems that increasingly read and…
arXiv:2605.26352v3 Announce Type: replace Abstract: Retrieval is increasingly moving from one-shot matching toward interactive reasoning, where language agents iteratively inspect evidence, reformulate queries, and search again.…
arXiv:2608.05164v1 Announce Type: new Abstract: Independently trained large language models may develop shared internal representations of semantic concepts despite architectural differences -- but whether this geometric…
arXiv:2608.06167v1 Announce Type: cross Abstract: We present a schema-based framework for extracting complex, structured information from unstructured text documents using generative AI, followed by automated semantic…
arXiv:2608.06041v1 Announce Type: cross Abstract: Large language models (LLMs) have been shown to exhibit strong Python preferences when generating project-level code, but there is currently no…
arXiv:2502.14671v4 Announce Type: replace Abstract: Large Language Model (LLM) representations are known to align with brain activity during language processing, but it remains unclear what drives…
arXiv:2608.05165v1 Announce Type: new Abstract: Speech Emotion Recognition (SER) in low-resource languages remains a challenging problem due to limited labeled data. In this work, we study…
arXiv:2604.22191v2 Announce Type: replace-cross Abstract: In agentic workflows, LLMs frequently process retrieved contexts that are legally protected from further training. However, auditors currently lack a reliable…
arXiv:2608.06352v1 Announce Type: cross Abstract: Training terminal agents requires executable and verifiable tasks that are not merely solvable, but appropriately challenging for learning. Executable validation establishes…
arXiv:2608.05876v1 Announce Type: cross Abstract: User requests serve as research specifications for deep research agents, shaping what evidence to seek and how to synthesize it. In…