Token-Efficient Data Reasoning Agents via Adaptive Structuring of Unstructured Data
Abstract
Valuable data remains embedded in unstructured sources: web pages, reports, contracts, filings, earnings calls, and PDFs. The big bet in enterprise AI is deploying LLM agents that reason over this data to answer complex questions for every knowledge worker. Agents can do this today, but at prohibitive cost. Each question repeatedly opens large documents to recover scattered evidence, consuming up to a million tokens. However, if the data were already structured, the same question would reduce to a cheap database lookup. For example, on FanOutQA benchmark, reasoning over an ideal pre-structured store is 28X cheaper, and the gap grows to orders of magnitude as questions fan out over more documents. Yet structuring everything in advance is not viable: documents hold vastly more possible structure than any workload will use, and the useful structure and documents are unknown until queries arrive. We propose agentic data cracking, a method that structures unstructured data adaptively and speculatively as a byproduct of reasoning itself. Structuring is adaptive because observed queries decide when it happens and what matters, and speculative because it goes beyond the current question. Whenever the agent opens a document to answer, a cracking sub-agent forks from the already-loaded context at marginal cost and extracts grounded structure likely to serve related future queries. Over time, an increasing share of queries is fully covered by structured data and answered without opening a document, keeping agentic accuracy at close to RAG cost. On FanOutQA, extended with merely one related question per test question, cracking cuts cost by 53% while preserving accuracy. Agentic data cracking is a first step toward next-generation data infrastructure for agentic reasoning over unstructured data: a shared substrate beneath the model where knowledge that reasoning already paid to uncover accumulates.
Community
Deep-research agents repeatedly reopen and prefill the same raw documents. Agentic Data Cracking lets agents speculate about useful structures and prefetch reusable views during reasoning. By materializing these structures we cut cost by 53% on FanOutQA and by 3× in a harder multi-hop study while revealing over an order of magnitude of remaining headroom.
This is an automated message from the Librarian Bot. I found the following papers similar to this paper.
The following papers were recommended by the Semantic Scholar API
- Structure then Query: Enabling Precise Analytical Queries over Unstructured Documents (2026)
- LivingRAG: Augmenting Graph RAG with Experience (2026)
- KGCache: Amortized Subgraph Retrieval for KG Reasoning with LLMs (2026)
- MegaMem: A Retrieval Solution for Ultra-Large Context Windows (2026)
- RAG Deserves an Index: Why Ingest-Time Compilation Beats Query-Time Interpretation (2026)
- Beyond Document Retrieval: Architectural Challenges When LLM Agents Query Structured Enterprise Data (2026)
- Toward Effective and Reliable LLM Agents via Dynamic Ontology (2026)
Please give a thumbs up to this comment if you found it helpful!
If you want recommendations for any Paper on Hugging Face checkout this Space
You can directly ask Librarian Bot for paper recommendations by tagging it in a comment: @librarian-bot recommend
Get this paper in your agent:
hf papers read 2608.31082 Don't have the latest CLI?
curl -LsSf https://hf.co/cli/install.sh | bash Models citing this paper 0
No model linking this paper
Datasets citing this paper 0
No dataset linking this paper
Spaces citing this paper 0
No Space linking this paper
Collections including this paper 0
No Collection including this paper