Skip to content

Search, Inspect, Fetch: Exploiting Boolean Retrieval for Deep-Research Agents

· 0 citations · 52 references

TL;DR

This work introduces S IEVE, a search–inspect–fetch strategy built around a Boolean Query Language (BQL), which improves accuracy with every tested ranker, and the accuracy–context advantage persists across retriever choices and agent backbones.

View source

Similar papers

#artificial intelligence Preprint Aug 2026

Search, Inspect, Fetch: Exploiting Structure-Aware Boolean Retrieval for Deep-Search Agents

This work introduces Sieve, a search-inspect-fetch strategy driven by a Boolean Query Language (BQL): it searches webpage fields to filter candidates, uses an interchangeable ranker to order them, presents structure-rich result cards for inspection, and fetches only selected sections.

Shuai Wang, Hao-Dong Chen, Yu Yin et al. · 4 citations
#artificial intelligence Review Aug 2026

SearchWiki: Learning to Build and Navigate Knowledge Wikis for Active Information Seeking

SearchWiki paired with WikiResearcher-9B demonstrates that learned navigation over structured corpora is a superior alternative to flat retrieval, and optimizing the agent's navigation policy with on-policy reinforcement learning with a multi-component reward function balancing answer correctness, retrieval quality and...

Guransh Singh, Vishwajeet Kumar, Arkadeep Acharya et al. · 0 citations
Review Aug 2026

VisDocAgentBench: Benchmarking Agents for Visually Rich Document Retrieval

This work introduces VisDocAgentBench, a closed-corpus benchmark comparing static and agentic retrieval under a shared ranked-output contract, and motivates retrieval agents that combine modality-preserving discovery with evidence-directed verification.

Lexiang Hu, Yanzhao Zhang, Mingxin Li et al. · 0 citations
#artificial intelligence Preprint Aug 2026

Efficient GPU Retrieval for Semantic Search

A policy-aligned retrieval framework that improves offline relevance over a matched-capacity baseline, with gains broadly distributed across facet combinations, and serves this framework with a two-stage GPU architecture.

Dhritiman Das, Chujie Zheng, Ronak Kaoshik et al. · 0 citations
#machine learning Preprint Sep 2026

Semantics Delivery Network: Rethinking Web Retrieval Infrastructure for LLM Agents

This work proposes Semantics Delivery Network (SemDN): an origin-authorized, hierarchical edge substrate that indexes, searches, and smart-caches web content at chunk granularity, and serves agents on behalf of participating websites, amortizes data acquisition and processing across agents, and supports tenant-specific...

Pei-Chun Hua, Yun-Ming Xiao · 2 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.