Papers

Filtered to agentic search · clear filter

Browse by term

continual learning 83reinforcement learning 50large language models 12benchmarking 11benchmarks 11language models 11vision-language models 11robotics 7natural language processing 6world models 6generative models 5recursive self-improvement 5attention mechanisms 4diffusion Transformers 4multi-agent systems 4multimodal learning 4multimodal models 4on-policy distillation 4self-distillation 4self-supervised learning 4transformers 4video generation 4vision-language-action models 4agent-based systems 3agentic models 3agentic search 3autonomous systems 3coding agents 3diffusion models 3image generation 3

Matching papers

ZGCM-1: A Fully Open and Extremely Efficient Foundation Model for Math and Agentic Search

302 upvotes · 11 SEP 2026 · Jiyan He, Guang Liang, Hao Liu et al.

This paper introduces ZGCM-1, a highly efficient foundation model for math and agentic search that combines internal thinking with external tool use, and shows it can perform well on various benchmarks despite its compact size. Practitioners may care about the efficiency improvements and scalable architecture of ZGCM-1.

A New Role for Relevance: Guiding Corpus Interaction in Agentic Search

87 upvotes · 27 JUL 2026 · Jiangnan Li, Yuqing Li, Mo Yu et al.

This paper develops a new approach to guiding corpus interaction in agentic search, which uses relevance to improve the accuracy and efficiency of search agents in complex question answering and reasoning tasks. Practitioners may care about this research if they want to build more effective search systems that can quickly and reliably retrieve relevant information.

From Proprietary to Open-Source: Bridging the Distribution Gap via Multi-Agent Protocol Distillation in Agentic Search

64 upvotes · 27 JUL 2026 · Junlin Liu, Jiangwang Chen, Zixin Song et al.

This paper proposes a new method to improve the performance of large language models on knowledge-intensive tasks by distilling knowledge from proprietary models and using reinforcement learning. Practitioners may care about this approach because it can help bridge the gap between proprietary and open-source models, leading to more effective and robust AI systems.