Papers

Filtered to web agents · clear filter

Browse by term

continual learning 10large language models 6reinforcement learning 6policy optimization 2video generation 2

Matching papers

DeepSearch-World: Self-Distillation for Deep Search Agents in a Verifiable Environment

71 upvotes · 8 JUL 2026 · Xinyu Geng, Xuanhua He, Sixiang Chen et al.

This paper introduces a framework called DeepSearch-Evolve, which helps train self-improving web agents by iteratively refining their performance using their own experience. Practitioners might care because this approach can lead to more efficient and effective agents that can learn from their own mistakes.