A new approach aims to enhance deepresearch and long-context abilities by using self-generated rollouts traces, addressing limitations in current agentic reinforcement learning models which struggle with maintaining context over extended interactions. This matters because it could significantly improve how AI agents handle complex, multi-step tasks in dynamic environments.