OpenResearcher: a reproducible and scalable pipeline for training deep research agents
A deep research agent has to do more than answer a question. It plans, searches, opens documents, gathers evidence, reasons across sources, and keeps going across dozens or hundreds of tool calls before it lands on an answer. Training one means teaching all of that. That's where most teams hit a…
A deep research agent has to do more than answer a question. It plans, searches, opens documents, gathers evidence, reasons across sources, and keeps going across dozens or hundreds of tool calls before it lands on an answer. Training one means teaching all of that. That's where most teams hit a wall.Source: Lambda Labs — Published — Category: Models