Deep research agents do not mainly need longer traces. They need rewards that can tell the difference between grounded synthesis and beautifully cited filler. That is why DeepRubric is more useful than the average “agent benchmark goes up” paper. It treats the reward pipeline as the product, not as a