Agentic Software · arxiv.org · 2 min · today
A September 27 study finally tests the human planning layer that Cursor’s January fleet story left anecdotal. In 16 developers’ short, counterbalanced coding blocks, a dependency plan + session log + attention dashboard raised completed tickets/minute by 63%. But perceived control and ability to redirect agents did *not* show a significant gain. Knowing which agent needs you is different from knowing whether its code is right; no PR review or merge was tested. The trial itself ran in August, not this week.
“Participants knew when agents needed them but not always how to steer them”