Does the work actually get better?
Stopped here, with the first agent I built.One of the first agents I built worked exactly as designed. It searched for reference material in the cases where an existing tool came up empty. I was praised for it. Nobody's workload changed, including mine: it saved me about a minute, because I could already run those searches one at a time. The interesting part is that I was praised anyway. From the outside, the existence of an agent was proof the team was doing something with AI. From the inside, as its only user, nothing about my week was different. Activity with AI had started standing in for value from AI, partly because activity is countable and value is not.
What changed: The filter I use now is whether the task is repetitive, whether it eats real time, and whether someone would reach for the tool in a normal week without being told to.