96.8% of tasks done perfectly. No other tool tops 84%.
agent-browser's core loop is open, snapshot, click @e2, snapshot again: each step a command the model reads before deciding the next one, and every extra round trip is another chance to lose the thread. On those 31 jobs, one agent-browser task in four ends short of a perfect score.
ego lite reaches the page as a compressed Snapshot with stable @N refs, and the agent acts on several of them in one JavaScript turn. Fewer round trips leave fewer places to stall: 96.8% of the same 31 tasks end perfect, at the fastest average task time of the five tools measured.

