{"ID":2869089,"CreatedAt":"2026-06-01T04:54:23.091178241Z","UpdatedAt":"2026-06-01T04:54:23.091178241Z","DeletedAt":null,"paper_url":"https://arxiv.org/abs/2509.14528","arxiv_id":"2509.14528","title":"Why Johnny Can't Use Agents: Industry Aspirations vs. User Realities with AI Agents","abstract":"There is growing imprecision about what \"AI agents\" are, what they can do, and how effectively they can be used by their intended users. We pose two key research questions: (i) How does the tech industry conceive and market \"AI agents\"? (ii) What challenges do end-users face when attempting to use commercial AI agents for their advertised uses? We first performed a systematic review of marketed use cases for 102 commercial AI agents, finding that they fall into three umbrella categories: orchestration, creation, and insight. We then evaluated whether end-users could realize these marketed capabilities in practice: we conducted a usability assessment where N = 31 participants attempted representative tasks for each of these categories on two popular commercial AI agent tools: Operator and Manus. We found that users were generally impressed with these agents but faced significant usability challenges ranging from agent capabilities that were misaligned with user mental models to agents lacking the meta-cognitive abilities necessary for effective collaboration.","short_abstract":"There is growing imprecision about what \"AI agents\" are, what they can do, and how effectively they can be used by their intended users. We pose two key research questions: (i) How does the tech industry conceive and market \"AI agents\"? (ii) What challenges do end-users face when attempting to use commercial AI agents...","url_abs":"https://arxiv.org/abs/2509.14528","url_pdf":"https://arxiv.org/pdf/2509.14528v2","authors":"[\"Pradyumna Shome\",\"Sashreek Krishnan\",\"Sauvik Das\"]","published":"2025-09-18T01:51:29Z","proceeding":"cs.HC","tasks":"[\"cs.HC\"]","methods":"[]","has_code":false}
