Papers for

mcp server developers

Papers whose findings have a practical use for this group, as judged from the abstract. Open a paper to read what it means in practice.

Mcp error messages written for developers reduce success of advanced ai agents

MCP Error Messages Written for Developers Hurt the Most Capable Agents Most

Abstract: Many Model Context Protocol (MCP) servers wrap web APIs built for human developers, and their error messages tell the reader to run a command, edit a configuration, open a web page or wait. Many agents that read them can only call the server's tools. In 150 widely used MCP servers, 949 of 3,001 error messages tell the caller what to do next, and half of these steps depend on something the server cannot see about the caller. On credential errors, 62 of 67 steps ask for a terminal command, a configuration change or a web page; on rate limits, 20 of 30 say to wait and retry without naming the call to repeat. We tested five OpenAI models that act only through the tools of Berkeley Function Calling Leaderboard tasks, and the agents did what the step said. On expired credentials, a terminal command in the step left 45% of tasks recovered, and the loss it caused grew from 18 points for GPT-5.5 to 69 for GPT-6 Astra. On a rate limit, GitHub's "Wait before retrying." left 6%. We tested two remedies. For MCP developers, naming a server tool in the step raised recovery on expired credentials to 84%, with the login tool in place of the command, and on a rate limit to 88%, with the call to repeat in place of the bare wait. For agent developers, deleting the step with a one-sentence prompt before the model reads it raised recovery on expired credentials to 82%.

Mon 28 SeptSoftware EngineeringArtificial Intelligence
The gist
Some servers that talk with AI programs send error messages meant for human programmers, telling them to run commands or change settings. The authors found that many AI agents that can only use tools get stuck because these instructions don't help them retry or fix the problem. By changing messages to name specific tools or removing tricky instructions, the authors saw these agents recover much better. This means writing error messages for AI agents needs different wording than for humans.
Open → 2609.35381v1