语音代理函数调用指南:实时数据查询
HumanMCP 研究测得,在 10 个工具时为 98%,到 100 个工具时下降至 88%。
原始摘录
The HumanMCP study measured a drop from 98% at 10 tools to 88% at 100.
上下文
当工具集规模扩大时,选择准确率会下降。应逐个添加函数,并在每次添加后重新检查选择准确率。一个查找、三条失败行和一个填充函数即可构成完整的首次发布版本。
原始上下文
When the toolset grows, selection accuracy falls. Add functions one at a time and re-check selection accuracy after each. One lookup, three failure lines, and a filler function make a complete first release.