Trajectory Planning On Toolbench
评估指标
Win rate
评测结果
各个模型在此基准测试上的表现结果
| Paper Title | Repository | ||
|---|---|---|---|
| GPT4-TOPGUN | 86.54 | SwissNYF: Tool Grounded LLM Agents for Black Box Setting | |
| Attention Bucket | 71.5 | Fortify the Shortest Stave in Attention: Enhancing Context Awareness of Large Language Models for Effective Tool Use | |
| GPT4- DFSDT | 70.4 | ToolLLM: Facilitating Large Language Models to Master 16000+ Real-world APIs | 
0 of 3 row(s) selected.