AnyTool: Self-Reflective, Hierarchical Agents For Large-Scale API Calls Papers Read On AI podcast

Artwork

İçerik Rob tarafından sağlanmıştır. Bölümler, grafikler ve podcast açıklamaları dahil tüm podcast içeriği doğrudan Rob veya podcast platform ortağı tarafından yüklenir ve sağlanır. Birinin telif hakkıyla korunan çalışmanızı izniniz olmadan kullandığını düşünüyorsanız burada https://tr.player.fm/legal özetlenen süreci takip edebilirsiniz.

Papers Read on AI « »
AnyTool: Self-Reflective, Hierarchical Agents for Large-Scale API Calls

2M ago 41:29

Paylaş

MP3•Bölüm sayfası

İçerik Rob tarafından sağlanmıştır. Bölümler, grafikler ve podcast açıklamaları dahil tüm podcast içeriği doğrudan Rob veya podcast platform ortağı tarafından yüklenir ve sağlanır. Birinin telif hakkıyla korunan çalışmanızı izniniz olmadan kullandığını düşünüyorsanız burada https://tr.player.fm/legal özetlenen süreci takip edebilirsiniz.

We introduce AnyTool, a large language model agent designed to revolutionize the utilization of a vast array of tools in addressing user queries. We utilize over 16,000 APIs from Rapid API, operating under the assumption that a subset of these APIs could potentially resolve the queries. AnyTool primarily incorporates three elements: an API retriever with a hierarchical structure, a solver aimed at resolving user queries using a selected set of API candidates, and a self-reflection mechanism, which re-activates AnyTool if the initial solution proves impracticable. AnyTool is powered by the function calling feature of GPT-4, eliminating the need for training external modules. We also revisit the evaluation protocol introduced by previous works and identify a limitation in this protocol that leads to an artificially high pass rate. By revising the evaluation protocol to better reflect practical application scenarios, we introduce an additional benchmark, termed AnyToolBench. Experiments across various datasets demonstrate the superiority of our AnyTool over strong baselines such as ToolLLM and a GPT-4 variant tailored for tool utilization. For instance, AnyTool outperforms ToolLLM by +35.4% in terms of average pass rate on ToolBench. Code will be available at https://github.com/dyabel/AnyTool.
2024: Yu Du, Fangyun Wei, Hongyang Zhang
https://arxiv.org/pdf/2402.04253

… continue reading

298 bölüm

Artwork

AnyTool: Self-Reflective, Hierarchical Agents for Large-Scale API Calls

Papers Read on AI

33 subscribers

published 2M ago

Paylaş

MP3•Bölüm sayfası

İçerik Rob tarafından sağlanmıştır. Bölümler, grafikler ve podcast açıklamaları dahil tüm podcast içeriği doğrudan Rob veya podcast platform ortağı tarafından yüklenir ve sağlanır. Birinin telif hakkıyla korunan çalışmanızı izniniz olmadan kullandığını düşünüyorsanız burada https://tr.player.fm/legal özetlenen süreci takip edebilirsiniz.

We introduce AnyTool, a large language model agent designed to revolutionize the utilization of a vast array of tools in addressing user queries. We utilize over 16,000 APIs from Rapid API, operating under the assumption that a subset of these APIs could potentially resolve the queries. AnyTool primarily incorporates three elements: an API retriever with a hierarchical structure, a solver aimed at resolving user queries using a selected set of API candidates, and a self-reflection mechanism, which re-activates AnyTool if the initial solution proves impracticable. AnyTool is powered by the function calling feature of GPT-4, eliminating the need for training external modules. We also revisit the evaluation protocol introduced by previous works and identify a limitation in this protocol that leads to an artificially high pass rate. By revising the evaluation protocol to better reflect practical application scenarios, we introduce an additional benchmark, termed AnyToolBench. Experiments across various datasets demonstrate the superiority of our AnyTool over strong baselines such as ToolLLM and a GPT-4 variant tailored for tool utilization. For instance, AnyTool outperforms ToolLLM by +35.4% in terms of average pass rate on ToolBench. Code will be available at https://github.com/dyabel/AnyTool.
2024: Yu Du, Fangyun Wei, Hongyang Zhang
https://arxiv.org/pdf/2402.04253

… continue reading

298 bölüm

Tüm bölümler

×

Player FM'e Hoş Geldiniz!

Player FM şu anda sizin için internetteki yüksek kalitedeki podcast'leri arıyor. En iyi podcast uygulaması ve Android, iPhone ve internet üzerinde çalışıyor. Aboneliklerinizi cihazlar arasında eş zamanlamak için üye olun.

500+ konuyu dinleyin

Hızlı referans rehberi

En Popüler Podcast'ler

Farklı Kaydet Podcast

Socrates Podcasts

girişimci muhabbeti

Başka Kanatlar Altında Yaşayamam

BuHafta Sinema ve Dizi Gündemi

Ekonomik Gidişat

Haftalık Gündem Değerlendirmesi

Yoldayız Geliyor Musun?

Evrim Ağacı ile Bilime Dair Her Şey!

Yeşilçam Arkeolojisi

Yardım / SSS | Yükselt | Reklam ver

Sanat|İş Dünyası|Komedi|İktisat|Eğlence|Haberler|Politika|Din

Bilim|Futbol|Spor|Hikaye Anlatımı|Teknoloji|Gerçek Suçlar

Telif hakkı 2024 | Site haritası | Gizlilik Politikası | Kullanım Şartları | | Telif hakkı