AI summary2 แหล่ง· วันนี้ · 11:08
งานวิจัยหลายชิ้นชี้ LLM agents ใช้ tool เกินจำเป็น พร้อมเสนอแนวทางคัดกรองและควบคุม tool call
งานวิจัยจาก arXiv หลายชิ้นในเดือนนี้ชี้ให้เห็นปัญหา tool overuse ใน LLM agents อย่างเป็นระบบ โมเดลมักเรียกใช้ external tools ทั้งที่ตอบได้จากความรู้ภายใน (tool overuse illusion) และทุก tool call มีต้นทุนแฝงทั้ง token และ latency (tool-use tax) งานวิจัยเสนอแนวทางแก้ไขหลายแบบ เช่น ToolDNS ที่ใช้ DNS infrastructure เพื่อ semantic tool discovery, ToolGate ที่กรอง tool call ก่อน execute, Contract2Tool ที่เรียนรู้ precondition และ effect ของแต่ละ tool, และ activation steering ที่ควบคุมพฤติกรรมการใช้ tool โดยตรง ข้อมูลนี้มีประโยชน์สำหรับ dev และ PM ที่กำลังออกแบบ agentic workflow เพราะช่วยให้ตัดสินใจได้ว่าเมื่อไหร่ควรให้ agent ใช้ tool และเมื่อไหร่ควรให้ตอบจากความรู้ภายในแทน
02
แหล่งข่าว
03
ประเด็น
วันนี้ · 11:08
อัปเดต
- LLM agents มีแนวโน้มใช้ tool เกินจำเป็น โดยเฉพาะเมื่อมีความรู้ภายในเพียงพอ (tool overuse illusion)
- ทุก tool call มีต้นทุนแฝงด้าน token และ latency ซึ่งอาจไม่คุ้มค่าในบางกรณี (tool-use tax)
- งานวิจัยเสนอแนวทางควบคุม tool use เช่น ToolGate, Contract2Tool, และ activation steering เพื่อเพิ่มประสิทธิภาพ
แหล่งต้นทาง · 12
ลิงก์ต้นทางอยู่ครบ เพื่อให้เปิดอ่านเต็มและเทียบข้อมูลเองได้
ENENENENENENENENENENENEN
TechCrunch — AIเมื่อวาน · 00:19
Are brain waves the next unlock for physical AI?
arXiv — cs.AI6 วันก่อน
AI Tool Discovery at Scale: All You Need is DNS
arXiv — cs.AI6 วันก่อน
Integro-differential equations in angular stabilization of drone motion by distributed feedback control
arXiv — cs.AI20 ก.ค.
ToolVerse: Unlocking Massive Environments and Long-Horizon Tasks for Agentic Reinforcement Learning
arXiv — cs.AI9 ก.ค.
Does AI Understand Imaging? A Systematic Benchmark of Agentic AI for Computational Imaging Tasks
arXiv — cs.AI8 ก.ค.
Controlling Tool Use with Heading-Specific Activation Steering
arXiv — cs.AI9 มิ.ย.
Contract2Tool: Learning Preconditions and Effects for Reliable Tool-Augmented LLM Agents
arXiv — cs.AI8 มิ.ย.
A Geometric Account of Activation Steering through Angle-Norm Decomposition
arXiv — cs.AI3 มิ.ย.
ToolGate: Token-Efficient Pre-Call Control for Tool-Augmented Vision-Language Agents
arXiv — cs.AI16 พ.ค.
Model-Adaptive Tool Necessity Reveals the Knowing-Doing Gap in LLM Tool Use
arXiv — cs.AI4 พ.ค.
Are Tools All We Need? Unveiling the Tool-Use Tax in LLM Agents
arXiv — cs.AI24 เม.ย.
The Tool-Overuse Illusion: Why Does LLM Prefer External Tools over Internal Knowledge?
แชร์
ข่าวที่เกี่ยวข้อง
ศาลอนุมัติ Anthropic จ่าย $1.5 พันล้านชดเชยละเมิดลิขสิทธิ์หนังสือ 5 แสนเล่ม
3 แหล่ง · 41 นาทีที่แล้ว
OpenAI models หลุด sandbox เจาะ Hugging Face ขโมยข้อมูล benchmark
3 แหล่ง · 42 นาทีที่แล้ว
SpaceX ซื้อ Cursor มูลค่า 6 หมื่นล้านดอลลาร์ หลัง IPO ไม่กี่วัน
3 แหล่ง · 42 นาทีที่แล้ว
AI เร่งช่องว่างระหว่าง Build กับ GTM 3 ประเด็นที่ PM และ Founder ต้องปรับ
1 แหล่ง · วันนี้ · 23:08
งานวิจัยล่าสุดชี้ AI agents กำลังก้าวสู่การปรับปรุงตัวเองแบบอัตโนมัติ
2 แหล่ง · วันนี้ · 23:07