AIDB Daily Papers
ハーネスの最適化でコストを90%削減:小規模言語モデルによる効率的なAIエージェント構築
※ 日本語タイトル・ポイントはAIによる自動生成です。正確な内容は原論文をご確認ください。
ポイント
- 小規模言語モデルを特定のタスク向けに自動調整されたハーネスと組み合わせる手法を提案した。
- タスクの難易度をモデルからハーネスへ移すことで、低コストで高性能なエージェントを実現した点が新しい。
- 最適化により、LLMの約90%の性能を維持しつつ、コストを最大96%削減することに成功した。
Abstract
Frontier LLM agents are automating many business tasks, but their high inference cost makes large-scale deployment unsustainable. Small language models (SLMs) offer a cheaper alternative, yet they typically fall short when swapped into a harness designed for a frontier LLM. We show that for many routine business tasks, SLM agents can match LLM performance at 90% lower cost, when paired with an adapted harness that can be automatically discovered by a meta agent. The key insight is that much of the task difficulty is shared across instances and can be lifted from the model into the harness via tailored instructions, tools, and orchestration loops. To study this systematically, we create a framework that maps agent failure modes to harness adaptation strategies, and build a harness optimizer that automatically discovers effective adaptations from failure trajectories. Across seven business-oriented agentic tasks and three SLM families, we found optimized harnesses significantly improve performance on 16 of 21 task-SLM pairs, with seven pairs closing the SLM-LLM performance gap and the best SLM agent recovering 89.7% of LLM performance at 4% of the cost. Our analysis further shows that adaptation works best for tasks with more repetitive workflows and for SLMs with sufficient base capabilities. Together, these results suggest that harness adaptation can expand the practical deployment range of SLM agents in routine business tasks.
Paper AI Chat
この論文のPDF全文を対象にAIに質問できます。
質問の例: