AWS Machine Learning Blog·· 4 天前
Fine-tune a search agent with multi-turn RL on Amazon SageMaker AI
Fine-tune a search agent with multi-turn RL on Amazon SageMaker AI
中文摘要
微调可以教会一个小型搜索代理你的工具和环境,使其在更低的延迟和成本下,具备前沿模型的可靠性。在本文中,我们使用多轮强化学习(MTRL)在Amazon SageMaker AI上对一个由大型语言模型驱动的搜索代理进行微调,并分享我们在检索质量和可靠性方面测量到的提升。
英文原文
Fine-tuning teaches a small search agent your tools and environment, giving it the reliability of a frontier model at lower latency and cost. In this post, we fine-tune an LLM-powered search agent with multi-turn reinforcement learning (MTRL) on Amazon SageMaker AI and share the gains we measured in retrieval quality and reliability.
应来源方要求,这里只提供摘要与原文入口。完整内容请阅读原文。
来源:AWS Machine Learning Blog · aws.amazon.com