跳到正文
AWS Machine Learning Blog·· 3 天前

Fine-tune a search agent with multi-turn RL on Amazon SageMaker AI

Fine-tune a search agent with multi-turn RL on Amazon SageMaker AI

摘要

Fine-tuning teaches a small search agent your tools and environment, giving it the reliability of a frontier model at lower latency and cost. In this post, we fine-tune an LLM-powered search agent with multi-turn reinforcement learning (MTRL) on Amazon SageMaker AI and share the gains we measured in retrieval quality and reliability.

应来源方要求,这里只提供摘要与原文入口。完整内容请阅读原文。

来源:AWS Machine Learning Blog · aws.amazon.com