🚀NEW LABGetting Started with Claude AgentsStart lab
Reinforcement Learning

Model Swarms

First page
Model Swarms
Paper summary

propose a new collaborative search algorithm to adapt LLM via swarm intelligence; a pool of LLM experts collaboratively move in the weight space and optimize a utility function representing various adaptation objectives; experiments demonstrate that Model Swarms could flexibly adapt LLM experts to a single task, multi-task domains, reward models, as well as diverse human interests. improves over 12 model composition baselines by up to 21.0% across tasks and contexts.

Ask this paper

Every Monday
Get next week’s papers.
Subscribe on Substack