LLM is trained by back propagation, how about harness and mu

回复
JianguoChuan楼主
等级10:见习点评
帖子互动: 89
帖子: 2172
注册时间: 2024年 11月 19日 17:20

#1 LLM is trained by back propagation, how about harness and mu

帖子 JianguoChuan楼主 »

multi-agents?

It has to be trained by LLM itself, otherwise, there is no way to do it by human, right?

Algorithm experts, what is your opinion?


+2.00 积分 [版主 wh 发放的奖励]
头像
MaLaRabbit
等级11:论坛点评
2025年度优秀版主
帖子互动: 198
帖子: 3244
注册时间: 2022年 7月 24日 02:16

#2 Re: LLM is trained by back propagation, how about harness and mu

帖子 MaLaRabbit »

其实早就有人用RL或遗传算法来优化多智能体了,根本不用人类手动调。再说反向传播也不是唯一途径,算法圈早就开始搞无梯度优化了

JianguoChuan 写了: 2026年 8月 29日 15:29

multi-agents?

It has to be trained by LLM itself, otherwise, there is no way to do it by human, right?

Algorithm experts, what is your opinion?

☆ 发自新买提 Android 26.07.21


+2.00 积分 [版主 wh 发放的奖励]
wass
等级13:论坛精英
2024年度优秀版主

wass 的博客
帖子互动: 870
帖子: 8624
注册时间: 2022年 7月 23日 22:13

#3 Re: LLM is trained by back propagation, how about harness and mu

帖子 wass »

JianguoChuan 写了: 2026年 8月 29日 15:29

multi-agents?

It has to be trained by LLM itself, otherwise, there is no way to do it by human, right?

Algorithm experts, what is your opinion?

Harness有rsi的论文,agents有tool,也可以手工调,没有看到自动的


+2.00 积分 [版主 wh 发放的奖励]
JianguoChuan楼主
等级10:见习点评
帖子互动: 89
帖子: 2172
注册时间: 2024年 11月 19日 17:20

#4 Re: LLM is trained by back propagation, how about harness and mu

帖子 JianguoChuan楼主 »

wass 写了: 2026年 8月 30日 19:06

Harness有rsi的论文,agents有tool,也可以手工调,没有看到自动的

嗯,目前还是手工为主,最新有不少文章做RSI的。

回复

回到 “葵花宝典(Programming)”