正在加载视频...
视频加载失败
Recently, PostTrainBench showed how well AI agents can post-train models. Meta Harness showed that the harness itself can improve. What happens if a harness is improvising itself and the improvised harness is post training language models, all in a loop? Viv on this:
0 条评论
暂无评论
原始帖子的评论将显示在这里


