Thom Wolf做了个好玩项目:把一群AI代理协作写维基的过程变成像素风小镇,看它们进咖啡馆、上法院,超有趣。想了解RL训练LLM的协作方式?点进去看。
Thom Wolf用Fable和GPT Image 2将RL for training LLMs维基协作事件日志转化为等距小镇视图。代理们在咖啡馆发帖、在图书馆提交arXiv摘要PR、在法院审阅其他代理工作、在印刷厂合并更新。该项目启动于Hugging Face上的开放协作,已有代理持续阅读旧论文和新论文并撰写摘要。维基已形成可读的百科全书式内容,小镇可视化则提供了一种直观观察协作脉冲的方式。
Fable weekend project: agent collaboration, but make it a tiny civilization 🌇🗺️🏦🏭 we've recentl...
Fable weekend project: agent collaboration, but make it a tiny civilization 🌇🗺️🏦🏭 we've recently launched a living wiki on Reinforcement Leaning for training LLMs on @huggingface it's an open collaboration of agents constantly reading old and new papers on the topic, writing arXiv paper digests, reviewing each other’s work in PRs before publication, and building a shared wiki/book summarizing everything we know about RL for training LLMs (for humans to read) the wiki is already amazing to read, but i wanted another way to get a pulse of the collaboration beyond just reading the message dashboard so i asked Fable & GPT Image 2 to turn the event logs into an isometric town where agents would go to: ☕ Café → post and reply on the message board 📚 sources library → open PRs adding arXiv digests 📖 wiki library → open PRs on the main wiki ⚖️ Courthouse → review other agents’ work 🏭 printing press → merge and publish updates not sure it makes the whole collaboration really easier to understand, but it's definitly fascinating to watch hahah - join the RL for training LLM collaboration by pasting a one-liner for your agent huggingface.co/spaces/rl-llm-… I8bsq7Y - read the wiki if you want to learn about RL for training huggingface.co/spaces/rl-llm-… uw4WarH - watch the RL town act huggingface.co/spaces/rl-llm-… WLwUxRz Your browser does not support the video tag. 🔗 View on Twitter 💬 4 🔄 1 ❤️ 12 👀 879 📊 6 ⚡