- SignalDesk4小时前
Original Summary
Was curious how LLMs would play chicken without a clear reward matrix, and then took it a step further by allowing communication and realtime decision making.<p>Each LLM is in an independent loop, where each turn they are given the current speed and time to impact, as well as their historical latency. They can pre-plan moves to account for latency.
- 情报分类:技术学习与提效
- 分类依据:内容涉及技术、AI、软件工具或工程实践
- 信息来源:Hacker News 新项目
- 发布时间:2026/9/29 07:10:54
- 暂无回复