I spend a good amount of time chasing the LLM for evidence that its answer is actually accurate. To be honest, I can usually get there, but it works great for fresh work and gets messy on months-old work. And it&#x27;s not just me; a lot of people have raised the same thing.<p>So for the last 6 months I&#x27;ve been trying to fix it. Today I can say it&#x27;s working for me — it flags me in advance. Obviously not perfect yet, and that&#x27;s where I need help: I need people to test it and tell me what else needs fixing.<p>PS: This isn&#x27;t a memory layer; it flags stale answers before you act on them. Give it a try, and if you think I&#x27;m solving a real problem, please drop a star.


  • 情报分类:技术价值
  • 命中依据:LLM输出准确性问题有技术讨论价值
  • 来源:Hacker News 新项目
  • 原作者:bhanuhai2
  • 发布时间:2026/9/9 18:22:22