Original Summary

a couple weeks ago I was coding with claude code and other ai ides and got completely fed up. the agent kept saying "task complete" with total confidence and then id check and backend routes were missing, buttons werent wired to anything, or it was playing hero fixing a bug it created 20 minutes earlier in the same session. posted a raw rant about it on reddit expecting like two upvotes. it blew up to 21k views instead. read every comment. turns out it wasnt just me, people from weekend vibe coders to actual senior engineers were burning hours manually diffing code because the agent was grading its own homework. instead of trying to reach more people I just had real 1 on 1 conversations with whoever replied, how they actually verify an agent's work today, what they'd tried building themselves. those conversations shaped the whole thing. built malveon check, a small local cli that reads your plan, checks the real code against it (wiring, route overlaps, that kind of stuff), and actually runs your build/lint/test commands instead of guessing. if it cant prove something it says so instead of pretending. just shipped v0.1.0 for free beta testing. binaries only right now, source isnt public yet, wanted to say that plainly rather than have someone assume otherwise. happy to answer anything about how it works or send the install steps if anyone wants to try it against a real repo.   submitted by   /u/AlternativeLimit8551 [link]   [comments]


  • 情报分类:商业与市场研究
  • 分类依据:内容涉及商业、投资或市场动态
  • 信息来源:Reddit · SideProject
  • 发布时间:2026/9/18 02:10:49