- SignalDesk4小时前
Original Summary
Everyone building an agent right now is one bad prompt injection away from it doing something it shouldn't. The usual fix is a second call to GPT-4o or Claude to ask "was that safe", which doubles your cost and latency for what's really a yes/no classification. I built a free API that does that check with Jev instead, the new TypeSafe classification-only model everyone's been talking about this week. No text generation, just probabilities against fixed criteria, so it runs stupidly fast, under 300ms most of the time, for a fraction of a cent per call. Full disclosure, the landing page HTML is vibecoded, I threw it together fast. The actual work went into the API integration, the quota system, and the Stripe billing behind it. Free tier is 200 checks/month, no card needed. Rate limited so don't worry about hammering the demo. Curious what breaks it, throw your worst jailbreak attempts at it. https://oraclemarin.fr/agent-guard   submitted by   /u/Zboubkiller [link]   [comments]
- 情报分类:商业与市场研究
- 分类依据:内容涉及商业、投资或市场动态
- 信息来源:Reddit · SaaS
- 发布时间:2026/9/19 07:40:47
- 暂无回复