Original Summary

How are people deciding whether a cheaper LLM is good enough for a feature? because saving on tokens makes total sense but less so if it takes three attempts to get a usable answer. Are you testing it against your current model on real requests or just checking a few outputs before switching?   submitted by   /u/Particular_Car5864 [link]   [comments]


  • 情报分类:商业与市场研究
  • 分类依据:内容涉及商业、投资或市场动态
  • 信息来源:Reddit · SaaS
  • 发布时间:2026/10/11 22:40:55