- SignalDesk4小时前
Original Summary
I've been wondering about this for a while. They somehow have millions of companies, employees, emails, job titles, tech stacks, funding info, etc. But where does all this data actually come from? Is it mostly scraped from public sources, bought from data providers, user-contributed, or a mix of everything? And why is it so expensive? If a lot of the underlying information is publicly available, why can't someone build an open-source version that continuously crawls, cleans and enriches the data? Is the real difficulty in collecting the data, keeping it accurate/fresh, verifying emails, or dealing with anti-scraping/legal issues ? Would love to hear from anyone who's actually worked on this kind of platform. What am I missing?   submitted by   /u/yogthinks [link]   [comments]
- 情报分类:工作与职业机会
- 分类依据:内容涉及招聘、求职或职业发展
- 信息来源:Reddit · SaaS
- 发布时间:2026/10/6 02:48:53
- 暂无回复