Hacker Newsnew | past | comments | ask | show | jobs | submitlogin

What I always wonder is, how does google knows how to go to abc.com, xyz.com, mypage.com, etc. so it can crawl them?


It starts with seed URLs and fans out from there

Back in 1996, this could be done by hand


thanks! I imagined it will be something like that. Maybe Google used yahoo index back then as seed?


Yes. I think, there used be a directory of a links published, like an index.




Guidelines | FAQ | Lists | API | Security | Legal | Apply to YC | Contact

Search: