What did you understand the most/least about today's lesson/independent practice?
10/25/10I understood most of today's lesson about why certain content is only available on the invisible web. One reason is that the search engine may not know about the page/website. The second reason is that the content is located deep in the web, so the search engine would not bother to look for it. The third reason is that the page changes so much, that the search engine doesn't bother to look for it all the time because its meaningless. The fourth reason is because the search engine is not able to search non-HTML content such as images or other file types like PDF. The fourth reason why certain content can only be search on the invisible web is because the search engine cannot access the pages to index them because the page requires a password or the site has a search box that is required to be filled out.
10/26/10
Today we learned 5 reason why search engines do not index deep web content. The first reason is because the search engine may not be able to access the page because it requires a password. If there's a site that we know about, we are able to search its content without a using a search engine. Another reason why search engines don't look for deep content is because the content is located in the last part of the last category that holds the most interest for searchers which are the sites that hold information in databases such as phone directories, literature databases and patent databases.
No comments:
Post a Comment