Spiders?

Status
Not open for further replies.

The Arcane

Scooby
Joined
Jan 2, 2005
Messages
1,686
This has been bugging me for a while so can somebody help me out? What are these Google and Yahoo! Spider things that you see listed when you look to see who is online? Is it just when somebody has viewed a page here from a search engine?
 
I believe they are actually bots from search engines. Not actual people. So when someone searches for something (say Buffy) and you get your list of results where that word appears on a particular Web page, well we here on our site will see it as a spider or bot.
 
Someone tried to explain this to me on another forum. Just confused me more 🙁
 
T
The Arcane
When it comes to all things websitey, you and be both, sweetheart ;)
This might help.

The Parts Of A Crawler-Based Search Engine

Crawler-based search engines have three major elements. First is the spider, also called the crawler. The spider visits a web page, reads it, and then follows links to other pages within the site. This is what it means when someone refers to a site being "spidered" or "crawled." The spider returns to the site on a regular basis, such as every month or two, to look for changes.

Everything the spider finds goes into the second part of the search engine, the index. The index, sometimes called the catalog, is like a giant book containing a copy of every web page that the spider finds. If a web page changes, then this book is updated with new information.

Sometimes it can take a while for new pages or changes that the spider finds to be added to the index. Thus, a web page may have been "spidered" but not yet "indexed." Until it is indexed -- added to the index -- it is not available to those searching with the search engine.

Search engine software is the third part of a search engine. This is the program that sifts through the millions of pages recorded in the index to find matches to a search and rank them in order of what it believes is most relevant.

http://searchenginewatch.com/showPage.html?page=2168031

also

Off the page factors are those that a webmasters cannot easily influence. Chief among these is link analysis. By analyzing how pages link to each other, a search engine can both determine what a page is about and whether that page is deemed to be "important" and thus deserving of a ranking boost. In addition, sophisticated techniques are used to screen out attempts by webmasters to build "artificial" links designed to boost their rankings.

Another off the page factor is clickthrough measurement. In short, this means that a search engine may watch what results someone selects for a particular search, then eventually drop high-ranking pages that aren't attracting clicks, while promoting lower-ranking pages that do pull in visitors. As with link analysis, systems are used to compensate for artificial links generated by eager webmasters.

http://searchenginewatch.com/showPage.html?page=2167961

I think nowadays spiders also follow links from other sites.
 
Thanks for that, N4H. I may actually be a little more confused now but that is usually the way when I ask a question about internet stuff and somebody actually gives me an answer. 😀
 
K
killerdwarf
LOL! Me too...
Status
Not open for further replies.
Back
Top Bottom