Related articles
Edit |
Discuss Article
Search engine search engine being used to find Wikipedia]]
A search engine is a program designed to help one access files stored on a computer, for example a public server on the World Wide Web. The search engine allows one to ask for media content meeting specific criteria (typically those containing a given word or phrase) and retrieving a list of files that match those criteria. Unlike an index document that organizes files in a predetermined way, a search engine looks for files only after the user has entered search criteria.
In the context of the Internet, search engines usually refer to the World Wide Web and not other protocols or areas. Furthermore search engines mine data available in newsgroups, large databases, or open directories like DMOZ.org. Because the data collection is automated, they are distinguished from Web directories, which are maintained by people.
The vast majority of search engine are run by private companies using proprietary algorithms and closed databases, the most popular currently being Google (with MSN Search and Yahoo closely behind). There have been several attempts to create open-source search engines, among which are Htdig, Nutch, Egothor and OpenFTS. [1]
How search engines work
Web search engines work by storing information about a large number of web pages, which they retrieve from the WWW itself. These pages are retrieved by a web crawler -- an automated web browser which follows every link it sees. The contents of each page are then analyzed to determine how it should be indexed (for example, words are extracted from the titles, headings, or special fields called meta tags). Data about web pages is stored in an index database for use in later queries. Some search engines, such as Google, store all or part of the source page (referred to as a cache) as well as information about the web pages.
When a user comes to the search engine and makes a query, typically by giving key words, the engine looks up the index and provides a listing of best-matching web pages according to its criteria, usually with a short summary containing the document's title and sometimes parts of the text.
The usefulness of a search engine depends on the relevance of the results it gives back. While there may be millions of Web pages that include a particular word or phrase, some pages may be more relevant, popular, or authoritative than others. Most search engines employ methods to rank the results to provide the "best" results first. How a search engine decides which pages are the best matches, and what order the results should be shown in, varies widely from one engine to another. The methods also change over time as Internet usage changes and new techniques evolve.
Most Web search engines are commercial ventures supported by advertising revenue and, as a result, some employ the controversial practice of allowing advertisers to pay money to have their listings ranked higher in search results.
History
The first Web search engine was Lycos, which started at Carnegie Mellon University as a research project in 1994.
Soon after, many search engines appeared and vied for popularity. These included WebCrawler, Hotbot, Excite, Infoseek, Inktomi, and AltaVista. In some ways they competed with popular directories such as Yahoo. Later, the directories integrated or added on search engine technology for greater functionality.
In 2002, Yahoo! acquired Inktomi and in 2003, Yahoo! aquired Overture, which
owned AlltheWeb and Altavista. In 2004, Yahoo! launched it's own search engine based on the combined technologies of its acquisitions and providing a service that gave pre-eminence to the Web search engine over the directory.
Search engines were also known as some of the brightest stars in the Internet investing frenzy that occurred in the late 1990s. Several companies entered the market spectacularly, recording record gains during their initial public offerings.
Before the advent of the Web, there were search engines for other protocols or uses, such as the Archie search engine for anonymous FTP sites and the Veronica search engine for the Gopher protocol.
Osmar R. Zaïane's From Resource Discovery to Knowledge Discovery on the Internet details the history of search engine technology prior to the emergence of Google.
Recent additions to the list of search engines include a9.com, AlltheWeb, Ask Jeeves, Gigablast, Teoma, Wisenut, Kartoo, and Vivisimo.
Google
In around 2001, the Google search engine rose to prominence. Its success was based in part on the concept of link popularity and PageRank. Each page is ranked by how many pages link to it, on the premise that good or desirable pages are linked to more than others. The PageRank of linking pages and the number of links on these pages contribute to the PageRank of the linked page. This makes it possible for Google to order its results by how many web sites link to each found page.
A factor in Google's success, which spawned many offsprings, is the simplicity of its user interface.
Researchers at NEC Research Institute claim to have improved upon Google's patented PageRank technology by using web crawlers to find "communities" of websites. Instead of ranking pages, this technology uses an algorithm that follows links on a webpage to find other pages that link back to the first one and so on from page to page. The algorithm "remembers" where it has been and indexes the number of cross-links and relates these into groupings. In this way virtual communities of webpages are found.
Challenges faced by search engines
- The web is growing much faster than any present-technology search engine can possibly index (see distributed crawling).
- Many web pages are updated frequently, which forces the search engine to revisit them periodically.
- The queries one can make are currently limited to searching for key words, which may results in many false positives.
- Dynamically generated sites, which may be slow or difficult to index, or may result in excessive results from a single site.
- Many dynamically generated sites are not indexable by search engines; this phenomenon is known as the invisible web.
- Some search engines do not order the results by relevance, but rather according to how much money the sites have paid them.
- Some sites use tricks to manipulate the search engine to display them as the first result returned for some keywords. This can lead to some search results being polluted, with more relevant links being pushed down in the result list.
See also
External links
Source | Copyright
|
 |
 |
 |
Webmasters: Add your website here:
Readers: Edit |
Discuss Listings
Internet Promotion Outlet Marketing and promotional software to help your website promotion efforts. http://www.shoplet.com/pr/index.html
Ecisive A full-service interactive advertising agency specializing in the design, development and marketing of interactive media; including websites, CD-ROMs, banner programs and opt-in email. http://www.pushinteractive.com
1 2 3 Link Free and low cost web site promotion services. Free web page; submit url to search engines; free banner exchange to increase traffic; reciprocal link exchange; add url to business directory; internet advertising and marketing. http://www.123link.com
Reciprocom Advertising package available in a high quality print format or as a state-of-the-art multimedia CD-ROM to reach targeted e-commerce consumers, through pooled consumer lists. http://www.reciprocom.com
NewGate Internet Provides corporate consulting for online public relations and internet marketing strategy. Audience development programs build online brand awareness and increase site traffic. http://www.newgate.net
Majon International Offers targeted web marketing promotions and advertising including mall linkings, opt-in email campaigns, web site exposure promotions, pop-up and pop-under advertising, and press release distribution services. http://www.majon.com
Sweepstakes Builder Traffic builder. http://www.sweepstakesbuilder.com
Ad Resource Features information on internet advertising and web site promotion. http://adres.internet.com/
Mortgage Promote Internet search engine management provider to the mortgage industry offering web site promotion services, Internet marketing techniques and search engine strategies to mortgage brokers, bankers, and lending institutions. http://www.mortgagepromote.com
Webmasters Heaven Everything you need to create, maintain and promote a website. http://home.swipnet.se/~w-59072/webmaster/
E@symail Interactive Providing opt-in email list brokering and email campaign management. http://www.easymailinteractive.com
TrafficJumper.com Provides web site promotion suite for a monthly fee. Includes email marketing, banner impressions, and search engine submissions. http://www.trafficjumper.com
List You Ltd Offers pay per click management, market research, email marketing, online PR, directory placement and affiliate marketing. http://www.listyou.co.uk
Pandia Per and Susanne Koch's comprehensive guide to Web searching and search engine optimization, with tutorials, news, tools, reviews and SE marketing resources. http://www.pandia.com/
Academy of Web Specialists Search engine marketing resources. http://www.academywebspecialists.com
Web Marketing and Site Promotion Internet marketing and website promotion tutorials, including search engine optimisation techniques. http://www.webmarketingplus.co.uk/
The EngineMage Site offers a list of search engines, articles relating to search engine optimization, key word suggestion tools, and contact details. http://enginemage.com/
Hits To Sales - Special Report Exploding search engine myths. http://www.hitstosales.com/2search.html
Search Engine Ethics Offers search engine optimization and submission articles and news. Describes how search engines work, ranking algorithms, and submission processes. http://www.searchengineethics.com/
Search Engine World Brett Tabke's SEO topics for mid-level to pro-level search engine promotion specialists. News updated daily. http://www.searchengineworld.com
Submit Corner Reviews and comparisons of the top search engines with free online tools to help optimize your website. http://www.submitcorner.com/Guide/SE/
LinkScout Customizable website promotion system. http://www.linkscout.com/
Eric Ward Specialized internet promotion and press release services for high-end web site launches. http://www.ericward.com/
Search Engine Marketing FAQ Answers questions about search engines, directories, spiders, submissions, protection of intellectual property, and effective search engine optimization and placement. http://www.seologic.com/faq/
Web Ranking Reports Offers real time, web based search engine ranking reports. http://www.webrankingreports.com/
|