BETA SAÚDE
Distributed search engine
Texto da Wikipédia (en), licença CC BY-SA. O BETARUBI mostra o verbete inteiro nesta página — a leitura não continua fora do site.
A distributed search engine is a search engine where there is no central server. Unlike traditional centralized search engines, work such as crawling, data mining, indexing, and query processing is distributed among several peers in a decentralized manner where there is no single point of control.
History
This section may require cleanup to meet Wikipedia's quality standards. The specific problem is: Sub-section order is not chronological, but potentially promotional. (February 2025) |
Presearch
Started in 2017, Presearch is an ERC20 powered (PRE) search engine powered by a distributed network of community operated nodes which aggregate results from a variety of sources. This powers the searches at presearch.com. This is planned to be a precursor where each node collaborates on a global decentralised index. [1] Presearch averages 5 million searches per day and has 2.2 million registered users. On Sept 1, 2021, Presearch was added as a default option to the search engine list on Android for the EU.[2] On May 27, 2022, Presearch officially transitioned from its Testnet to a Mainnet. This means all search traffic through the service now runs over Presearch's decentralized network of volunteer-run nodes.[3]
YaCy
On December 15, 2003, Michael Christen announced development of a P2P-based search engine, eventually named YaCy, on the heise online forums.[4][5]
Seeks
Seeks was an open source websearch proxy and collaborative distributed tool for websearch. It ceased to have a usable release in 2016.
InfraSearch
In April 2000 several programmers (including Gene Kan, Steve Waterhouse) built a prototype P2P web search engine based on Gnutella called InfraSearch. The technology was later acquired by Sun Microsystems and incorporated into the JXTA project.[6] It was meant to run inside the participating websites' databases creating a P2P network that could be accessed through the InfraSearch website.[7][8][9]
Opencola
On May 31, 2000 Steelbridge Inc. announced development of OpenCOLA a collaborative distributive open source search engine.[10] It runs on the user's computer and crawls the web pages and links the user puts in their opencola folder and shares resulting index over its P2P network.[11]
Faroo
In February 2001 Wolf Garbe published an idea of a peer-to-peer search engine,[12] started the Faroo prototype in 2004,[13] and released it in 2005.[14][15]
Goals
The goals of building a distributed search engine include:
- to create an independent search engine powered by the community;
- to make the search operation open and transparent by relying on open-source software;
- to distribute the advertising revenue to node maintainers, which may help create more robust web infrastructure;
- to allow researchers to contribute to the development of open-source and publicly-maintainable ranking algorithms and to oversee the training of the algorithm parameters.
Challenges
- The amount of data to be processed is enormous. The size of the visible web is estimated at 5PB spread around 10 billion pages.
- The latency of the distributed operation must be competitive with the latency of the commercial search engines.
- A mechanism that prevents malicious users from corrupting the distributed data structures or the rank needs to be developed.
See also
References
- ↑ "Presearch is a Decentralized Search Engine".
- ↑ 297shares; 4.3kreads (2021-09-01). "Google Adds Presearch As A Default Option on Android Devices in EU". Search Engine Journal. Retrieved 2021-11-10.
{{cite web}}: CS1 maint: numeric names: authors list (link) - ↑ Kan, Michael (2022-05-26). "The Next Google? Decentralized Search Engine 'Presearch' Exits Testing Phase". PC Magazine.
- ↑ "YaCy: News". Archived from the original on 2005-11-24.
- ↑ Michael Christen. "Ich entwickle eine P2P-basierende Suchmaschine. Wer macht mit?". heise online.
- ↑ Justin Hibbard. "Can peer-to-peer grow up?". Red Herring.[permanent dead link]
- ↑ Simon Foust. "Move Over Yahoo, Here Comes InfraSearch". Dmusic. Archived from the original on 2000-10-13.
- ↑ Sean M. Dugan. "Peer-to-peer networking is poised to revolutionize the Internet once again". InfoWorld. Archived from the original on 2000-10-18.
- ↑ John Borland. "Napster-like technology takes Web search to new level". Cnet.
- ↑ David Akin. "Software launched with a little pop". Financial Post.[dead link]
- ↑ Paul Heltzel. "OpenCola-Have Some Code and a Smile". Technology Review.
- ↑
Wolf Garbe. "BINGOOO - Die Transformation des World Wide Web zur virtuellen Datenbank" (in German). Wirtschaftinformatik. Archived from the original on 2014-02-02. Retrieved 2010-12-21.
... Wir setzen dem das Konzept einer verteilten Peer-to-Peer-Suchmaschine entgegen [We counter with the concept of a distributed peer-to-peer search engine] ...
- ↑
Bernard Lunn. "Technical Q&A With FAROO Founder". ReadWriteWeb. Archived from the original on 2011-02-14.
... When I started to work on the first prototype in 2004 ...
- ↑ "FAROO: History". Archived from the original on 2008-03-22.
- ↑ "Revisited: Deriving crawler start points from visited pages by monitoring HTTP traffic". Faroo.
