Forum for Science, Industry and Business

Sponsored by:     3M 
Search our Site:

 

Search engine makes social calls

07.03.2002


New algorithm exploits community structure of the web.



The web has spontaneously organized itself into communities. A new search algorithm that pinpoints these could help surfers find what they want and avoid offensive content.

Page builders can link anywhere. But they don’t, Gary Flake, of the NEC Research Institute in Princeton, and his colleagues have found. Instead, pages congregate into social groups that focus most of their attention on each other.


Web directories compiled by hand, such as Yahoo!, recognize this to an extent. Flake’s team has automated the process. "We find extremely high-quality sites that Yahoo! and Google don’t know about," he says.

The new search ignores a page’s text, looking only at its links. It crawls from a starting page to others it links to, and so on out into the web, picking out islands of expertise in the sea of information.

A test of the algorithm starting at the home pages of biologist Francis Crick, astrophysicist Stephen Hawking, and computer scientist Ronald Rivest yielded groups of sites that are tightly focused on each researcher’s life, work and field. "The sites are remarkably topically related - the clusters’ properties are completely intuitive," says Flake.

"You can extract a lot of meaning from links," agrees Mike Thelwall, who studies search engines at the University of Wolverhampton, UK. The new approach is, he says, a clever way to find meaningful groups among the effectively infinite number of ways to subdivide linked pages.

The first application of community searching may be to fence off areas of the web such as pornography or hate-speech communities, says Flake. Current content filters are largely text-based; these are easy to dodge and require intensive human management.

Community service

Google pioneered the use of links to deduce pages’ relevance. Its PageRank technology counts a link from site A to site B as a vote for B from A. But it does not take account of all the other sites to which A has links, as NEC’s new technique does.

Flake does not expect to displace the market leader - "Google’s a great search engine," he says. Rather, he wants to add an extra dimension to searching.

Using link structures could lead to more efficient, customized searching, particularly for scientists, who are careful to link to each other’s pages. "For academics it’s going to be a big improvement," says Thelwall.

Flake’s team is now mapping out communities in the web as a whole, without using a starter page. This could detect hitherto unsuspected communities, says computer scientist and network researcher Jon Kleinberg of Cornell University in Ithaca, New York.

"It could bring together people with common interests that may not know of each other’s existence. You could also catch the early stages of new trends," says Kleinberg.

References
  1. Flake, G. W., Lawrence, S., Giles, C. L. & Coetzee, F. Self-organization and identification of communities. IEEE Computer, 35, 66 - 71, (2002).

JOHN WHITFIELD | © Nature News Service

More articles from Information Technology:

nachricht Cutting edge research for the industries of tomorrow – DFKI and NICT expand cooperation
21.03.2017 | Deutsches Forschungszentrum für Künstliche Intelligenz GmbH, DFKI

nachricht Molecular motor-powered biocomputers
20.03.2017 | Technische Universität Dresden

All articles from Information Technology >>>

The most recent press releases about innovation >>>

Die letzten 5 Focus-News des innovations-reports im Überblick:

Im Focus: Giant Magnetic Fields in the Universe

Astronomers from Bonn and Tautenburg in Thuringia (Germany) used the 100-m radio telescope at Effelsberg to observe several galaxy clusters. At the edges of these large accumulations of dark matter, stellar systems (galaxies), hot gas, and charged particles, they found magnetic fields that are exceptionally ordered over distances of many million light years. This makes them the most extended magnetic fields in the universe known so far.

The results will be published on March 22 in the journal „Astronomy & Astrophysics“.

Galaxy clusters are the largest gravitationally bound structures in the universe. With a typical extent of about 10 million light years, i.e. 100 times the...

Im Focus: Tracing down linear ubiquitination

Researchers at the Goethe University Frankfurt, together with partners from the University of Tübingen in Germany and Queen Mary University as well as Francis Crick Institute from London (UK) have developed a novel technology to decipher the secret ubiquitin code.

Ubiquitin is a small protein that can be linked to other cellular proteins, thereby controlling and modulating their functions. The attachment occurs in many...

Im Focus: Perovskite edges can be tuned for optoelectronic performance

Layered 2D material improves efficiency for solar cells and LEDs

In the eternal search for next generation high-efficiency solar cells and LEDs, scientists at Los Alamos National Laboratory and their partners are creating...

Im Focus: Polymer-coated silicon nanosheets as alternative to graphene: A perfect team for nanoelectronics

Silicon nanosheets are thin, two-dimensional layers with exceptional optoelectronic properties very similar to those of graphene. Albeit, the nanosheets are less stable. Now researchers at the Technical University of Munich (TUM) have, for the first time ever, produced a composite material combining silicon nanosheets and a polymer that is both UV-resistant and easy to process. This brings the scientists a significant step closer to industrial applications like flexible displays and photosensors.

Silicon nanosheets are thin, two-dimensional layers with exceptional optoelectronic properties very similar to those of graphene. Albeit, the nanosheets are...

Im Focus: Researchers Imitate Molecular Crowding in Cells

Enzymes behave differently in a test tube compared with the molecular scrum of a living cell. Chemists from the University of Basel have now been able to simulate these confined natural conditions in artificial vesicles for the first time. As reported in the academic journal Small, the results are offering better insight into the development of nanoreactors and artificial organelles.

Enzymes behave differently in a test tube compared with the molecular scrum of a living cell. Chemists from the University of Basel have now been able to...

All Focus news of the innovation-report >>>

Anzeige

Anzeige

Event News

International Land Use Symposium ILUS 2017: Call for Abstracts and Registration open

20.03.2017 | Event News

CONNECT 2017: International congress on connective tissue

14.03.2017 | Event News

ICTM Conference: Turbine Construction between Big Data and Additive Manufacturing

07.03.2017 | Event News

 
Latest News

Argon is not the 'dope' for metallic hydrogen

24.03.2017 | Materials Sciences

Astronomers find unexpected, dust-obscured star formation in distant galaxy

24.03.2017 | Physics and Astronomy

Gravitational wave kicks monster black hole out of galactic core

24.03.2017 | Physics and Astronomy

VideoLinks
B2B-VideoLinks
More VideoLinks >>>