Forum for Science, Industry and Business

Sponsored by:     3M 
Search our Site:

 

New method for building multilingual ontologies that can be applied to the Semantic Web

19.09.2008
Researchers from the Validation and Business Applications Group based at the Universidad Politécnica de Madrid’s School of Computing (FIUPM) have developed a new method for building multilingual ontologies that can be applied to the Semantic Web.

An ontology is a structured set of terms and concepts underpinning the meaning of a subject area. Artificial intelligence and knowledge representation systems are the principal users of ontologies. Researchers from all over the world are now working on applying ontologies to the Internet with the aim of building a Semantic Web and providing users with a intelligent tool for using the information on the web.

The importance of the proposal by Jesús Cardeñosa, Carolina Gallardo, Luis Iraola and Miguel Ángel de la Villa is that it revolutionizes current ontology building systems. Until now, these systems have had a major stumbling block: the multilingual component. The application of ontologies to the Internet comes up against serious problems triggered by linguistic breadth and diversity. This diversity stands in the way of users making intelligent use of the web.

So far all approaches to solve the multilingual component have run into serious difficulties. Some were based on experts agreeing on the terms to be used in each language. These endeavours failed to take off due to the difficulty of finding experts in many languages. Other approaches set out to use one language (almost always English) as a pivot. Again the results were held back, this time because the use of a natural language as an interlanguage occasions ambiguity. As the researchers note, early attempts at using a natural reference language to build machine translation systems go back 20 years ago, and the results were no good.

Multilingual thesauruses

Neither has the 1985 international standard on the creation and development of multilingual thesauruses (ISO 5964:1985) solved the problem. A thesaurus is a list of hierarchically interrelated (general and subordinate terms) terms, possibly containing more than one word, used to index (for archiving purposes) and retrieve documents.

Although an ontology is not exactly the same thing as a thesaurus, the researchers claim that it is very similar because a great many ontologies are actually a simplified version of its huge potential for representation, confined to three basic relations (“is-part-of”, “is-a-type-of”, “is-a”). The above ISO standard on thesauruses covers these three basic relationships.

According to these researchers, the current approaches for solving the multilingual component in ontology building are mostly applied to ontologies with the same representation power as a thesaurus, without even applying the above ISO 5964:1985 standard to deal with multilingualism. But, even if this standard is applied, the approach always has to use a reference language, which, as mentioned, has never worked.

Language-independent ontologies

The method proposed by the School of Computing researchers solves this problem because it is based on building ontologies that can represent information irrespective of the language. They are therefore applicable to multilingual systems.

This method is an advance in that, after analysing the natural language, the method looks for linguistic patterns (grammatical structures) that match the exact ontological structures. This it does multilingually, as the linguistic patterns are capable of building language-independent structures.

The innovative thing about what these researchers are proposing is the construction of multilingual ontologies using what are known as universal words as the concept name. The concept of universal word stems from the United Nations University’s UNL Project (Universal Networking Language). This project was set up to break down the linguistic barriers on the Internet. The researchers claim that the characteristics of UNL also very close match the features of an ontology.

These researchers assume that ordinary texts contain more information than what can be extracted using just the domain terms, as any text has implicit ontological relations that can be extracted by analysing certain grammatical structures of the sentences making up the text.

A case study

In an article presented last July at The 2008 International Conference on Semantic Web and Web Services (SWWS'08), these researchers describe their approach and explain a case study demonstrating the validity of their method. The case study used the contents of the current catalogue of Spanish monuments as part of the Patrilex project, funded by the Spanish National Research Plan in conjunction with the Underdirectorate General of Cultural Heritage.

In this case study, the sentences from the catalogue of Spanish monuments were coded in UNL language. This codification is a semantic representation of the catalogue contents. The researchers then searched for predefined linguistic patterns in the semantic representation. After identifying the contents matching the patterns, they instantiated the contents as ontological structures.

The big advantage of using the UNL system is that the universal words are independent of the language and are not ambiguous. The non-ambiguity makes the translation of the ontology built this way to any language extremely precise.

The universal words used can be viewed and used publicly at the public access repository developed by the researchers. Universal words are usually used to develop multilingual dictionaries, as explained in another press release.

Eduardo Martínez | alfa
Further information:
http://www.fi.upm.es/?pagina=737&idioma=english

More articles from Information Technology:

nachricht Smarter robot vacuum cleaners for automated office cleaning
15.08.2017 | Fraunhofer-Institut für Arbeitswirtschaft und Organisation IAO

nachricht Researchers 3-D print first truly microfluidic 'lab on a chipl devices
15.08.2017 | Brigham Young University

All articles from Information Technology >>>

The most recent press releases about innovation >>>

Die letzten 5 Focus-News des innovations-reports im Überblick:

Im Focus: Exotic quantum states made from light: Physicists create optical “wells” for a super-photon

Physicists at the University of Bonn have managed to create optical hollows and more complex patterns into which the light of a Bose-Einstein condensate flows. The creation of such highly low-loss structures for light is a prerequisite for complex light circuits, such as for quantum information processing for a new generation of computers. The researchers are now presenting their results in the journal Nature Photonics.

Light particles (photons) occur as tiny, indivisible portions. Many thousands of these light portions can be merged to form a single super-photon if they are...

Im Focus: Circular RNA linked to brain function

For the first time, scientists have shown that circular RNA is linked to brain function. When a RNA molecule called Cdr1as was deleted from the genome of mice, the animals had problems filtering out unnecessary information – like patients suffering from neuropsychiatric disorders.

While hundreds of circular RNAs (circRNAs) are abundant in mammalian brains, one big question has remained unanswered: What are they actually good for? In the...

Im Focus: RAVAN CubeSat measures Earth's outgoing energy

An experimental small satellite has successfully collected and delivered data on a key measurement for predicting changes in Earth's climate.

The Radiometer Assessment using Vertically Aligned Nanotubes (RAVAN) CubeSat was launched into low-Earth orbit on Nov. 11, 2016, in order to test new...

Im Focus: Scientists shine new light on the “other high temperature superconductor”

A study led by scientists of the Max Planck Institute for the Structure and Dynamics of Matter (MPSD) at the Center for Free-Electron Laser Science in Hamburg presents evidence of the coexistence of superconductivity and “charge-density-waves” in compounds of the poorly-studied family of bismuthates. This observation opens up new perspectives for a deeper understanding of the phenomenon of high-temperature superconductivity, a topic which is at the core of condensed matter research since more than 30 years. The paper by Nicoletti et al has been published in the PNAS.

Since the beginning of the 20th century, superconductivity had been observed in some metals at temperatures only a few degrees above the absolute zero (minus...

Im Focus: Scientists improve forecast of increasing hazard on Ecuadorian volcano

Researchers from the University of Miami (UM) Rosenstiel School of Marine and Atmospheric Science, the Italian Space Agency (ASI), and the Instituto Geofisico--Escuela Politecnica Nacional (IGEPN) of Ecuador, showed an increasing volcanic danger on Cotopaxi in Ecuador using a powerful technique known as Interferometric Synthetic Aperture Radar (InSAR).

The Andes region in which Cotopaxi volcano is located is known to contain some of the world's most serious volcanic hazard. A mid- to large-size eruption has...

All Focus news of the innovation-report >>>

Anzeige

Anzeige

Event News

Call for Papers – ICNFT 2018, 5th International Conference on New Forming Technology

16.08.2017 | Event News

Sustainability is the business model of tomorrow

04.08.2017 | Event News

Clash of Realities 2017: Registration now open. International Conference at TH Köln

26.07.2017 | Event News

 
Latest News

New thruster design increases efficiency for future spaceflight

16.08.2017 | Physics and Astronomy

Transporting spin: A graphene and boron nitride heterostructure creates large spin signals

16.08.2017 | Materials Sciences

A new method for the 3-D printing of living tissues

16.08.2017 | Interdisciplinary Research

VideoLinks
B2B-VideoLinks
More VideoLinks >>>