Researchers at the Fraunhofer Institute for Algorithms and Scientific Computing (SCAI) and at the Jülich Supercomputing Centre (JSC) of Forschungszentrum Jülich have used their substantial computing grid infrastructures for a new application in scientific computing: the large-scale annotation of biomedical and chemical texts and images in pharmaceutical patents. This will allow patent searches of an unparalleled power. Now, queries provide interesting insights into intersections between biology and chemistry, and the analysis of chemistry is truly multi-modal in the sense that text- and image-based information can be analyzed simultaneously.
More than 50,000 patents describing inventions in pharmaceutical chemistry have been processed on the large-scale computing grid infrastructures at SCAI and JSC. Automated "named entity recognition" services have identified and annotated:
o biological entities in text (e.g. protein names; gene names; gene polymorphisms; cell types),
o medical entities in text (e.g. disease names; pathology terms; risk factor terminology) as well as
o chemical information in text (e.g. drug names; expressions following the naming standards of the International Union of Pure and Applied Chemistry (IUPAC)) and
o images (e.g. chemical structure depictions).
The grid middleware UNICORE (Uniform Interface to Computing Resources) was used to manage the annotation services in the grid infrastructure, to control the streams of input and output data from the patents database to the annotation services, and to monitor the overall progress.
"This large-scale experiment opens new perspectives in scientific computing," says Prof. Dr. Martin Hofmann-Apitius, head of the Department of Bioinformatics at Fraunhofer SCAI. "This type of application goes way beyond the usual simulation applications that we are used to in the scientific computing community."
So far, text mining applications have only been run on bibliographic databases of life sciences and biomedical information such as MEDLINE. But the extension towards a multimodal analysis including annotation of text- and image-based information in full text documents on grid infrastructures has never been done before.
"We are pleased to see that our institute, which has a strong record in numerical simulation, has contributed to a new field of applications for supercomputers: what we call knowledge computing is likely to become a new discipline on its own," emphasizes Prof. Dr. Ulrich Trottenberg, Director of Fraunhofer SCAI.
"UNICORE made it possible to run this experiment at such a large scale in computing grid infrastructures at SCAI and JSC," says Dr. Achim Streit, head of Distributed Systems and Grid Computing at JSC. "The powerful workflow and data management capabilities of UNICORE allowed to annotate the patents in a seamless and automated way. A supercomputer connected by UNICORE to the infrastructure of the German Grid Initiative (D-Grid) was used to perform the knowledge extraction. This initial step of the experiment demonstrates what is possible today and shows the potential for more complex production runs in the future, using HPC systems connected in grid infrastructures".
"This is a very good example of how powerful supercomputers at JSC equipped with world-class grid technologies like UNICORE can generate synergies to enable new fields of research. I am proud that JSC is a member of the international UNICORE open source community and leads its development," explains Prof. Dr. Dr. Thomas Lippert, Director of JSC.
The team at SCAI, led by Dr. Marc Zimmermann for the image analysis annotators and by Dr. Juliane Fluck and Dr. Christoph Friedrich for the text analytics part, is currently working on the in-depth analysis of the meta-information generated in the course of this large-scale in silico-experiment. Their colleague on the side of JSC in Jülich, Mathilde Romberg, is happy that after weeks of intensive work the first "production runs" have been completed. However, the teams on both sides know that there are another 1.5 million patents waiting for them.Contact:
Michael Krapp | Fraunhofer Gesellschaft
Powerful IT security for the car of the future – research alliance develops new approaches
25.05.2018 | Universität Ulm
Supercomputing the emergence of material behavior
18.05.2018 | University of Texas at Austin, Texas Advanced Computing Center
The more electronics steer, accelerate and brake cars, the more important it is to protect them against cyber-attacks. That is why 15 partners from industry and academia will work together over the next three years on new approaches to IT security in self-driving cars. The joint project goes by the name Security For Connected, Autonomous Cars (SecForCARs) and has funding of €7.2 million from the German Federal Ministry of Education and Research. Infineon is leading the project.
Vehicles already offer diverse communication interfaces and more and more automated functions, such as distance and lane-keeping assist systems. At the same...
A research team led by physicists at the Technical University of Munich (TUM) has developed molecular nanoswitches that can be toggled between two structurally different states using an applied voltage. They can serve as the basis for a pioneering class of devices that could replace silicon-based components with organic molecules.
The development of new electronic technologies drives the incessant reduction of functional component sizes. In the context of an international collaborative...
At the LASYS 2018, from June 5th to 7th, the Laser Zentrum Hannover e.V. (LZH) will be showcasing processes for the laser material processing of tomorrow in hall 4 at stand 4E75. With blown bomb shells the LZH will present first results of a research project on civil security.
At this year's LASYS, the LZH will exhibit light-based processes such as cutting, welding, ablation and structuring as well as additive manufacturing for...
There are videos on the internet that can make one marvel at technology. For example, a smartphone is casually bent around the arm or a thin-film display is rolled in all directions and with almost every diameter. From the user's point of view, this looks fantastic. From a professional point of view, however, the question arises: Is that already possible?
At Display Week 2018, scientists from the Fraunhofer Institute for Applied Polymer Research IAP will be demonstrating today’s technological possibilities and...
So-called quantum many-body scars allow quantum systems to stay out of equilibrium much longer, explaining experiment | Study published in Nature Physics
Recently, researchers from Harvard and MIT succeeded in trapping a record 53 atoms and individually controlling their quantum state, realizing what is called a...
25.05.2018 | Event News
02.05.2018 | Event News
13.04.2018 | Event News
25.05.2018 | Event News
25.05.2018 | Machine Engineering
25.05.2018 | Life Sciences