July 18, 2013

Mining for meaning: Getting computers to understand natural language texts

Programs that can understand language and can identify meaningful links between the various parts of a text is the focus of work being carried out in Saarbrücken by researchers like Ivan Titov. The computer scientist is currently developing a procedure that will enable computers to learn to identify semantically relevant relationships within texts. This research could mean that in future we will be able to ask our computer specific questions about the content of a text. The computer would then analyse the text and supply the user with the right answers.

Every student who was ever written a homework assignment or an academic essay is familiar with the problem: before you can start to write anything yourself, you usually have to battle through numerous texts and through pages and pages of references and academic literature. A computer program that can quickly process the text, provide a meaningful summary of its content or even answer questions about it would obviously be of great practical value in such a situation.

Ivan Titov and his team of researchers, splitting their time between Saarland University and the University of Amsterdam, are currently working on this problem. Titov is interested in how computers can learn to understand the meaning and the relationships between words in sentences and within texts. 'The model that we have developed simulates how humans create texts. In order to understand texts, we get our computers to work through this process but in the reverse direction: given the text the computer will uncover its meaning or even intent of the writer' explained Dr Titov. However, Titov and his group do not themselves stipulate a fully detailed model and the rules contained within it, instead, they use millions of sentences to generate both the model and the rules. The sentences that are analysed are drawn from large collections, such as Wikipedia. Analysis of this massive dataset requires a lot of computing power, with the specially developed algorithms running on around one hundred computers.

The idea is to develop software that enables computers to identify hidden, context-dependent relationships between words and clauses in texts, as the following example shows. Looking at the two sentences: 'John has just graduated from Saarland University. He is now working for Google.' it is clear even to a computer that John and Saarland University are linked by the relationship 'has graduated' and that John and Google are connected by the relationship "is working for". But the model developed by the Saarbrücken computer scientists can also recognize that John studied at Saarland University; very probably in the Department of Computer Science and Informatics. Once computers can understand these patterns in human language, the next step for the researchers is to apply the method to get machines to automatically produce meaningful summaries of short texts and to answer questions about the text content.

Beside Ivan Titov Hans Uszkoreit is awarded with a Google Focused Award worth US$ 220.000. Uszkoreit is professor of Computational Linguistics at Saarland University and Scientific Director at the German Research Center for Artificial Intelligence. He is interested in how computers can indentify linguistic relationships in large text collections.

Through its Focused Research Award program, the search engine provider Google supports research of major interest to the company itself and to the field of informatics. Prize winners receive free access to Google tools and technologies.

Provided by Saarland University

Citation: Mining for meaning: Getting computers to understand natural language texts (2013, July 18) retrieved 30 June 2024 from https://phys.org/news/2013-07-natural-language-texts.html

This document is subject to copyright. Apart from any fair dealing for the purpose of private study or research, no part may be reproduced without the written permission. The content is provided for information purposes only.

Explore further

New computer-based tool measures readability for different readers

0 shares

Feedback to editors

The Milky Way's eROSITA bubbles are large and distant

Jun 29, 2024

Saturday Citations: Armadillos are everywhere; Neanderthals still surprising anthropologists; kids are egalitarian

Jun 29, 2024

NASA astronauts will stay at the space station longer for more troubleshooting of Boeing capsule

Jun 29, 2024

The beginnings of fashion: Paleolithic eyed needles and the evolution of dress

Jun 28, 2024

Analysis of NASA InSight data suggests Mars hit by meteoroids more often than thought

Jun 28, 2024

New computational microscopy technique provides more direct route to crisp images

Jun 28, 2024

A harmless asteroid will whiz past Earth Saturday. Here's how to spot it

Jun 28, 2024

Tiny bright objects discovered at dawn of universe baffle scientists

Jun 28, 2024

New method for generating monochromatic light in storage rings

Jun 28, 2024

Soft, stretchy electrode simulates touch sensations using electrical signals

Jun 28, 2024

Load comments (0)

Mining for meaning: Getting computers to understand natural language texts

The Milky Way's eROSITA bubbles are large and distant

Saturday Citations: Armadillos are everywhere; Neanderthals still surprising anthropologists; kids are egalitarian

NASA astronauts will stay at the space station longer for more troubleshooting of Boeing capsule

The beginnings of fashion: Paleolithic eyed needles and the evolution of dress

Analysis of NASA InSight data suggests Mars hit by meteoroids more often than thought

New computational microscopy technique provides more direct route to crisp images

A harmless asteroid will whiz past Earth Saturday. Here's how to spot it

Tiny bright objects discovered at dawn of universe baffle scientists

New method for generating monochromatic light in storage rings

Soft, stretchy electrode simulates touch sensations using electrical signals

Relevant PhysicsForums posts

Newbie question about deep learning

Who can find the largest prime number with their own programmed code?

Math Major Trying to Learn CS

Parallelizing N-Queens

How to test locally hosted websites on mobile?

Question about learning programming

New computer-based tool measures readability for different readers

Our ambiguous world of words

Researchers using Kinect to allow deaf people to communicate via computer (w/ Video)

Researcher develops computational text analysis method made possible regardless of language or domain

Linguists, computer scientists use supercomputers to improve natural language processing

New study suggests Voynich text is not a hoax

Hyphens in paper titles harm citation counts and journal impact factors

A big step toward the practical application of 3-D holography with high-performance computers

Combining multiple CCTV images could help catch suspects

Applying deep learning to motion capture with DeepLabCut

Training artificial intelligence with artificial X-rays

New model for large-scale 3-D facial recognition

Medical Xpress

Tech Xplore

Science X

Mining for meaning: Getting computers to understand natural language texts

The Milky Way's eROSITA bubbles are large and distant

Saturday Citations: Armadillos are everywhere; Neanderthals still surprising anthropologists; kids are egalitarian

NASA astronauts will stay at the space station longer for more troubleshooting of Boeing capsule

The beginnings of fashion: Paleolithic eyed needles and the evolution of dress

Analysis of NASA InSight data suggests Mars hit by meteoroids more often than thought

New computational microscopy technique provides more direct route to crisp images

A harmless asteroid will whiz past Earth Saturday. Here's how to spot it

Tiny bright objects discovered at dawn of universe baffle scientists

New method for generating monochromatic light in storage rings

Soft, stretchy electrode simulates touch sensations using electrical signals

Relevant PhysicsForums posts

Related Stories

New computer-based tool measures readability for different readers

Our ambiguous world of words

Researchers using Kinect to allow deaf people to communicate via computer (w/ Video)

Researcher develops computational text analysis method made possible regardless of language or domain

Linguists, computer scientists use supercomputers to improve natural language processing

New study suggests Voynich text is not a hoax

Recommended for you

Hyphens in paper titles harm citation counts and journal impact factors

A big step toward the practical application of 3-D holography with high-performance computers

Combining multiple CCTV images could help catch suspects

Applying deep learning to motion capture with DeepLabCut

Training artificial intelligence with artificial X-rays

New model for large-scale 3-D facial recognition

Newsletter sign up

Donate and enjoy an ad-free experience