August 3, 2007

New program color-codes text in Wikipedia entries to indicate trustworthiness

The online reference site Wikipedia enjoys immense popularity despite nagging doubts about the reliability of entries written by its all-volunteer team. A new program developed at the University of California, Santa Cruz, aims to help with the problem by color-coding an entry's individual phrases based on contributors' past performance.

The program analyzes Wikipedia's entire editing history--nearly two million pages and some 40 million edits for the English-language site alone--to estimate the trustworthiness of each page. It then shades the text in deepening hues of orange to signal dubious content. A 1,000-page demonstration version is already available on a web page operated by the program's creator, Luca de Alfaro, associate professor of computer engineering at UCSC.

Other sites already employ user ratings as a measure of reliability, but they typically depend on users' feedback about each other. This method makes the ratings vulnerable to grudges and subjectivity. The new program takes a radically different approach, using the longevity of the content itself to learn what information is useful and which contributors are the most reliable.

"The idea is very simple," de Alfaro said. "If your contribution lasts, you gain reputation. If your contribution is reverted [to the previous version], your reputation falls." De Alfaro will speak about his new program this Saturday, August 4, at the Wikimania conference in Taipei, Taiwan.

The program works from a user's history of edits to calculate his or her reputation score. The trustworthiness of newly inserted text is computed as a function of the reputation of its author. As subsequent contributors vet the text, their own reputations contribute to the text's trustworthiness score. So an entry created by an unknown author can quickly gain (or lose) trust after a few known users have reviewed the pages.

A benefit of calculating author reputation in this way is that de Alfaro can test how well his reliability scores work. He does so by comparing users' reliability scores with how long their subsequent edits last on the site. So far, the program flags as suspect more than 80 percent of edits that turn out to be poor. It's not overly accusatory, either: 60 to 70 percent of the edits it flags do end up being quickly corrected by the Wikipedia community.

The exhaustive analysis of Wikipedia's seven-year edit history takes de Alfaro's desktop PC about a week to complete. At present he is working from copies of the site that Wikipedia periodically distributes. Once the initial backlog of edits is calculated, however, de Alfaro said that updating reliability scores in real time should be fairly simple.

While the program prominently displays text trustworthiness, de Alfaro favors keeping hidden the reputation ratings of individual users. Displaying reputations could lead to competitiveness that would detract from Wikipedia's collaborative culture, he said, and could demoralize knowledgeable contributors whose scores remain low simply because they post infrequently and on few topics.

"We didn't want to modify the experience of a user going in to Wikipedia," de Alfaro said. "It is very relaxing right now and we didn't want to modify what has worked so well and is so welcoming to the new user."

Source: UC Santa Cruz

Citation: New program color-codes text in Wikipedia entries to indicate trustworthiness (2007, August 3) retrieved 16 July 2024 from https://phys.org/news/2007-08-color-codes-text-wikipedia-entries-trustworthiness.html

This document is subject to copyright. Apart from any fair dealing for the purpose of private study or research, no part may be reproduced without the written permission. The content is provided for information purposes only.

Explore further

New drone imagery reveals 97% of coral dead at a Lizard Island reef after last summer's mass bleaching

0 shares

Feedback to editors

Researchers achieve unprecedented nanostructuring inside silicon

44 minutes ago

The current international poverty line is a 'misleading shortcut method,' say experts

44 minutes ago

Animal researchers develop digital dog and cat skull database

44 minutes ago

World's rarest whale may have washed up on New Zealand beach, possibly shedding clues on species

1 hour ago

Silicon photonics light the way toward large-scale applications in quantum information

13 hours ago

Earth system scientists discover missing piece in climate models

13 hours ago

Research team uses satellite data and machine learning to predict typhoon intensity

14 hours ago

Researchers directly simulate the fusion of oxygen and carbon nuclei

14 hours ago

New tool can predict bitterness in foods without prior knowledge of their chemical structures

14 hours ago

Nano-confinement may be key to improving hydrogen production

14 hours ago

Load comments (0)

New program color-codes text in Wikipedia entries to indicate trustworthiness

Researchers achieve unprecedented nanostructuring inside silicon

The current international poverty line is a 'misleading shortcut method,' say experts

Animal researchers develop digital dog and cat skull database

World's rarest whale may have washed up on New Zealand beach, possibly shedding clues on species

Silicon photonics light the way toward large-scale applications in quantum information

Earth system scientists discover missing piece in climate models

Research team uses satellite data and machine learning to predict typhoon intensity

Researchers directly simulate the fusion of oxygen and carbon nuclei

New tool can predict bitterness in foods without prior knowledge of their chemical structures

Nano-confinement may be key to improving hydrogen production

Relevant PhysicsForums posts

Help solving a geometrical matching issue with Graph Neural Networks

5 GHz PC WiFi connection Cybersecurity question

Help with some optimization code for Block Matrices

Is an API Always Necessary for Server-Client Communication?

I did this POST message configuration damage to my wifi internet, help

Number of Multiplications in the FFT Algorithm

New drone imagery reveals 97% of coral dead at a Lizard Island reef after last summer's mass bleaching

The more medals Canadian athletes win, the fewer Canadians participate in organized sport

Industrial fleets operating in the Indian Ocean turn off monitoring systems, fail reporting obligations

New calculation approach allows more accurate predictions of how atoms ionize when impacted by high-energy electrons

From takeoff to flight, the wiring of a fly's nervous system is mapped

Wolves reintroduced to Isle Royale temporarily affect other carnivores, humans have influence as well

Hyphens in paper titles harm citation counts and journal impact factors

A big step toward the practical application of 3-D holography with high-performance computers

Combining multiple CCTV images could help catch suspects

Applying deep learning to motion capture with DeepLabCut

Training artificial intelligence with artificial X-rays

New model for large-scale 3-D facial recognition

Medical Xpress

Tech Xplore

Science X

New program color-codes text in Wikipedia entries to indicate trustworthiness

Researchers achieve unprecedented nanostructuring inside silicon

The current international poverty line is a 'misleading shortcut method,' say experts

Animal researchers develop digital dog and cat skull database

World's rarest whale may have washed up on New Zealand beach, possibly shedding clues on species

Silicon photonics light the way toward large-scale applications in quantum information

Earth system scientists discover missing piece in climate models

Research team uses satellite data and machine learning to predict typhoon intensity

Researchers directly simulate the fusion of oxygen and carbon nuclei

New tool can predict bitterness in foods without prior knowledge of their chemical structures

Nano-confinement may be key to improving hydrogen production

Relevant PhysicsForums posts

Related Stories

New drone imagery reveals 97% of coral dead at a Lizard Island reef after last summer's mass bleaching

The more medals Canadian athletes win, the fewer Canadians participate in organized sport

Industrial fleets operating in the Indian Ocean turn off monitoring systems, fail reporting obligations

New calculation approach allows more accurate predictions of how atoms ionize when impacted by high-energy electrons

From takeoff to flight, the wiring of a fly's nervous system is mapped

Wolves reintroduced to Isle Royale temporarily affect other carnivores, humans have influence as well

Recommended for you

Hyphens in paper titles harm citation counts and journal impact factors

A big step toward the practical application of 3-D holography with high-performance computers

Combining multiple CCTV images could help catch suspects

Applying deep learning to motion capture with DeepLabCut

Training artificial intelligence with artificial X-rays

New model for large-scale 3-D facial recognition

Newsletter sign up

Donate and enjoy an ad-free experience