The newest AI computing tool: people

June 28, 2007

A USC Information Sciences Institute researcher is among a growing group of computer scientists learning to solve difficult IT problems of information classification, reliability and meaning by datamining public websites like Digg, del.icio.us and Flickr.

That tool, according to ISI computer scientist Kristina Lerman, is people, human intelligence at work on the social web, the network of blogs, bookmark, photo and video- sharing sites, and other meeting places now involving hundreds of thousands of individuals daily, recording observations and sharing opinions and information.

Lerman shared her recent work with others in the burgeoning new field of social information processing a special AAAI-sponsored symposium on the subject March 26-28 at Stanford.

She says that extracting 'metadata' about transactions -- who is talking to whom, who is listening, how conclusions are reached, and how they spread -- can help researchers answer currently refractory problems about documents: their accuracy and quality, their categorization, the relation of their embedded terminology.

One benefit, according to Lerman, who in addition to her ISI appointment, is a research assistant professor at the Viterbi School of Engineering Department of Computer Science at University of Southern California, is automatic determination of the semantics of content from one kind of metadata: tags.

Tags play a crucial role in a longrunning project called the Semantic Web.

For about a decade, she notes, researchers sought a way to organize data so that someone searching for a specific kind of "check" wouldn't have to weed out unwanted references to chess, symbols, verification procedures, financial documents, political science theories and many more.

Tagging seeks to eliminate ambiguities by affixing 'tags,' computer labels peeling apart the multiple meanings of ordinary language into discreet indicators of meaning, guiding computer searches.

But with natural language being as complex as it is, making sense of tags is not easy. Attempts to manually attack the vocabulary and build in the intricate interconnections that signal different word meanings have proved frustrating.

Lerman hopes she's onto another way. Hundreds of thousands of users are now online, chattering away on all kinds of topics. This volume of directed discourse provides a new way to extracting meaning from tags —statistical models.

The process has been called "folksonomy," a collectively constructed informal classification system. Unlike the traditional approach to the Semantic Web, in which a few knowledge professionals try to agree on a formal classification system which will then be used to annotate data, folksonomy emerges from collective tagging activities of many individuals.

New social websites aimed at sharing information such as del.icio.us and Flickr organically grow ways for site members to access each others holdings. Typically, the members themselves spontaneously create a tagging system, encouraged by the site architecture.

The tags emerging from such systems, Lerman and collaborators have found, can be turned to broader purposes.

One of Lerman's initial tagging investigations used the photo-sharing site Flickr, analyzing results returned by a request for images of 'beetles,' including some pictures of insects, some pictures of Volkswagens, and a few other entries.

By extracting the tags that Flickr users had described the images with, and applyng a mathematical technique called the "Expectation-maximization (EM) algorithm," Lerman found it possible to quite accurately separate pictures of insects from pictures of cars returned by the “beetle” search.

Lerman has gone beyond tagging to using metadata to acquire more and more accurate information about the content of documents in social networking situations.

A Lerman paper now in pre-publication on "Social Information Processing in Social News Aggregation" notes: The rise of the social media sites, such as blogs, wikis, Digg and Flickr among others, underscores the transformation of the Web to a participatory medium in which users are collaboratively creating, evaluating and distributing information.

The innovations introduced by social media have lead to a new paradigm for interacting with information, what we call 'social information processing'.

In the paper, Lerman argues that "by tracking stories over time, that social networks play an important role in document recommendation." In addition to providing a platform for document recommendation, social Web enables researchers to study collective user behavior quantitatively.

In the same paper, Lerman also presented a mathematical model of how collaborative rating and promotion of stories emerges from the independent decisions made by many users. She found good agreement between predictions of the model and user data gathered from Digg.

In another paper, examining de.licio.us, Lerman and collaborators "describe a probabalistic model of the user annotation process," and then used the model "to automatically find resources relevant to a particular information domain ... with promising results."

Source: University of Southern California


print this article email this article download pdf blog this article bookmark this article     Stumble it Digg this share on Facebook retweet share on Reddit add to delicious    
Rate this story - 3.4 /5 (16 votes)


June 28, 2007 all stories

Comments: 0

3.4 /5 (16 votes)
  • Stumble this up

  • Digg this

  • share this

  • hide
  • Related Stories

  • eStadium application brings multimedia sports features to smartphones
    created Nov 06, 2009 | popularity not rated yet | comments 0
  • Psychiatric impact of torture could be amplified by head injury
    created Nov 06, 2009 | popularity not rated yet | comments 0
  • The politics of climate fixes
    created Nov 06, 2009 | popularity not rated yet | comments 0
  • Perceived parent-pressure causes excessive antibiotic prescription
    created Nov 06, 2009 | popularity not rated yet | comments 0
  • Cultural Beliefs About Pesticides Put Mexican Farmworkers at Risk
    created Nov 05, 2009 | popularity not rated yet | comments 0



  • hide
  • Relevant PhysicsForums posts

  • Read multiple binary files to ascii
    created 21 hours ago
  • Engineering Translation software
    created Nov 06, 2009
  • Changing the language options on your phone.
    created Nov 03, 2009
  • HP strange RPN operation???
    created Nov 02, 2009
  • Computational physics problems that involve nontrivial CS concepts?
    created Nov 01, 2009
  • Databases in physics
    created Oct 31, 2009
  • More from Physics Forums - Computing & Technology

Other News

Microsoft websites were the most visited in September

Microsoft websites top spots in September: comScore

Technology / Internet

created 17 hours ago | popularity 2.3 / 5 (3) | comments 0

Industry tracker comScore on Friday released a study showing that Internet users in September spent more time at Microsoft websites that at any other online properties.


Brazil blackouts result of cyber hacking: report

Technology / Internet

created 17 hours ago | popularity 3 / 5 (3) | comments 0

Massive power outages in Brazil in 2005 and 2007 that impacted millions were caused by cyber hackers attacking control systems, the US television network CBS said Sunday.


airpod

Car That Runs on Compressed Air Questioned by Critics (w/ Video)

Technology / Energy

created Nov 03, 2009 | popularity 3.8 / 5 (18) | comments 26

(PhysOrg.com) -- As electric cars begin breaking into the short-distance vehicle market, one French company thinks that it has an alternative to the electric vehicle: a car that runs on compressed air. Motor ...


The Beatles perform in 1964 at the Olympia in Paris

Bluebeat to battle EMI over Beatles songs

Technology / Internet

created 16 hours ago | popularity 4 / 5 (1) | comments 0

US online music service Bluebeat said it plans to fight British recording label EMI over rights to stream and sell versions of Beatles songs.


Sahara

Will Europe Be Powered by the Sahara

Technology / Energy

created Nov 04, 2009 | popularity 4.1 / 5 (19) | comments 24

(PhysOrg.com) -- Europe has long been interested in developing alternative energy sources. And, one of the more interesting places that some Europeans are looking for solar power is the Sahara. With the vast ...