The newest AI computing tool: people
June 28, 2007A USC Information Sciences Institute researcher is among a growing group of computer scientists learning to solve difficult IT problems of information classification, reliability and meaning by datamining public websites like Digg, del.icio.us and Flickr.
That tool, according to ISI computer scientist Kristina Lerman, is people, human intelligence at work on the social web, the network of blogs, bookmark, photo and video- sharing sites, and other meeting places now involving hundreds of thousands of individuals daily, recording observations and sharing opinions and information.
Lerman shared her recent work with others in the burgeoning new field of social information processing a special AAAI-sponsored symposium on the subject March 26-28 at Stanford.
She says that extracting 'metadata' about transactions -- who is talking to whom, who is listening, how conclusions are reached, and how they spread -- can help researchers answer currently refractory problems about documents: their accuracy and quality, their categorization, the relation of their embedded terminology.
One benefit, according to Lerman, who in addition to her ISI appointment, is a research assistant professor at the Viterbi School of Engineering Department of Computer Science at University of Southern California, is automatic determination of the semantics of content from one kind of metadata: tags.
Tags play a crucial role in a longrunning project called the Semantic Web.
For about a decade, she notes, researchers sought a way to organize data so that someone searching for a specific kind of "check" wouldn't have to weed out unwanted references to chess, symbols, verification procedures, financial documents, political science theories and many more.
Tagging seeks to eliminate ambiguities by affixing 'tags,' computer labels peeling apart the multiple meanings of ordinary language into discreet indicators of meaning, guiding computer searches.
But with natural language being as complex as it is, making sense of tags is not easy. Attempts to manually attack the vocabulary and build in the intricate interconnections that signal different word meanings have proved frustrating.
Lerman hopes she's onto another way. Hundreds of thousands of users are now online, chattering away on all kinds of topics. This volume of directed discourse provides a new way to extracting meaning from tags —statistical models.
The process has been called "folksonomy," a collectively constructed informal classification system. Unlike the traditional approach to the Semantic Web, in which a few knowledge professionals try to agree on a formal classification system which will then be used to annotate data, folksonomy emerges from collective tagging activities of many individuals.
New social websites aimed at sharing information such as del.icio.us and Flickr organically grow ways for site members to access each others holdings. Typically, the members themselves spontaneously create a tagging system, encouraged by the site architecture.
The tags emerging from such systems, Lerman and collaborators have found, can be turned to broader purposes.
One of Lerman's initial tagging investigations used the photo-sharing site Flickr, analyzing results returned by a request for images of 'beetles,' including some pictures of insects, some pictures of Volkswagens, and a few other entries.
By extracting the tags that Flickr users had described the images with, and applyng a mathematical technique called the "Expectation-maximization (EM) algorithm," Lerman found it possible to quite accurately separate pictures of insects from pictures of cars returned by the “beetle” search.
Lerman has gone beyond tagging to using metadata to acquire more and more accurate information about the content of documents in social networking situations.
A Lerman paper now in pre-publication on "Social Information Processing in Social News Aggregation" notes: The rise of the social media sites, such as blogs, wikis, Digg and Flickr among others, underscores the transformation of the Web to a participatory medium in which users are collaboratively creating, evaluating and distributing information.
The innovations introduced by social media have lead to a new paradigm for interacting with information, what we call 'social information processing'.
In the paper, Lerman argues that "by tracking stories over time, that social networks play an important role in document recommendation." In addition to providing a platform for document recommendation, social Web enables researchers to study collective user behavior quantitatively.
In the same paper, Lerman also presented a mathematical model of how collaborative rating and promotion of stories emerges from the independent decisions made by many users. She found good agreement between predictions of the model and user data gathered from Digg.
In another paper, examining de.licio.us, Lerman and collaborators "describe a probabalistic model of the user annotation process," and then used the model "to automatically find resources relevant to a particular information domain ... with promising results."
Source: University of Southern California
-
Engineers build first sub-10-nm carbon nanotube transistor
Feb 01, 2012 |
4.9 / 5 (30) |
30
-
Something old, something new: Evolution and the structural divergence of duplicate genes
Jan 31, 2012 |
4.6 / 5 (7) |
1
-
The hidden nanoworld of ice crystals: Revealing the dynamic behavior of quasi-liquid layers
Jan 30, 2012 |
5 / 5 (3) |
1
-
Stock market network reveals investor clustering
Jan 27, 2012 |
3.9 / 5 (23) |
8
-
Of microchemistry and molecules: Electronic microfluidic device synthesizes biocompatible probes
Jan 26, 2012 |
5 / 5 (1) |
0
-
Synergistic relations between computer science and technology.
Feb 06, 2012
-
how do iphone gloves work?
Feb 05, 2012
-
iPhone battery over time
Jan 30, 2012
-
Best alternate Tablet to an iPad for writing math or physics equations?
Jan 26, 2012
-
Sending SMS to a website
Jan 20, 2012
-
Need help with my technical fest!
Jan 19, 2012
- More from Physics Forums - Computing & Technology
More news stories
Netflix light on flicks as viewers soak up TV shows
Like most fresh faces that arrive in Hollywood, Netflix wanted to be a movie star. But now it's learning what many in Tinseltown have known for decades: Movies are sexy, but the real money is in television.
36 minutes ago |
not rated yet |
1
Sony's Hirai refuses to abandon dire TV business
Struggling Japanese entertainment giant Sony will not abandon its cash-bleeding television business, its incoming CEO says, but he acknowledges tough decisions lie ahead including over redundancies.
1 hour ago |
not rated yet |
0
New error-correcting codes guarantee the fastest possible rate of data transmission
Error-correcting codes are one of the triumphs of the digital age. Theyre a way of encoding information so that it can be transmitted across a communication channel such as an optical fiber o ...
Technology / Computer Sciences
3 hours ago |
5 / 5 (3) |
2
|
Small modular reactor design could be a 'SUPERSTAR'
(PhysOrg.com) -- Though most of today's nuclear reactors are cooled by water, we've long known that there are alternatives; in fact, the world's first nuclear-powered electricity in 1951 came from a reactor ...
Technology / Energy & Green Tech
3 hours ago |
5 / 5 (5) |
9
|
Advanced power-grid model finds low-cost, low-carbon future in West
(PhysOrg.com) -- The least expensive way for the Western U.S. to reduce greenhouse gas emissions enough to help prevent the worst consequences of global warming is to replace coal with renewable and other ...
Technology / Energy & Green Tech
3 hours ago |
5 / 5 (1) |
3
|
Curry spice component may help slow prostate tumor growth
Curcumin, an active component of the Indian curry spice turmeric, may help slow down tumor growth in castration-resistant prostate cancer patients on androgen deprivation therapy (ADT), a study from researchers ...
Antidepressants and pregnancy: Women must consider the impact of drugs on baby, and of depression on baby, themselves
Upon learning they are pregnant, most women dutifully nix the alcohol, sushi and caffeine. But what about antidepressants?
To avoid early labor and delivery, weight and diet changes not the answer
One of the strongest known risk factors for spontaneous or unexpected preterm birth any birth that occurs before the 37th week of pregnancy, most often without a known cause is already having had one. For women ...
Arthritic knees, but not hips, have robust repair response
Researchers at Duke University Medical Center used new tools they developed to analyze knees and hips and discovered that osteoarthritic knee joints are in a constant state of repair, while hip joints are not.
The power of estrogen -- male snakes attract other males
A new study has shown that boosting the estrogen levels of male garter snakes causes them to secrete the same pheromones that females use to attract suitors, and turned the males into just about the sexiest ...
Fool's gold may prove an unlikely alternative to overexploited catalytic materials
Catalytic materials, which lower the energy barriers for chemical reactions, are used in everything from the commercial production of chemicals to catalytic converters in car engines. However, with current catalytic materials ...